Difference between revisions of "Trust Region Policy Optimization (TRPO)"

From
Jump to: navigation, search
m
Line 9: Line 9:
  
 
* [[Deep Reinforcement Learning (DRL)]]
 
* [[Deep Reinforcement Learning (DRL)]]
 +
* [[Policy]]
 +
* [[Assistants]] ... [[Hybrid Assistants]]  ... [[Agents]]  ... [[Negotiation]] ... [[LangChain]]
 +
* [[Generative AI]]  ... [[OpenAI]]'s [[ChatGPT]] ... [[Perplexity]]  ... [[Microsoft]]'s [[BingAI]] ... [[You]] ...[[Google]]'s [[Bard]] ... [[Baidu]]'s [[Ernie]]
  
 
<youtube>xvRrgxcpaHY</youtube>
 
<youtube>xvRrgxcpaHY</youtube>

Revision as of 10:51, 26 March 2023