AI Weekly Malaysia

Back to items Summaries

Evolved Policy Gradients

ID
934
Status
new
Published
18 Apr 2018, 3:00 PM
Fetched
27 Jun 2026, 7:47 PM
Provider
OpenAI News
Category
ai-labs
Original URL
https://openai.com/index/evolved-policy-gradients
Source URL
https://openai.com/news/rss.xml

Excerpt

We’re releasing an experimental metalearning approach called Evolved Policy Gradients, a method that evolves the loss function of learning agents, which can enable fast training on novel tasks. Agents trained with EPG can succeed at basic tasks at test time that were outside their training regime, like learning to navigate to an object on a different side of the room from where it was placed during training.

Summary

No summary yet. It will appear after the daemon summarizes this item.

Top