AI Weekly Malaysia

Back to items Summaries

Learning from human preferences

ID
976
Status
new
Published
13 Jun 2017, 3:00 PM
Fetched
27 Jun 2026, 7:47 PM
Provider
OpenAI News
Category
ai-labs
Original URL
https://openai.com/index/learning-from-human-preferences
Source URL
https://openai.com/news/rss.xml

Excerpt

One step towards building safe AI systems is to remove the need for humans to write goal functions, since using a simple proxy for a complex goal, or getting the complex goal a bit wrong, can lead to undesirable and even dangerous behavior. In collaboration with DeepMind’s safety team, we’ve developed an algorithm which can infer what humans want by being told which of two proposed behaviors is better.

Summary

No summary yet. It will appear after the daemon summarizes this item.

Top