AI Weekly Malaysia

Back to items Summaries

Estimating worst case frontier risks of open weight LLMs

ID
458
Status
new
Published
05 Aug 2025, 8:00 AM
Fetched
27 Jun 2026, 7:47 PM
Provider
OpenAI News
Category
ai-labs
Original URL
https://openai.com/index/estimating-worst-case-frontier-risks-of-open-weight-llms
Source URL
https://openai.com/news/rss.xml

Excerpt

In this paper, we study the worst-case frontier risks of releasing gpt-oss. We introduce malicious fine-tuning (MFT), where we attempt to elicit maximum capabilities by fine-tuning gpt-oss to be as capable as possible in two domains: biology and cybersecurity.

Summary

No summary yet. It will appear after the daemon summarizes this item.

Top