Home / AI Tech / Article
AI Tech

Estimating worst case frontier risks of open weight LLMs

OP

OpenAI Blog

August 04, 2025 at 08:00 PM

📌 In this paper, we study the worst-case frontier risks of releasing gpt-oss. We introduce malicious fine-tuning (MFT), where we attempt to elicit maximum capabilities by fine-tuning gpt-oss to be as capable as possible in two domains: biology and cybersecurity.

In this paper, we study the worst-case frontier risks of releasing gpt-oss. We introduce malicious fine-tuning (MFT), where we attempt to elicit maximum capabilities by fine-tuning gpt-oss to be as capable as possible in two domains: biology and cybersecurity.

Read the full article at

OpenAI Blog

Visit Source ↗