Skip to main content
Back to News Hub
šŸ¤–OpenAI
October 29, 2025
Funding & Investment

gpt-oss-safeguard technical report

Overview

gpt-oss-safeguard-120b and gpt-oss-safeguard-20b are two open-weight reasoning models post-trained from the gpt-oss models and trained to reason from a provided policy in order to label content under that policy. In this report, we describe gpt-oss-safeguard's capabilities and provide our baseline safety evaluations on the gpt-oss-safeguard models, using the underlying gpt-oss models as a baseline. For more information about the development and architecture of the underlying gpt-oss models, see the original gpt-oss model model card⁠.

Read the full story at OpenAI

This publisher only syndicates a short excerpt by RSS. The full article, with all the detail, quotes, and context, lives on their site.

Open original article

Continue Learning

Comments

Comments appear only after moderation. Your email identifies your submission to the moderator and is never displayed here.

No approved comments yet.

Originally published by OpenAI
Read the original