TLDRocket
Sign in

Open-Weight Models

32 summarised stories about Open-Weight Models, each linking back to the original source. Browse all topics →

Wednesday, 29 October 2025

Introducing gpt-oss-safeguard

OpenAI Blog 8 months ago 2 sources

OpenAI released gpt-oss-safeguard, an open-weight model designed to classify content safety issues and let developers implement custom safety policies. The model is available as open weights, enabling developers to run it locally and adapt it to their specific safety requirements. Developers can now build customized content moderation systems without relying on closed APIs or OpenAI's default policies.

gpt-oss-safeguard technical report

OpenAI Blog 8 months ago 2 sources

Two open-weight reasoning models, gpt-oss-safeguard-120b and gpt-oss-safeguard-20b, were developed by post-training gpt-oss models to apply provided safety policies for content labeling. The models come in 120 billion and 20 billion parameter sizes. The technical report establishes baseline safety evaluations comparing the new safeguard models against their underlying gpt-oss predecessors.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.