Skip to main content
AI Socratic
gpt-oss-safeguard-20b

OpenAI released gpt-oss-safeguard-20b on September 18, 2025. The registry currently records it as open weights.

Model publication

Open weights
openai/gpt-oss-safeguard-20b Hugging Face model card
openai/gpt-oss-safeguard-20b · model preview · Hugging Face model page

gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b. This open-weight, 21B-parameter Mixture-of-Experts (MoE) model offers lower latency for safety tasks like content classification, LLM filtering, and trust & safety labeling. Learn more about this model in OpenAI's gpt-oss-safeguard [user guide](https://cookbook.openai.com/articles/gpt-oss-safeguard-guide).

Licence
apache-2.0
Architecture
mixture_of_experts
Parameters
21.5B
Context
131.1K tokens

Benchmarks

No benchmark observation recorded for this model.

About the Authors

A

AI Socratic