This repository contains OpenAI's 20B parameter safeguard model designed to evaluate and classify text for safety, toxicity, and policy violations. It serves as a guardrail for LLM applications. The model weights are openly available on Hugging Face for integration into developer pipelines.
Gpt Oss Safeguard 20b sits in PulseGate's Foundation models & chat category. It focuses on detecting and filtering unsafe or harmful content generated by large language models. It is built as an open-source project for AI application developers. Gpt Oss Safeguard 20b is open source under the Apache-2.0 license. Gpt Oss Safeguard 20b is available on the web, the command line, and API.
OpenAI builds and maintains Gpt Oss Safeguard 20b, and the product first shipped in 2025. Development happens publicly on GitHub with 4.5k stars. Key capabilities include Content Moderation, Safety Classification, and LLM Guardrails.
Latest indexed changes and source events
openai/gpt-oss-safeguard-20b verified by the PulseGate indexer
Other apps tracked under the same category.