Skip to content
Back to the index

Wildguard

huggingface.coInfrastructure🇺🇸

PulseGate's liveness check found it on 14 Sep 2026; it is registered on Hugging Face and has been in the index since 19 Jul 2026. How this is checked

WildGuard is an open-source model developed by AllenAI for classifying harmful content, detecting jailbreak attempts, and performing safety evaluations on text. It is designed to help developers build safer AI applications by identifying problematic inputs before they reach production LLMs. The model is available on Hugging Face and can be used with standard transformer libraries.

Inferred · not functionally tested

Open SourceWebCLIAPI
Wildguard preview
Visit huggingface.co

Overview

3 features

Purpose: Identifying and filtering unsafe, harmful, or adversarial prompts in LLM applications.

Inferred · not functionally tested

Audience: developers

Inferred · not functionally tested

Functions: data_extraction

Inferred · not functionally tested

Interfaces: API: indicated (inferred, not tested) · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: unknown

Recorded constraints: pricing: open_source · license: Open Source · platforms: CLI, WEB · deployment: browser, cli, api_only

Constraint provenance is unknown; confirm requirements with the publisher.

Record sources: huggingface.co. These links do not verify the individual claims.

In the AI security & guardrails space, Wildguard takes a focused approach. Inferred · not functionally tested: It focuses on identifying and filtering unsafe, harmful, or adversarial prompts in LLM applications. Inferred · not functionally tested: It is built as an open-source project for developers. Basis unknown · not verified: Wildguard is open source under the Open Source license. Basis unknown · not verified: It ships for the web, the command line, and API.

It is developed by AllenAI (United States), and it first shipped in 2024. Inferred · not functionally tested: Among its 3 catalogued features are content moderation, jailbreak detection, and safety classification.

Summary written by a language model from the project’s public pages.

Tasks: Inferred · not functionally tested

  • Content moderation
  • Jailbreak detection
  • Safety classification

Topics: Inferred · not functionally tested

Tags
safety-modeljailbreak-detectioncontent-moderationharm-classifier
AI capabilities
Text
Inference: LocalWeights: Open

JSON profile · Text profile · Access guide

Built with & integrations

Hosting
AWS
AI providers
local_oss
Runs on
BrowserCLIAPI-only
Detected from
AWS
x-amz-cf-id header · x-amz-cf-pop header · via header
local_oss
bvllm in the HTML · text-generation- in the HTML

Trust & compliance

License
Open Source
Public signals
HTTPSPrivacy PolicyTerms of ServiceOpen SourceFree tier

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed19 Jul · 20:19 UTC
    allenai/wildguard seen via Hugging Face Enumerator
    Source: Hugging Face Enumerator · Open

Frequently asked questions about Wildguard

What does Wildguard do?
Inferred · not functionally tested: Wildguard focuses on identifying and filtering unsafe, harmful, or adversarial prompts in LLM applications. It is catalogued under AI security & guardrails on PulseGate.
Who is Wildguard for?
Inferred · not functionally tested: Wildguard is an open-source project built for developers.
Does Wildguard have a free plan?
Basis unknown · not verified: Yes — Wildguard is open source under the Open Source license and free to use.
What platforms does Wildguard run on?
Basis unknown · not verified: Wildguard runs on the web, the command line, and API.
Is Wildguard still active?
PulseGate's liveness check found it on 14 Sep 2026.
Who develops Wildguard?
Wildguard is developed by AllenAI, based in the United States.
When did Wildguard launch?
Wildguard first shipped in 2024.
Is Wildguard open source?
Basis unknown · not verified: Yes — Wildguard is open source under the Open Source license.

Similar projects

Closest matches by what these projects do