Llama Guard 4 12B

Safety-focused model for content moderation.

llama-guard-4-12b
STABLEModel DeactivatedGet StartedView uptime
131,072 context
Released April 30, 2025
Starting at $0.20/M input tokens
Starting at $0.20/M output tokens
Streaming
No ratings yetSign in to rate

Select Provider

All Providers for Llama Guard 4 12B

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Groq
Context: 131.1k
Deactivated since Mar 29, 2026
Input
$0.2
/M tokens
Cache Read
/M tokens
Output
$0.2
/M tokens
Get Started

Frequently asked questions

What is Llama Guard 4 12B?

Safety-focused model for content moderation. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Llama Guard 4 12B cost?

Pricing for Llama Guard 4 12B on LLM Gateway starts at $0.20 per million input tokens and $0.20 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Llama Guard 4 12B?

Llama Guard 4 12B supports a context window of up to 131,072 tokens on its largest provider deployment.

Which providers serve Llama Guard 4 12B?

Llama Guard 4 12B is served by Groq through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Llama Guard 4 12B support tool calling and structured outputs?

No. Llama Guard 4 12B does not currently support tool calling or structured JSON outputs through LLM Gateway.

When was Llama Guard 4 12B released?

Llama Guard 4 12B was released on April 30, 2025.