Llama Guard 4 12B

Safety-focused model for content moderation.

llama-guard-4-12b
STABLEGet StartedView uptime
163,840 context
Released April 30, 2025
Starting at $0.18/M input tokens
Starting at $0.18/M output tokens
Streaming
Vision
JSON Output
No ratings yetSign in to rate

Select Provider

All Providers for Llama Guard 4 12B

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

S

OpenRouter
Context: 163.8k
Input
$0.18
/M tokens
Cache Read
/M tokens
Output
$0.18
/M tokens
Get Started

Frequently asked questions

What is Llama Guard 4 12B?

Safety-focused model for content moderation. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Llama Guard 4 12B cost?

Pricing for Llama Guard 4 12B on LLM Gateway starts at $0.18 per million input tokens and $0.18 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Llama Guard 4 12B?

Llama Guard 4 12B supports a context window of up to 163,840 tokens on its largest provider deployment.

Which providers serve Llama Guard 4 12B?

Llama Guard 4 12B is served by OpenRouter through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Llama Guard 4 12B support tool calling and structured outputs?

Llama Guard 4 12B supports structured JSON outputs, but not tool calling.

When was Llama Guard 4 12B released?

Llama Guard 4 12B was released on April 30, 2025.