How does Llama Guard 4 12B rank among AI models?
Llama Guard 4 12B has no public ModelCap Index position as of 22 September 2026: the snapshot lacks enough matched benchmark evidence, or the identity match is unresolved. Its catalogue facts (price, context, providers) are shown without a rank.
How much does Llama Guard 4 12B cost per 1M tokens?
Llama Guard 4 12B is listed at $0.18 per 1M input tokens and $0.18 per 1M output tokens as of 22 September 2026. Prices come from the OpenRouter catalogue and re-observe on every refresh. For 1,000 requests using 2,000 input and 500 output tokens each, the listed rates imply $0.45 total (2M input + 0.5M output). This excludes caching, batch discounts, prompt-length tiers, tool charges and retries; verify the chosen endpoint's rate and limits before budgeting.
What is the context window of Llama Guard 4 12B?
Llama Guard 4 12B has a published context window of 164K tokens (163,840) and a maximum output of 16K tokens. Individual providers can serve less than the published maximum; the providers table lists each endpoint's own limit.
Which API providers serve Llama Guard 4 12B?
One provider serves Llama Guard 4 12B through OpenRouter as of 22 September 2026: DeepInfra. Each provider's price, context limit, quantization and measured uptime are in the providers table above.
When was Llama Guard 4 12B released?
Llama Guard 4 12B first appeared in the catalogue on 30 Apr 2025, with a published knowledge cutoff of 2024-08-31. It is the current version in its family. Rank and price movements since then are recorded on the site's changes feed.