How does Gemma 3 4B rank among AI models?
Gemma 3 4B holds ModelCap Index position #183 of 248 ranked language models as of 22 September 2026, with an Index score of 22.1 (interval 16.1–28.1); evidence: measured. The Index combines public benchmark boards with published uncertainty; the methodology page explains the weighting.
How much does Gemma 3 4B cost per 1M tokens?
Gemma 3 4B is listed at $0.05 per 1M input tokens and $0.10 per 1M output tokens as of 22 September 2026. Prices come from the OpenRouter catalogue and re-observe on every refresh. For 1,000 requests using 2,000 input and 500 output tokens each, the listed rates imply $0.15 total (2M input + 0.5M output). This excludes caching, batch discounts, prompt-length tiers, tool charges and retries; verify the chosen endpoint's rate and limits before budgeting.
What is the context window of Gemma 3 4B?
Gemma 3 4B has a published context window of 131K tokens (131,072) and a maximum output of 16K tokens. Individual providers can serve less than the published maximum; the providers table lists each endpoint's own limit.
Which API providers serve Gemma 3 4B?
One provider serves Gemma 3 4B through OpenRouter as of 22 September 2026: DeepInfra. Each provider's price, context limit, quantization and measured uptime are in the providers table above.
How does Gemma 3 4B compare with granite guardian 4.1 8b?
Gemma 3 4B ranks #183 (Index 22.1) and granite guardian 4.1 8b ranks #182 (Index 23.5) on the ModelCap Index as of 22 September 2026. The head-to-head page lines up their benchmarks, listed prices, context windows and provider counts side by side.
When was Gemma 3 4B released?
Gemma 3 4B first appeared in the catalogue on 13 Mar 2025, with a published knowledge cutoff of 2024-08-31. It is the current version in its family. Rank and price movements since then are recorded on the site's changes feed.