Skip to content
ModelCap

Llama 4 Maverick 17B 128E Instruct FP8

Official build · attached, unrankedGated access

Meta·meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8

Llama 4 Maverick 17B 128E Instruct FP8 has a published context limit of 1M tokens in the ModelCap catalogue as of 22 Sept 2026, 21:31 UTC.

Llama 4 Maverick 17B 128E Instruct FP8 weight access is gated access, from ModelCap classification of published repository metadata as of 22 Sept 2026, 21:31 UTC.

Snapshot facts · download the public dataset · methodology

ModelCap Index
See primary model · Llama 4 Maverick
This release is visible for discovery, not ranked as a peer model
Market Gravity
8.1
Secondary market signal

This separately configured or accelerated route is reviewed as the same underlying model identity as Llama 4 Maverick. That canonical product owns rank and capability evidence; pricing, providers and other operational differences remain route-specific below.

OpenRouter endpoint status could not be refreshed. This model is not treated as currently served until a successful observation arrives.

Input
Self-hosted
weights only · no listed API price
Output
Self-hosted
weights only · no listed API price
Context
1M
tokens
Providers
0
endpoints
Gateway spend
latest · Vercel share

Providers

Uptime measured over the last 30 minutes

Current provider status is unavailable.

Hugging Face Inference Providers

Exact repository identity

Provider economics and telemetry published by the Hugging Face Router. These are operational signals, not benchmark results.

ProviderInputOutputContextTTFTThroughputFeatures
novita
live
$0.27$0.851M581 ms79.5 tok/sstructured

Reported evaluation leads (1)

Not independent measurement

Structured rows reported in this exact model repository at the pinned revision below. They are the publisher talking about itself: never measured evidence, never a score input on their own. Rows marked as used were laddered against independently measured peers inside the launch-card channel of this model's modeled placement, with the replay-measured optimism haircut applied; the others played no part.

Dataset / taskMetricValueRevisionProvenance
Idavidrein/gpqa
diamond
53.0303
Unpinned
not published
Reported
Not used in launch estimate

Deployment readiness

Metadata index
58.3
partial · 69% metadata coverage
Not capability
Artifact reproducibility
100.0
Access & legal clarity
68.3
Deployability
48.7
Evaluation provenance
0.0

Missing: accessLegalClarity.resolvable-license-terms, deployability.measured-provider-uptime, evaluationProvenance.admitted-evaluation-observations, evaluationProvenance.capability-family-breadth, evaluationProvenance.evaluation-source-diversity, evaluationProvenance.identity-and-source-confidence, evaluationProvenance.evaluation-dataset-revisions.

A metadata completeness and deployability index—not a safety certification, quality grade, or production-readiness claim.

Market activity

Venue coverage
Market Gravity
8.1
OpenRouter weekly popularity
Not listed

Market Gravity uses OpenRouter’s full-catalogue popularity order. Sparse Vercel top-ten observations appear only when published and do not affect the score.

Market Gravity breakdown

Usage55%
0.0

OpenRouter weekly popularity across the full model catalogue.

Liquidity25%
0.0

Independent providers versus the model's open or closed cohort, weighted by uptime.

Open reach15%
50.0

Hugging Face 30-day downloads for open models; neutral for closed models.

Freshness5%
12.8

Time since first public availability, on a six-month half-life.

Specification

Context window
1M tokens
Max output
Inputs
Text, Image
Outputs
Text
Tokenizer
Cached input tokens
Not offered
First seen on OpenRouter
1 Apr 2025

Capabilities

  • Not supported: Reasoning
  • Not supported: Tool use
  • Not supported: Structured output
  • Not supported: Response format
  • Not supported: Moderated

Weights & access

Gated access. The weights are hosted behind an access request or acceptance step. Review the publisher's access requirements and license before use or redistribution.

Access
Gated access
Licence
llama4
Pinned revision
94125d2bd830
Parameters
401.6B
Architecture
Llama4ForConditionalGeneration
Model type
llama4
Hugging Face downloads (30d)
52K
Downloads (all time)
1.7M
Likes
180
Repository updated
22 May 2025, 23:46 UTC