NVIDIA·nvidia/Llama-3.1-405B-Instruct-FP8
Llama 3.1 405B Instruct FP8 weight access is restricted license, from ModelCap classification of published repository metadata as of 22 Sept 2026, 03:10 UTC.
Snapshot facts · download the public dataset · methodology
OpenRouter endpoint status could not be refreshed. This model is not treated as currently served until a successful observation arrives.
Providers
Uptime measured over the last 30 minutesCurrent provider status is unavailable.
Deployment readiness
Metadata index- Artifact reproducibility
- 100.0
- Access & legal clarity
- 67.5
- Deployability
- 8.0
- Evaluation provenance
- 0.0
Missing: accessLegalClarity.resolvable-license-terms, deployability.independent-provider-records, deployability.measured-provider-uptime, deployability.published-provider-pricing, deployability.declared-capacity, evaluationProvenance.admitted-evaluation-observations, evaluationProvenance.capability-family-breadth, evaluationProvenance.evaluation-source-diversity, evaluationProvenance.identity-and-source-confidence, evaluationProvenance.evaluation-dataset-revisions.
A metadata completeness and deployability index—not a safety certification, quality grade, or production-readiness claim.
Market activity
Venue coverage- Market Gravity
- 7.8
- OpenRouter weekly popularity
- Not listed
Market Gravity uses OpenRouter’s full-catalogue popularity order. Sparse Vercel top-ten observations appear only when published and do not affect the score.
Market Gravity breakdown
- Usage55%
- 0.0
- Liquidity25%
- 0.0
- Open reach15%
- 50.0
- Freshness5%
- 5.7
OpenRouter weekly popularity across the full model catalogue.
Independent providers versus the model's open or closed cohort, weighted by uptime.
Hugging Face 30-day downloads for open models; neutral for closed models.
Time since first public availability, on a six-month half-life.
Specification
- Context window
- — tokens
- Max output
- —
- Inputs
- Text
- Outputs
- Text
- Tokenizer
- —
- Cached input tokens
- Not offered
- First seen on OpenRouter
- 29 Aug 2024
Capabilities
- Not supported: Reasoning
- Not supported: Tool use
- Not supported: Structured output
- Not supported: Response format
- Not supported: Moderated
Weights & access
Restricted license. Downloadable weights use a model-specific or use-restricted license. Review the linked license before use, redistribution, or commercial deployment.
- Access
- Restricted license
- Licence
- llama3.1
- Pinned revision
- 55a1cd4cc02c
- Parameters
- 405.9B
- Architecture
- LlamaForCausalLM
- Model type
- llama
- Hugging Face downloads (30d)
- —
- Downloads (all time)
- —
- Likes
- —
- Repository
- nvidia/Llama-3.1-405B-Instruct-FP8
- Repository updated
- 22 Aug 2025, 06:13 UTC
- Base lineage
- meta-llama/Llama-3.1-405B-Instruct
Metadata status: stale · Hugging Face quota reset exceeds the metadata stage deadline.