NVIDIA·nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4
NVIDIA Nemotron 3 Ultra 550B A55B NVFP4 is a language model published by NVIDIA. The latest OpenRouter endpoint refresh was unavailable, so no current serving claim is made. It accepts text input and returns text. The published context limit is 512K tokens. The catalogue declares support for tool use.
OpenRouter endpoint status could not be refreshed. This model is not treated as currently served until a successful observation arrives.
Providers
Uptime measured over the last 30 minutesCurrent provider status is unavailable.
Hugging Face Inference Providers
Exact repository identityProvider economics and telemetry published by the Hugging Face Router. These are operational signals, not benchmark results.
Reported evaluation leads (5)
Excluded from scoreStructured rows reported in this exact model repository at the pinned revision below. They are useful research leads, but remain quarantined until dataset revision, harness, configuration and identity pass ModelCap’s independent admission review.
ModelCap Index · Current board
Methodology- Public rank basis
- Measured (measured)
- Index support
- 76.2% · strong
- Index interval
- 71.3–79.1
- Rank posterior
- Not available
- Top 5 / Top 10 probability
- Not available / Not available
- Posterior as of
- Snapshot timestamp unavailable
- Identity binding
- Exact catalogue product · nvidia/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4 · aggregates disclosed benchmark configurations · endpoint configuration not claimed · artifact metadata nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4@183968f87ae4cedce3039313cac1fd43d112c578 (not the evaluation revision)
- Direct benchmark observations
- 2
- Breadth-adjusted BES point estimate
- 75.2
BES normalizes admitted public benchmark evidence for measured Index rows. Its legacy score and observables prior do not define the public language rank.
- Observed capability
- 75.2
- Evidence status
- confirmed · 81% mass
- General preference
- 75.0
- Coding
- 75.8
- Agents & tools
- —
- Reasoning
- —
- Evidence breadth
- 92%
- Evidence coverage
- 76%
- Usage (OpenRouter popularity)
- —
- Liquidity (providers × uptime)
- —
- Open reach (HF downloads)
- —
- Surface (context / tools / modalities)
- 55.1
- Freshness
- 78.0
- · Catalogue: context 512288, tools
- · First seen 2026-06-03
These adoption and deployment observations remain context only. They do not change this model's ModelCap Index score or rank.
- Arena Overall75.0
- Arena Coding75.8
The bounded 0–100 capability score blends each source’s competitive placement with its published achievement, then combines capability families. Evidence breadth adds a modest uncertainty adjustment; price and popularity are not part of benchmark-led rank.
Deployment readiness
Metadata index- Artifact reproducibility
- 85.0
- Access & legal clarity
- 92.5
- Deployability
- 64.8
- Evaluation provenance
- 50.1
Missing: artifactReproducibility.base-lineage, deployability.measured-provider-uptime, evaluationProvenance.evaluation-dataset-revisions.
A metadata completeness and deployability index—not a safety certification, quality grade, or production-readiness claim.
Market activity
Venue coverage- Market Gravity
- 11.4
- OpenRouter weekly popularity
- Not listed
Market Gravity uses OpenRouter’s full-catalogue popularity order. Sparse Vercel top-ten observations appear only when published and do not affect the score.
Arena preference
- Status
- Ranked
- Source
- Arena
- Official rank
- #96 of 385
- Rating
- 1426.0
- 95% confidence interval
- 1419–1433
- Votes
- 11K
- Source snapshot
- 6 Aug 2026, 13:00 UTC
- Category
- overall
- Identity match confidence
- 100%
Market Gravity breakdown
- Usage55%
- 0.0
- Liquidity25%
- 0.0
- Open reach15%
- 50.0
- Freshness5%
- 78.0
OpenRouter weekly popularity across the full model catalogue.
Independent providers versus the model's open or closed cohort, weighted by uptime.
Hugging Face 30-day downloads for open models; neutral for closed models.
Time since first public availability, on a six-month half-life.
Specification
- Context window
- 512K tokens
- Max output
- —
- Inputs
- Text
- Outputs
- Text
- Tokenizer
- —
- Cached input tokens
- Not offered
- First seen on OpenRouter
- 3 Jun 2026
Capabilities
- Not supported: Reasoning
- Supported: Tool use
- Not supported: Structured output
- Not supported: Response format
- Not supported: Moderated
Weights & access
Restricted license. Downloadable weights use a model-specific or use-restricted license. Review the linked license before use, redistribution, or commercial deployment.
- Access
- Restricted license
- Licence
- openmdw-1.1
- Pinned revision
- 183968f87ae4
- Parameters
- 335.0B
- Architecture
- NemotronHForCausalLM
- Model type
- nemotron_h
- Hugging Face downloads (30d)
- 209K
- Downloads (all time)
- 768K
- Likes
- 282
- Repository updated
- 24 Jun 2026, 19:23 UTC