NVIDIA · Nvidia Nemotron 3 Ultra 550b A55b Bf16 · active

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 is an open-weight language model published by NVIDIA.

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
Statusactive
InputText
OutputText
Context256K
Total / active parameters560.5B / 55B
Licenseother
APIUnknown
Open weights / self-hostYes / Yes

Provider availability

Serving endpoints and pricing

No directly sourced provider offerings yetPricing and availability stay unknown until collected from each serving provider.

This page represents one immutable developer model ID. Serving endpoints, pricing, regions, quantizations, and service tiers attach separately and never create duplicate model records.

Technical specification

Developer model IDnvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
VersionNVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
ReleasedJun 4, 2026
ArchitectureNemotronHForCausalLM
Maximum outputUnknown
Knowledge cutoffNot reported

Capabilities

chatgenerationreasoningtools
ReasoningYes
VisionNo
Tool callingYes

Version lineage

PredecessorUnknown
SuccessorsUnknown
Aliasesnvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16, nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16

Explicit unknowns

  • Maximum output
  • Knowledge cutoff
  • API availability
  • Provider pricing

Comparable peers

Related model comparisons

All comparisons →

Questions

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 FAQs

What is NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16?+

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 is an open-weight language model published by NVIDIA. It is developed by NVIDIA and its current sourced lifecycle status is active.

What inputs and outputs does NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 support?+

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16's current record lists Text as input and Text as output.

How much context does NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 support?+

The current record lists a 262,144-token context window and does not report a maximum output length.

Is NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 available through an API?+

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16's current primary-source record does not establish API availability. No provider endpoint is treated as available without direct provider evidence.

How large is NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16?+

The current record lists 560.5B total parameters and 55B active parameters. Its architecture is NemotronHForCausalLM.

What capabilities and tasks does NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 support?+

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16's recorded capabilities are chat, generation, reasoning, and tools. Its supported tasks are text.

Can NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 be self-hosted?+

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 is an open-weight model. Self-hosting is supported, and the recorded license is other.

Does NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 have related versions or aliases?+

No predecessor is recorded. No successor is recorded. Recorded aliases are nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16.

Which models can NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 be compared with?+

The current compatible peer set includes NVIDIA-Nemotron-3-Super-120B-A12B-BF16, NVIDIA-Nemotron-3-Nano-30B-A3B-BF16, NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16, NVIDIA-Nemotron-3-Super-120B-A12B-Base-BF16, and GLM-5. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons

How current is the Model Markets record for NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16?+

NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 was last verified Aug 28, 2026 from NVIDIA's primary source, “nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 developer model repository.” The evidence is classified as official fact.

Primary evidence

Sources and observation date