MODELS
by NVIDIA
NVIDIA Nemotron 3 Ultra is an open 550B-parameter MoE language model with 55B active parameters, 1M context, configurable reasoning, and tool-use support.
NVIDIA Research presents Nemotron 3 Ultra as its most capable Nemotron 3 model, with 550B total parameters and 55B active parameters. NVIDIA’s launch blog describes it as an open MoE model for reasoning, orchestration, long-running agents, coding, tool use, and long-context work. The NVIDIA NIM model card lists text input/output, 1M context, configurable reasoning, tool use, supported languages, OpenMDW-1.1 license, and a June 4, 2026 release date. NVIDIA API documentation provides the nvidia/nemotron-3-ultra-550b-a55b chat-completions endpoint.
Evaluate NVIDIA’s highest-capability Nemotron 3 model profile for large-scale reasoning and long-context workloads. Build agentic workflows that need reasoning orchestration coding tool use or long-running agent behavior as described by NVIDIA. Use an NVIDIA NIM/API-hosted model with official model card API reference deployment guide and downloadable checkpoint coverage.
Last updated Jul 23, 2026