Nemotron 3 Ultra 550B A55B
Nemotron 3 Ultra 550B A55B
📌 Model Overview
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Nemotron 3 Ultra 550B A55B is a Open Weights model developed by NVIDIA, released on 2026-06-04. It is designed primarily for Chat Reasoning workloads. Featuring a 1M tokens context window and 550B parameter count, it offers robust performance for enterprise integration, developers, and researchers.
✨ Key Features & Capabilities
| Feature | Description |
|---|---|
| Context Window | 1M tokens capacity for extended prompts and multi-turn workflows |
| Primary Task | Optimized for Chat Reasoning |
| Deployment | self-hostable, api-only |
| Modality | text |
| Max Output | Max Output: 128K tokens |
| Native reasoning capability | Native reasoning capability |
| Tool / function calling support | Tool / function calling support |
⚙️ Technical Specifications
| Specification | Details |
|---|---|
| Developer / Lab | NVIDIA |
| Release Date | 2026-06-04 |
| Model Type | Open Weights |
| Parameters | 550B |
| Context Window | 1M tokens |
| License | proprietary |
| Model Family | nemotron |
📊 Benchmarks & Performance
| Benchmark | Score | Source |
|---|---|---|
| SWE-Bench Verified | 70.7% | Independent Eval |
| SWE-Bench Multilingual | 67.7 (resolve rate) | Independent Eval |
| Terminal-Bench | 56.4% | Independent Eval |
| GPQA | 87 (accuracy) | Independent Eval |
| Humanity's Last Exam | 26.7 (accuracy) | Independent Eval |
| Humanity's Last Exam | 37.4 (accuracy) | Independent Eval |
| LiveCodeBench | 89 (pass@1) | Independent Eval |
| MMLU-Pro | 86.8 (accuracy) | Independent Eval |
| BrowseComp | 44.4 (accuracy) | Independent Eval |
| IFBench | 81.7 (accuracy) | Independent Eval |
| GDPval | 46.7 (wins or ties) | Independent Eval |
💰 Pricing
| Tier / Unit | Rate (USD) |
|---|---|
| 1M input tokens | $0.5 |
| 1M output tokens | $2.5 |
| 1M cache read tokens | $0.15 |
📜 License & Usage
This model is governed by the proprietary license. Please check official developer guidelines before commercial deployment.
Key Features
Max Output: 128K tokens
Native reasoning capability
Tool / function calling support
You might also want to compare
Verified Sources
Tags
Model Specs
Parameters
550B
Context Window
1M tokens
License
proprietary
Deployment
Pricing
- 1M input tokens$0.5
- 1M output tokens$2.5
- 1M cache read tokens$0.15
as of 2026-07-26
Lineage
Model Family
Part of the nemotron family
Curator Notes
Skeleton entry imported from models.dev backfill (2026-07-26). Needs manual enrichment and primary-source verification.
Compare Specs
Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.
Compare Model