Nemotron 3 Super 120B A12B
Nemotron 3 Super 120B A12B
📌 Model Overview
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Nemotron 3 Super 120B A12B is a Open Weights model developed by NVIDIA, released on 2026-03-11. It is designed primarily for Chat Reasoning workloads. Featuring a 262K tokens context window and 120B parameter count, it offers robust performance for enterprise integration, developers, and researchers.
✨ Key Features & Capabilities
| Feature | Description |
|---|---|
| Context Window | 262K tokens capacity for extended prompts and multi-turn workflows |
| Primary Task | Optimized for Chat Reasoning |
| Deployment | self-hostable, api-only |
| Modality | text |
| Max Output | Max Output: 262K tokens |
| Native reasoning capability | Native reasoning capability |
| Tool / function calling support | Tool / function calling support |
⚙️ Technical Specifications
| Specification | Details |
|---|---|
| Developer / Lab | NVIDIA |
| Release Date | 2026-03-11 |
| Model Type | Open Weights |
| Parameters | 120B |
| Context Window | 262K tokens |
| License | proprietary |
| Model Family | nemotron |
💰 Pricing
| Tier / Unit | Rate (USD) |
|---|---|
| 1M input tokens | $0.2 |
| 1M output tokens | $0.8 |
📜 License & Usage
This model is governed by the proprietary license. Please check official developer guidelines before commercial deployment.
Key Features
Max Output: 262K tokens
Native reasoning capability
Tool / function calling support
You might also want to compare
Verified Sources
Tags
Model Specs
Parameters
120B
Context Window
262K tokens
License
proprietary
Deployment
Pricing
- 1M input tokens$0.2
- 1M output tokens$0.8
as of 2026-07-26
Lineage
Model Family
Part of the nemotron family
Curator Notes
Skeleton entry imported from models.dev backfill (2026-07-26). Needs manual enrichment and primary-source verification.
Compare Specs
Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.
Compare Model