Nemotron Cascade 2 30B A3B
Nemotron Cascade 2 30B A3B
📌 Model Overview
Nemotron model for efficient reasoning, coding, and specialized AI agents
Nemotron Cascade 2 30B A3B is a Open Weights model developed by NVIDIA, released on 2026-03-24. It is designed primarily for Chat Reasoning workloads. Featuring a 256K tokens context window and 30B parameter count, it offers robust performance for enterprise integration, developers, and researchers.
✨ Key Features & Capabilities
| Feature | Description |
|---|---|
| Context Window | 256K tokens capacity for extended prompts and multi-turn workflows |
| Primary Task | Optimized for Chat Reasoning |
| Deployment | self-hostable, api-only |
| Modality | text |
| Max Output | Max Output: 33K tokens |
| Native reasoning capability | Native reasoning capability |
| Tool / function calling support | Tool / function calling support |
⚙️ Technical Specifications
| Specification | Details |
|---|---|
| Developer / Lab | NVIDIA |
| Release Date | 2026-03-24 |
| Model Type | Open Weights |
| Parameters | 30B |
| Context Window | 256K tokens |
| License | proprietary |
| Model Family | nemotron |
📜 License & Usage
This model is governed by the proprietary license. Please check official developer guidelines before commercial deployment.
Key Features
Max Output: 33K tokens
Native reasoning capability
Tool / function calling support
You might also want to compare
Verified Sources
Tags
Model Specs
Parameters
30B
Context Window
256K tokens
License
proprietary
Deployment
Lineage
Model Family
Part of the nemotron family
Curator Notes
Skeleton entry imported from models.dev backfill (2026-07-26). Needs manual enrichment and primary-source verification.
Compare Specs
Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.
Compare Model