Llama 3.3 Nemotron Super 49B v1.5
Llama 3.3 Nemotron Super 49B v1.5
📌 Model Overview
Nemotron model for efficient reasoning, coding, and specialized AI agents
Llama 3.3 Nemotron Super 49B v1.5 is a Open Weights model developed by NVIDIA, released on 2025-07-25. It is designed primarily for Chat Reasoning workloads. Featuring a 131K tokens context window and 49B parameter count, it offers robust performance for enterprise integration, developers, and researchers.
✨ Key Features & Capabilities
| Feature | Description |
|---|---|
| Context Window | 131K tokens capacity for extended prompts and multi-turn workflows |
| Primary Task | Optimized for Chat Reasoning |
| Deployment | self-hostable, api-only |
| Modality | text |
| Max Output | Max Output: 131K tokens |
| Native reasoning capability | Native reasoning capability |
| Tool / function calling support | Tool / function calling support |
⚙️ Technical Specifications
| Specification | Details |
|---|---|
| Developer / Lab | NVIDIA |
| Release Date | 2025-07-25 |
| Model Type | Open Weights |
| Parameters | 49B |
| Context Window | 131K tokens |
| License | proprietary |
| Model Family | nemotron |
💰 Pricing
| Tier / Unit | Rate (USD) |
|---|---|
| 1M input tokens | $0.1 |
| 1M output tokens | $0.4 |
📜 License & Usage
This model is governed by the proprietary license. Please check official developer guidelines before commercial deployment.
Key Features
Max Output: 131K tokens
Native reasoning capability
Tool / function calling support
You might also want to compare
Verified Sources
Tags
Model Specs
Parameters
49B
Context Window
131K tokens
License
proprietary
Deployment
Pricing
- 1M input tokens$0.1
- 1M output tokens$0.4
as of 2026-07-26
Lineage
Model Family
Part of the nemotron family
Curator Notes
Skeleton entry imported from models.dev backfill (2026-07-26). Needs manual enrichment and primary-source verification.
Compare Specs
Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.
Compare Model