Compare Models
Evaluate leading AI models side-by-side. Compare context windows, open-weights licensing, parameters, and verified benchmarks to choose the right model for your application.
Comparing 1 of 4 models
Model Comparison Matrix
Add a model | Add a model | Add a model | ||
|---|---|---|---|---|
| Type | API Only | |||
| Release Date | Feb 19, 2026 | |||
| License | Proprietary | |||
| Parameters | undisclosed | |||
| Modalities | textimagevideoaudiopdf | |||
| Context Window | 1.0M tokens | |||
| Primary Task | chat reasoning | |||
| Release Date | 2026-02-19 | |||
| ARC-AGI-2 | 77.1 (accuracy) | |||
| Artificial Analysis Coding Agent Index | 43 (average pass@1) | |||
| CharXiv Reasoning | 83.3 (accuracy) | |||
| GDPval-AA | 1314 (Elo) | |||
| GPQA Diamond | 94.3 (accuracy) | |||
| Humanity's Last Exam | 44.4 (accuracy) | |||
| MCP Atlas | 78.2% | |||
| MMMU Pro | 80.5 (accuracy) | |||
| OSWorld-Verified | 76.2% | |||
| SWE-Atlas Codebase QnA | 13.5 (score) | |||
| SWE-Atlas Refactoring | 33.81 (score) | |||
| SWE-Atlas Test Writing | 29.84 (score) | |||
| SWE-Bench Pro | 54.2 (resolve rate) | |||
| Terminal-Bench | 70.3% |