Compare Models
Evaluate leading AI models side-by-side. Compare context windows, open-weights licensing, parameters, and verified benchmarks to choose the right model for your application.
Comparing 1 of 4 models
Model Comparison Matrix
Add a model | Add a model | Add a model | ||
|---|---|---|---|---|
| Type | API Only | |||
| Release Date | Jul 9, 2026 | |||
| License | Proprietary | |||
| Parameters | undisclosed | |||
| Modalities | textimagepdf | |||
| Context Window | 1.1M tokens | |||
| Primary Task | chat reasoning | |||
| Release Date | 2026-07-09 | |||
| Agents' Last Exam | 50.3 (score) | |||
| Artificial Analysis Coding Agent Index | 74.6 (index score) | |||
| Artificial Analysis Intelligence Index | 51.2 (index score) | |||
| BrowseComp | 83.3 (accuracy) | |||
| DeepSWE | 67.2 (resolve rate) | |||
| FrontierMath | 78.6 (accuracy) | |||
| GPQA Diamond | 92.3 (accuracy) | |||
| MMMU Pro | 78.4 (accuracy) | |||
| OSWorld | 45.6% | |||
| SWE-Bench Pro | 62.7 (resolve rate) | |||
| Terminal-Bench | 84.7% | |||
| Toolathlon | 53.4% |