Documentation Navigation
Models & pricing / Models / Qwen3.8-2.4T-A95B
Qwen3.8-2.4T-A95B Overview
This model's data has not been fully verified and benchmarks may be missing.
Qwen3.8-2.4T-A95B is a chat-reasoning AI model created by Qwen, featuring 2.4T parameters and a context window of unknown.
Qwen3.8-2.4T-A95B is a 2.4T-parameter Mixture-of-Experts (MoE) text generation model developed by Qwen. Built on the Qwen3_5MoeForCausalLM architecture using transformers. Released on 2026-08-08 with 335 likes and 978 downloads on Hugging Face.
Model Lineage & Specification
qwen38-24t-a95b is Qwen's primary release in the current family.
qwen38-24t-a95bComparable models
| Feature | Qwen3.8-2.4T-A95B | Solar Pro 4 | Qwen3.8 Max Preview | Nemotron 3 Ultra |
|---|---|---|---|---|
| Description | Qwen3.8-2.4T-A95B is a 2.4T-parameter Mixture-of-Experts (MoE) text generation m... | Solar Pro 4 is a large language model from Upstage built for enterprise workflow... | Preview Qwen flagship for million-token multimodal reasoning and long-horizon ag... | NVIDIA Nemotron 3 Ultra is a flagship open-weight frontier LLM family engineered... |
| API Identifier | qwen38-24t-a95b | upstage-solar-pro-4 | alibaba-qwen3.8-max-preview | nvidia-nemotron-3-ultra |
| Parameters | 2.4T | — | — | 550B total / 55B active |
| Context Window | unknown | 524288 | 1M tokens | 1M tokens |
| License / Type | open weights | api only | api only | open weights |
Qwen3.8-2.4T-A95B
Qwen3.8-2.4T-A95B is a 2.4T-parameter Mixture-of-Experts (MoE) text generation m...
qwen38-24t-a95bSolar Pro 4
Solar Pro 4 is a large language model from Upstage built for enterprise workflow...
upstage-solar-pro-4Qwen3.8 Max Preview
Preview Qwen flagship for million-token multimodal reasoning and long-horizon ag...
alibaba-qwen3.8-max-previewNemotron 3 Ultra
NVIDIA Nemotron 3 Ultra is a flagship open-weight frontier LLM family engineered...
nvidia-nemotron-3-ultraQwen3.8-2.4T-A95B is a 2.4T-parameter Mixture-of-Experts (MoE) text generation model developed by Qwen. Built on the Qwen3_5MoeForCausalLM architecture using transformers. Released on 2026-08-08 with 335 likes and 978 downloads on Hugging Face.