Documentation Navigation

Models & pricing / Models / Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95B Overview

This model's data has not been fully verified and benchmarks may be missing.

Qwen3.8-2.4T-A95B is a chat-reasoning AI model created by Qwen, featuring 2.4T parameters and a context window of unknown.

Qwen3.8-2.4T-A95B is a 2.4T-parameter Mixture-of-Experts (MoE) text generation model developed by Qwen. Built on the Qwen3_5MoeForCausalLM architecture using transformers. Released on 2026-08-08 with 335 likes and 978 downloads on Hugging Face.

Model Lineage & Specification

qwen38-24t-a95b is Qwen's primary release in the current family.

DeveloperQwen
API Identifierqwen38-24t-a95b
Parameters2.4T
Context Windowunknown
LicenseUnknown

Comparable models

Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95B is a 2.4T-parameter Mixture-of-Experts (MoE) text generation m...

API IDqwen38-24t-a95b
Typeopen weights
Parameters2.4T
Contextunknown

Solar Pro 4

Solar Pro 4 is a large language model from Upstage built for enterprise workflow...

API IDupstage-solar-pro-4
Typeapi only
Parameters
Context524288

Qwen3.8 Max Preview

Preview Qwen flagship for million-token multimodal reasoning and long-horizon ag...

API IDalibaba-qwen3.8-max-preview
Typeapi only
Parameters
Context1M tokens

Nemotron 3 Ultra

NVIDIA Nemotron 3 Ultra is a flagship open-weight frontier LLM family engineered...

API IDnvidia-nemotron-3-ultra
Typeopen weights
Parameters550B total / 55B active
Context1M tokens

Qwen3.8-2.4T-A95B is a 2.4T-parameter Mixture-of-Experts (MoE) text generation model developed by Qwen. Built on the Qwen3_5MoeForCausalLM architecture using transformers. Released on 2026-08-08 with 335 likes and 978 downloads on Hugging Face.