MODEL SIGNAL · QWEN · NEW
Qwen3.8-2.4T-A95B
Qwen3.8-2.4T-A95B is a sparse Mixture-of-Experts causal language model from Qwen with 2.4 trillion total parameters, 95 billion activated parameters per token, and open-weight availability.
CATEGORYMultimodal
CONTEXT1010000
RELEASEDAugust 13, 2026
Key Features
- Sparse Mixture-of-Experts (MoE) causal LM
- 2.4 trillion total parameters
- 95 billion activated parameters per token
- 262,144-token native context window
- Extensible up to 1,010,000 tokens
- Open-weight availability
Read the Model Signal report →
The designed Model Signal report for Qwen3.8-2.4T-A95B is still in review. It will publish here after the fact and quality gates clear.