Qwen3.8-2.4T-A95B is the open source version of Tongyi Qianwen’s latest flagship series, released in August 2026. It uses a sparse MoE architecture with a total parameter volume of 2.4 trillion, activating approximately 95 billion parameters at each step, and supports 1 million token context and hybrid attention mechanisms. Core benchmarks include GPQA Diamond 92.6, PaperBench 93.0, OSWorld 86.1, and BabyVision 82.0. Ranked 4th globally on CodeArena.
- Input:
- Output:
- Input:
- $2$1.8/1M
- Output:
- $6$5.4/1M
- Context length:
- 1M
- Max output:
- 64K