Qwen releases open-weight Qwen3.8-2.4T-A95B, a 2.4T-parameter MoE model
Qwen/Qwen3.8-2.4T-A95B
AISummary
Qwen has released the Qwen3.8-2.4T-A95B model weights on Hugging Face, with 2.4T total and 95B activated parameters in a mixture-of-experts design. The release supports reasoning_effort levels and a 262,144-token native context extensible to 1,010,000 tokens, and it is text-only with thinking mode always on. The source reports benchmark results against Opus 4.8, Fable 5, GPT 5.6 Sol, and Qwen3.7-Max, and says the official Qwen3.8-Max API adds vision input and a 1M default context.
AIWhy it matters
The model card gives parameters, architecture, reasoning controls, and benchmark tables against named rival models, showing what an open release of this scale actually offers.
Source: Qwen · new models on Hugging Face · huggingface.co