Qwen3.8-2.4T
- AI
- Open Source
- Infrastructure
- Developer Tools
Qwen3.8-2.4T-A95B is an open-weight mixture-of-experts model with 2.4 trillion total parameters and about 95 billion active parameters per inference step. Alibaba published FP8 and BF16 checkpoints, and the model card positions it near frontier proprietary systems on coding and reasoning benchmarks. The catch is that the open release is not the full product. Vision, built-in tools, non-thinking mode, and the default 1M-token context stay with Qwen3.8-Max, while the open weights ship with a 250k context cap and a more awkward serving profile than rivals like Kimi K3, which launched with friendlier 4-bit quantization.
Treat this as a strong new option for API-based use and for providers benchmarking the frontier, not as a practical local model for most teams. If you care about self-hosting or edge deployment, the more consequential release to watch is the promised 27B variant and whether good 4-bit quants appear quickly.
-
huggingface.co
- Discuss on HN