Qwen3.8-Max
Qwen3.8-Max is Alibaba's flagship 2.4T-parameter MoE model (95B active) with a 1M-token context, accepting text, image, and video inputs. Designed for real-world work, research, and autonomous coding, it reportedly ran for 16 days without human intervention to build a coding tool; API access costs $2/M input and $6/M output via QwenCloud or Alibaba Cloud Model Studio, with open weights on Hugging Face (Qwen/Qwen3.8-2.4T-A95B) and a 27B distillation, plus the QwenWork enterprise platform in public beta.