Home / AI & Machine Learning

Photo of tablet, television, cryptocurrency
Image: via imageio.forbes.com
AI & Machine Learning

Alibaba Unveils Qwen3.8-Max AI Model at Competitive Price

WireByte Staff · August 5, 2026

Alibaba has released Qwen3.8-Max, a 2.4 trillion parameter generative AI model, at a lower price point than OpenAI and Anthropic's top-tier offerings. The model costs $2 per million input tokens and $6 per million output tokens, sparking a price war in the AI market. This move is seen as a subsidy for lower-layer companies and a challenge to the dominance of OpenAI and Anthropic.

Key points

  • Alibaba released Qwen3.8-Max, a 2.4 trillion parameter generative AI model, on February 12, 2026.
  • The model costs $2 per million input tokens and $6 per million output tokens, making it cheaper than OpenAI's GPT-5.6 Sol and Anthropic's Claude Opus 5.
  • Alibaba's pricing strategy is seen as a subsidy for lower-layer companies and a challenge to the dominance of OpenAI and Anthropic.
  • Qwen3.8-Max is built on a sparse mixture-of-experts design, routing each query through 95 billion active parameters to keep serving costs and latency down.
  • The model handles text, images, and video across a context window of one million tokens.

Alibaba's Qwen3.8-Max AI model has been released, marking a significant development in the generative AI market. The model boasts 2.4 trillion parameters, making it the largest built by Alibaba to date. In a surprising move, Alibaba has priced Qwen3.8-Max at $2 per million input tokens and $6 per million output tokens, undercutting OpenAI's GPT-5.6 Sol and Anthropic's Claude Opus 5.

This pricing strategy is seen as a challenge to the dominance of OpenAI and Anthropic, with analysts predicting a price war in the AI market. Alibaba's move is also seen as a subsidy for lower-layer companies, as the cheap tokens unlocked by Qwen3.8-Max will be metered in compute, benefiting companies that rely on AI services.

Qwen3.8-Max is built on a sparse mixture-of-experts design, which routes each query through 95 billion active parameters instead of the full network. This approach keeps serving costs and latency down while allowing the model to handle text, images, and video across a context window of one million tokens.

The release of Qwen3.8-Max marks a significant milestone in the development of generative AI. As the market continues to evolve, it will be interesting to see how other companies respond to Alibaba's pricing strategy.

Sources

WireByte Staff — Editorial Team

The WireByte editorial team synthesises technology news from multiple primary sources, verifies the facts, and links every source. Articles are produced with AI assistance and reviewed under our editorial policy.