Alibaba Unveils Qwen3.8-Max: 2.4 Trillion Parameters, Claims Parity With Western Frontier Models at One-Fifth Cost

Alibaba released Qwen3.8-Max, a 2.4-trillion-parameter MoE model claiming parity with OpenAI and Anthropic flagships at one-fifth the cost. Priced at $2/M input and $6/M output with a 1M-token context, open weights arrive August 10.

Monday August 3, 2026 Source: fortune.com
TL;DR — Quick Answer

Alibaba released Qwen3.8-Max, its most capable model yet, a 2.4-trillion-parameter Mixture-of-Experts model activating up to 95B parameters per query. Priced at $2/M input and $6/M output tokens, about one-fifth the cost of comparable Western models, it supports a 1M-token context window. Open weights for Qwen3.8-Max and a smaller Qwen3.8-27B are due August 10.

Key Takeaways

Alibaba Unveils Qwen3.8-Max: 2.4 Trillion Parameters, Claims Parity With Western Frontier Models at One-Fifth Cost — AI news article illustration

The Model

Qwen3.8-Max specs:

The claim of parity with OpenAI and Anthropic’s flagship models at one-fifth the cost is the most aggressive value proposition yet from a Chinese lab.

The 27B Variant

The smaller Qwen3.8-27B is significant for a different reason:

This model targets the local AI market — developers who want open-weight models running on their own hardware.

Market Context

Qwen3.8-Max arrives as Chinese labs dominate global token usage:

Alibaba’s strategy combines frontier-scale capability with aggressive pricing and open-weight availability — a combination Western labs have struggled to match.

Implications

For the AI industry, Qwen3.8-Max represents the clearest signal yet that the gap between Chinese and Western frontier models is closing — while the price gap remains enormous.

Frequently Asked Questions

What is Qwen3.8-Max?

Alibaba's most capable model, a 2.4-trillion-parameter Mixture-of-Experts model activating up to 95B parameters per query.

How much does Qwen3.8-Max cost to use?

$2 per million input tokens and $6 per million output tokens, about one-fifth the cost of comparable Western models.

When do the open weights arrive?

Open weights for Qwen3.8-Max and the smaller Qwen3.8-27B are due August 10.

Can Qwen3.8-27B run locally?

Yes, it needs a 17GB RAM footprint and runs on a single RTX 4090, a Mac with 24GB unified memory, or combined RAM plus VRAM.

This article is based on the official announcement from fortune.com . Read the original for full technical details.

Related Articles

Back to all news