Alibaba debuts Qwen3.8-Max model with 2.4T parameters
today debuted a new addition to its Qwen series of open-source large language models. 8-Max is the Chinese e-commerce giant's most capable LLM to date. 4 trillion parameters, about seven times more than the Qwen3.
Key Takeaways
- SiliconANGLE UPDATED 16:14 EDT / AUGUST 03 2026 AI Alibaba debuts Qwen3.
- 6, the previous two entries in the model lineup, shared certain technical properties including a so-called Gated DeltaNet attention mechanism.
- Gated DeltaNet is not the only implementation of attention that offers linear scaling.
What sets it apart is the use of two data processing techniques known as the gated update rule and the delta rule.
- 8-Max outperformed more than a dozen other frontier models including Meta Platforms Inc.
The model is available via Alibaba's cloud platform on launch.
- com/aws-marketplace/ About SiliconANGLE Media SiliconANGLE Media is a recognized leader in digital media innovation, uniting breakthrough technology, strategic insights and real-time audience engagement.
Stats & Key Facts
- #The LLM activates 95 billion of its parameters [...
- #The LLM activates 95 billion of its parameters when answering a query.
- #8-Max supports prompts with up to 1 million tokens worth of data.
- #According to Alibaba, that enables the model to analyze more than 200 pages of text or about 100 hours of footage per request.

SiliconANGLE UPDATED 16:14 EDT / AUGUST 03 2026 AI Alibaba debuts Qwen3. 4T parameters by Maria Deutscher Alibaba Group Holding Ltd. today debuted a new addition to its Qwen series of open-source large language models.
8-Max is the Chinese e-commerce giant's most capable LLM to date. 4 trillion parameters, about seven times more than the Qwen3. 5 model that Alibaba released in February.
The LLM activates 95 billion of its parameters when answering a query. 8-Max supports prompts with up to 1 million tokens worth of data. According to Alibaba, that enables the model to analyze more than 200 pages of text or about 100 hours of footage per request.
For more details please read the original article at SiliconANGLE AI.
Continue Learning
Comments
Sign in to join the conversation