Alibaba's largest model
2.4 trillion parameters, previewed at the World AI Conference in Shanghai and released on August 3, 2026.
Qwen3.8 Max is Alibaba's biggest model yet: 2.4 trillion parameters in a mixture of experts, text, images and video in, a million tokens of context. It launched on August 3, 2026, and the September 2 snapshot sharpened coding on engineering-scale projects and long autonomous agent runs. On avots.ai you use Qwen3.8 Max and the budget Qwen3.8 Flash pay per use, no plan, and the same balance also covers GPT-6 Astra, Claude Fable 5.1, Gemini and 40+ image, video and music models.
| Qwen3.8 Max | Qwen3.8 Flash | Qwen3.7 Max | |
|---|---|---|---|
| Best for | big coding projects, long agents, vision | volume work, documents, video analysis | the previous flagship |
| Price per 1M in / out | $2 / $6 | $0.15 / $0.47 | $2 / $6 |
| Context window | 1M tokens | 1M tokens | 1M tokens |
| Max output | 131k tokens | 131k tokens | 65k tokens |
| Input | text, images, video | text, images, video | text |
| On avots.ai | pay per use | pay per use | pay per use |
The number is the generation, the name is the tier: Max is the 2.4 trillion parameter flagship, Flash the fast multimodal tier at a thirteenth of the price. Both read video. Switch freely, one balance covers the whole Qwen family.
2.4 trillion parameters, previewed at the World AI Conference in Shanghai and released on August 3, 2026.
Max and Flash read images and video natively: describe a clip, pull facts from a recording, read charts and long documents.
The September 2 snapshot targets engineering-scale software projects and longer autonomous development runs, with stronger vision.
A whole codebase or a long video transcript in one conversation, and answers twice as long as Qwen3.7 allowed.
Sign up on app.avots.ai with email, Google or Telegram. Takes a minute.
1,000 tokens cost 0.99 EUR. No plan to choose, no renewal, the balance sits there until you use it.
Choose a Qwen3.8 tier in the model picker, or let Agent Avots route the task for you. Every reply shows its token cost.
app.avots.ai on desktop and mobile, no install needed.
Qwen3.8 in @AvotsAIbot, same balance, from any phone.
Use it inside Claude Desktop, Cursor or Cline via the MCP server.
The OpenAI compatible API takes your existing client, one key for every model.
1,000 tokens cost 0.99 EUR and a Qwen3.8 Flash reply costs a few of them. The same balance unlocks Max, GPT-6 Astra, Claude Fable 5.1, Gemini, Veo video and Nano Banana images.
Sign up and chat ✦Every LLM on avots.ai and which one to pick per task.
The previous generation: Max, Plus and Flash, still in the picker.
Open-weight MoE, 1M context, top LiveCodeBench coding scores.
Moonshot's frontier open model, number one at coding, 1M context.
OpenAI's new flagship and its Pro mode, pay per use.
What to use instead of ChatGPT Plus and how one balance replaces it.
avots.ai runs 40+ AI models for chat, image, video and audio on one balance. See the full list on the Tools page or open the app and start creating.