Qwen3.8-Max
Alibaba's top model. Apsara 2026 (2026-09-22): Alibaba says an updated Qwen3.8-Max went through 33 automated self-improvement cycles, raising its Artificial Analysis score from 40 to 45 (company claim, https://www.alibabacloud.com/en/press-room/alibaba-unveils-roadmap-on-full-stack-ai-strategy). Snapshot qwen3.8-max-0902; fast tier qwen3.8-max-prime (OpenRouter qwen/qwen3.8-max-prime, Beijing 3.301/9.902). Singapore endpoint needs your WorkspaceId (old dashscope-intl domain is being migrated). Also sold via Qwen Cloud (qwencloud.com). Release day not verified (weights on HF 2026-08-08). Knowledge cutoff not published.
- Context window
- 1,000,000 tokens
- Max output
- 131,072 tokens
- Input
- text, image, video
- Output
- text
- License
- qwen3.8-max (custom, for open weights Qwen3.8-2.4T-A95B)
- Pricing
- input: $2 · output: $6 · cache read: $0.25 (per 1M tokens (USD), Singapore/International region, input up to 1M; Beijing/Global regions 1.65/4.951) source
- Verified
- 2026-09-29
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| Alibaba Cloud Model Studio (DashScope, Singapore/Intl) | qwen3.8-max | https://{WorkspaceId}.ap-southeast-1.maas.aliyuncs.com/compatible-mode/v1 | docs |
| Alibaba Cloud Model Studio (US Virginia) | qwen3.8-max | https://dashscope-us.aliyuncs.com/compatible-mode/v1 | docs |
| OpenRouter | qwen/qwen3.8-max-0902 | openrouter.ai/qwen/qwen3.8-max-0902 | — |
| OpenRouter (open-weight base) | qwen/qwen3.8-2.4t-a95b | openrouter.ai/qwen/qwen3.8-2.4t-a95b | — |
| Hugging Face | — | huggingface.co/Qwen/Qwen3.8-2.4T-A95B | — |
| Web app | — | chat.qwen.ai | — |
Notable capabilities (4)
- First open-weight Qwen-Max-class model: Qwen3.8 brings a Max-class model to open release for the first time (Qwen3.8-2.4T-A95B, 2.4T total / 95B active MoE); the API version adds vision input, non-thinking mode, 1M context and built-in tools. source
- Multi-day autonomous coding: Alibaba markets it as able to code autonomously for over ten days to deliver complete projects, with closed-loop planning and iteration. source
- Native vision in the agent loop: Image and video understanding used throughout planning, execution and verification; parses ultra-long documents and long videos. source
- Tunable and preserved thinking: reasoning_effort controls depth; preserve_thinking keeps reasoning context from earlier turns. source
Alibaba's flagship for hard reasoning, long-horizon coding and professional work (law, finance, design); 1M context, text/image/video in.
curl "https://{WorkspaceId}.ap-southeast-1.maas.aliyuncs.com/compatible-mode/v1/chat/completions" \
-H "Authorization: Bearer $DASHSCOPE_API_KEY" -H "Content-Type: application/json" \
-d '{"model":"qwen3.8-max","messages":[{"role":"user","content":"Hello"}],"enable_thinking":true}'
Sources: https://www.alibabacloud.com/help/en/model-studio/qwen3-8-max · https://www.alibabacloud.com/help/en/model-studio/model-pricing · https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B
Other Alibaba (Qwen) models
Qwen-Audio-3.1-ASR (Flash) · Qwen-Audio-3.1-Realtime (Plus) · Qwen-Audio-3.1-TTS-Next · Qwen3.8-LiveTranslate (Flash Realtime) · Qwen3.8-Omni-Flash · Qwen3.8-27B · Qwen3.8-Flash · Qwen-Image-3.0 (Pro) · Qwen3.7-Plus · Qwen3-ASR (0.6B / 1.7B) + Qwen3-ForcedAligner · Qwen3-TTS (open weights 0.6B / 1.7B; API qwen3-tts-flash)