Alibaba
Qwen3.8-Max
Alibaba's hosted Qwen3.8-Max accepts text, images and video within a 1M-token context window. QwenCloud lists $2/$6 per MTok (input/output). Alibaba Cloud added the dated qwen3.8-max-0902 snapshot on September 2, 2026, also named qwen3.8-max-2026-09-02. The related Qwen3.8-2.4T-A95B weights are a text-only post-trained model; their license and local context limits should not be conflated with the hosted service.
1M tokens
Multimodal
Max
New
- Provider
- Alibaba
- Context
- 1MCatalog · As of Aug 13, 2026Stale · 35d old
- yno subscription tier
- Max
- Released
- Aug 3, 2026
- Speed
- Medium
- Reasoning
- Expert
- Modality
- Multimodal
Benchmarks
Quality
- AA Intelligence Index
- AA Coding Index
- GPQA Diamond
- 92.7%Artificial Analysis · Date unknownAge unknown
Speed and latency
- Output speed
- 40.74tokens/sConfiguration: Prompt length: 1,000 · Parallel queries: 1Artificial Analysis · Date unknownAge unknown
- Time to first token
- 1.79sConfiguration: Prompt length: 1,000 · Parallel queries: 1Artificial Analysis · Date unknownAge unknown
- Time to first answer token
- 50.89sConfiguration: Prompt length: 1,000 · Parallel queries: 1Artificial Analysis · Date unknownAge unknown
API pricing
- API input cost
- API output cost
Benchmark variant: Qwen3.8 Max
AA data retrieved Sep 17, 2026 · Artificial Analysis
Indexes use points; evaluations use accuracy percentages. API measurements do not measure yno application performance. Methodology
Capabilities
- Reasoning
- Code
- Vision
- Video
- Agentic
- Long Context
Use cases
- Long-horizon coding agents
- Multimodal analysis
- Repo-wide refactors
- Document/video understanding
Strengths
- 2.4T MoE / 95B active
- Text, image, and video input
- 1M context
- Related text-only weights available
Best for
Teams building hosted coding agents and document or video analysis workflows
Use Qwen3.8-Max in yno.ai
No credit card required
Related models
Frequently asked
- What is Qwen3.8-Max?
- Alibaba's hosted Qwen3.8-Max accepts text, images and video within a 1M-token context window. QwenCloud lists $2/$6 per MTok (input/output). Alibaba Cloud added the dated qwen3.8-max-0902 snapshot on September 2, 2026, also named qwen3.8-max-2026-09-02. The related Qwen3.8-2.4T-A95B weights are a text-only post-trained model; their license and local context limits should not be conflated with the hosted service.
- How much does Qwen3.8-Max cost in yno.ai?
- Qwen3.8-Max is available on the Max tier of yno.ai.
- What can Qwen3.8-Max do?
- Qwen3.8-Max is best for Teams building hosted coding agents and document or video analysis workflows. Its main capabilities include Reasoning, Code, Vision, Video, Agentic, Long Context.
- How does Qwen3.8-Max compare to other models?
- Qwen3.8-Max excels at 2.4T MoE / 95B active and is recommended for Long-horizon coding agents, Multimodal analysis, Repo-wide refactors, Document/video understanding. See related models below.
- How do I use Qwen3.8-Max in yno.ai?
- Sign up for yno.ai, select Qwen3.8-Max from the model picker, and start chatting.