Parameters
2.4T (Estimated)
Context Length
1M
Modality
Multimodal
Architecture
Undisclosed
License
Proprietary
Release Date
23 Sept 2026
Knowledge Cutoff
-
API Pricing (per 1M)
Input: $4.00 · Output: $12.00
No evaluation benchmarks for Qwen 3.8 Max Prime available.
Overall Rank
-
Coding Rank
-
Qwen 3.8 Max Prime is a high-throughput enterprise tier of Alibaba's Qwen3.8 Max flagship model. It supports text, image, and video input with a 1M-token context window for demanding production workflows.
Architecture specifications are undisclosed for proprietary models.
Alibaba's Qwen 3.8 generation represents the frontier hybrid Mixture-of-Experts architecture designed for coding, professional work, research, and long-horizon agentic tasks. It features a 2.4-trillion parameter architecture (95B active per token) combining Gated DeltaNet linear attention with standard Gated Attention, available both as open weights and as a hosted flagship service.
Assistant
Online