Parameters
Undisclosed
Context Length
1M
Modality
Multimodal
Architecture
Undisclosed
License
Proprietary
Release Date
21 Sept 2026
Knowledge Cutoff
-
API Pricing (per 1M)
Input: $0.15 · Output: $0.47
No evaluation benchmarks for Qwen 3.8 Omni Flash available.
Overall Rank
-
Coding Rank
-
Qwen 3.8 Omni Flash is Alibaba's native omni-modal reasoning model with integrated audio, video, image, and text capabilities. It is optimized for high-speed agentic interaction, real-time multimedia analysis, and long-context summarization.
Architecture specifications are undisclosed for proprietary models.
Alibaba's Qwen 3.8 generation represents the frontier hybrid Mixture-of-Experts architecture designed for coding, professional work, research, and long-horizon agentic tasks. It features a 2.4-trillion parameter architecture (95B active per token) combining Gated DeltaNet linear attention with standard Gated Attention, available both as open weights and as a hosted flagship service.
Assistant
Online