ApX logoApX logo

Qwen3.7 Plus

Parameters

Undisclosed

Context Length

1M

Modality

Multimodal

Architecture

Dense

License

Proprietary

Release Date

1 Jun 2026

Knowledge Cutoff

-

API Pricing (per 1M)

Input: $0.40 · Output: $1.60

Evaluation Benchmarks

Rank

#50

BenchmarkScoreRank

Professional Knowledge

MMLU Pro

0.885

🥉

3

Graduate-Level QA

GPQA

0.903

24

Agent Arena

Agent Arena

-5.38

41

Web Development

WebDev Arena

1460

50

General Text

Text Arena

1456

50

0.56

78

Agentic Index

Artificial Analysis

0.20

79

Intelligence Index

Artificial Analysis

0.26

96

Rankings

Overall Rank

#50

Coding Rank

#71

About Qwen3.7 Plus

Alibaba's multimodal agent model released June 1, 2026, unifying vision and language into a single versatile agent foundation built on the Qwen3.7 text backbone. Qwen3.7-Plus delivers comprehensive upgrades in vision-language understanding, enabling the model to reason over images, screenshots, documents, and visual inputs while retaining Qwen3.7-Max's powerful long-horizon agentic execution capabilities.

Technical Specifications

Architecture specifications are undisclosed for proprietary models.

Attention

Attention Structure

Multi-Head Attention

Attention Heads

Undisclosed

Key-Value Heads

Undisclosed

Attention Head Dimension

Undisclosed

Position Embedding

Absolute Position Embedding

RoPE Theta

Undisclosed

Sliding Window Attention

Undisclosed

Sliding Window Size

Undisclosed

Sliding Window Ratio

Undisclosed

Linear Attention

Undisclosed

Linear Attention Ratio

Undisclosed

Normalization

Undisclosed

Activation Function

Undisclosed

Dimensions

Hidden Dimension Size

Undisclosed

Number of Layers

Undisclosed

FFN Intermediate Size (Dense)

Undisclosed

Multi-Token Prediction Heads

Undisclosed

Tokenizer

Vocabulary Size

Undisclosed

About Qwen 3.7

Alibaba's Qwen 3.7 generation is designed for the agent era, delivering frontier-level agentic reasoning and long-horizon autonomous execution. Qwen3.7 models combine deep coding agent capabilities with broad cross-scaffold generalization, sustaining productive execution over multi-hour sessions with thousands of tool calls. The family includes both text-focused and full multimodal variants.


Other Qwen 3.7 Models