Active Parameters
770B
Context Length
1.05M
Modality
Text
Architecture
Mixture of Experts (MoE)
License
Apache 2.0
Release Date
27 Aug 2026
Knowledge Cutoff
-
VRAM requirements for different quantization methods and context sizes
1,024 tokens
Consumer
102x RTX 4090
24GB VRAM
Datacenter
26x NVIDIA A100
80GB VRAM
Apple Silicon
22x Apple M3 Max
128GB VRAM
1,048,576 tokens
Consumer
115x RTX 4090
24GB VRAM
Datacenter
29x NVIDIA A100
80GB VRAM
Apple Silicon
25x Apple M3 Max
128GB VRAM
Rank
#5
| Benchmark | Score | Rank |
|---|---|---|
Web Development | 1627 | ⭐ 7 |
Overall Rank
#5
Coding Rank
#9
Hunyuan 4 Preview is Tencent's next-generation foundation model preview showcasing advanced Chinese and English language understanding. It features enhanced architectural scalability for enterprise reasoning and text generation tasks.
Attention
Attention Structure
DeepSeek Sparse Attention
Attention Heads
64
Key-Value Heads
8
Attention Head Dimension
64
Position Embedding
ROPE
RoPE Theta
10,000,000
Sliding Window Attention
No
Sliding Window Size
-
Sliding Window Ratio
-
Linear Attention
No
Linear Attention Ratio
-
Normalization
RMS Normalization
Activation Function
SwigLU
Dimensions
Hidden Dimension Size
6,144
Number of Layers
78
FFN Intermediate Size (Dense)
18,432
Multi-Token Prediction Heads
1
Tokenizer
Vocabulary Size
120,832
Mixture of Experts
Total Expert Parameters
49.0B
Number of Experts
256
Active Experts
8
Shared Experts
1
FFN Intermediate Size (per Expert)
2,048
Dense Layers Before MoE
1
Tencent Hunyuan large language models with various capabilities.
APX AI
Online