ApX logoApX logo

Hunyuan Large

Active Parameters

389B

Context Length

28K

Modality

Text

Architecture

Mixture of Experts (MoE)

License

Tencent Hunyuan Community License

Release Date

5 Nov 2024

Knowledge Cutoff

Sep 2024

API Pricing (per 1M)

Self-hosted only

System Requirements

VRAM requirements for different quantization methods and context sizes

1,024 tokens

820.51 GB VRAM

Consumer

46x RTX 4090

24GB VRAM

Datacenter

12x NVIDIA A100

80GB VRAM

Apple Silicon

10x Apple M3 Max

128GB VRAM

28,000 tokens

876.20 GB VRAM

Consumer

50x RTX 4090

24GB VRAM

Datacenter

13x NVIDIA A100

80GB VRAM

Apple Silicon

10x Apple M3 Max

128GB VRAM

Architecture Diagram

Input TokensToken EmbeddingPosition: AbsoluteHidden: 4.1k · Context: 28Kx 60 layersLayerNormPre-AttentionMulti-Head Attention64Q / 64KV headsHead dim: 64+LayerNormPre-FFNSparse MoE FFN (2/32 experts)GELU+Final LayerNormOutput Logits

Evaluation Benchmarks

Rank

#127

BenchmarkScoreRank

General Text

Text Arena

1326

133

Rankings

Overall Rank

#127

Coding Rank

-

About Hunyuan Large

Hunyuan-DiT is Tencent's large-scale Mixture-of-Experts diffusion transformer engineered for high-fidelity text-to-image synthesis up to 4096x4096 resolution. It features bilingual CLIP and T5 encoders for fine-grained prompt comprehension and multi-turn editing.

Technical Specifications

Attention

Attention Structure

Multi-Head Attention

Attention Heads

64

Key-Value Heads

64

Attention Head Dimension

-

Position Embedding

Absolute Position Embedding

RoPE Theta

-

Sliding Window Attention

-

Sliding Window Size

-

Sliding Window Ratio

-

Linear Attention

-

Linear Attention Ratio

-

Normalization

Layer Normalization

Activation Function

GELU

Dimensions

Hidden Dimension Size

4,096

Number of Layers

60

FFN Intermediate Size (Dense)

-

Multi-Token Prediction Heads

-

Tokenizer

Vocabulary Size

-

Mixture of Experts

Total Expert Parameters

52.0B

Number of Experts

32

Active Experts

2

Shared Experts

-

FFN Intermediate Size (per Expert)

-

Dense Layers Before MoE

-

About Hunyuan

Tencent Hunyuan large language models with various capabilities.


Other Hunyuan Models
Hunyuan Large: Specifications and GPU VRAM Requirements