Parameters
Undisclosed
Context Length
400K
Modality
Multimodal
Architecture
Undisclosed
License
Proprietary
Release Date
17 Mar 2026
Knowledge Cutoff
-
API Pricing (per 1M)
Input: $0.75 · Output: $4.50
Rank
#72
| Benchmark | Score | Rank |
|---|---|---|
Graduate-Level QA | 0.88 | 25 |
Data Analysis | max 0.71 | 30 |
Coding | max 0.72 | 35 |
Agentic Coding | max 0.42 | 38 |
General | max 0.66 | 39 |
Reasoning | max 0.71 | 39 |
LiveBench Average | max 0.66 | 39 |
Mathematics | max 0.78 | 40 |
General Text | high 1448 | 54 |
Web Development | high 1397 | 60 |
Coding Index | max 0.56 | 67 |
Agentic Index | max 0.20 | 70 |
Intelligence Index | max 0.25 medium 0.20 Standard 0.11 | 96 127 185 |
Overall Rank
#72
Coding Rank
#77
GPT-5.4 mini is OpenAI's fast, efficient small model that brings many capabilities of GPT-5.4 to high-volume workloads. It significantly improves over GPT-5 mini across coding, reasoning, multimodal understanding, and tool use while running more than 2x faster. Approaches GPT-5.4 performance on SWE-Bench Pro (54.4%) and OSWorld-Verified (72.1%). Supports text and image inputs, tool use, function calling, web search, file search, computer use, and skills. 400K context window. Pricing: $0.75/M input, $4.50/M output. Available via API as gpt-5.4-mini. Released March 17, 2026.
Architecture specifications are undisclosed for proprietary models.
Attention
Attention Structure
Multi-Head Attention
Attention Heads
Undisclosed
Key-Value Heads
Undisclosed
Attention Head Dimension
Undisclosed
Position Embedding
Absolute Position Embedding
RoPE Theta
Undisclosed
Sliding Window Attention
Undisclosed
Sliding Window Size
Undisclosed
Sliding Window Ratio
Undisclosed
Linear Attention
Undisclosed
Linear Attention Ratio
Undisclosed
Normalization
Undisclosed
Activation Function
Undisclosed
Dimensions
Hidden Dimension Size
Undisclosed
Number of Layers
Undisclosed
FFN Intermediate Size (Dense)
Undisclosed
Multi-Token Prediction Heads
Undisclosed
Tokenizer
Vocabulary Size
Undisclosed
GPT-5.4 is OpenAI's most capable and efficient frontier model for professional work, combining the industry-leading coding capabilities of GPT-5.3-Codex with major advances in reasoning, computer use, and agentic workflows. It introduces native computer-use capabilities, tool search for large tool ecosystems, substantially improved knowledge work (spreadsheets, presentations, documents), and is OpenAI's most factual and token-efficient reasoning model. Supports up to 1M context tokens in Codex. Released March 5, 2026.
APX AI
Online