ApX 标志ApX 标志

趋近智

GPT-5.4

参数

未公开

上下文长度

272K

模态

Multimodal

架构

未公开

许可证

Proprietary

发布日期

5 Mar 2026

训练数据截止日期

-

API 价格 (每 1M)

输入: $2.50 · 输出: $15.00

评估基准

排名

#30

基准分数排名

研究生级问答

GPQA

0.928

10

max

0.79

12

max

0.94

16

max

0.78

17

LiveBench 平均分

LiveBench Average
max

0.78

17

max

0.88

21

智能编程

LiveBench Agentic
max

0.54

30

智能体指数

Artificial Analysis
max

0.44

30

通用文本

Text Arena
high

1475

标准

1465

31

42

max

0.78

32

Agent Arena

Agent Arena
high

0.31

33

max

0.71

44

Web 开发

WebDev Arena
high

1465

medium

1443

标准

1395

56

61

76

max

0.39

low

0.28

标准

0.18

56

109

170

排名

排名

#30

编程排名

#46

关于 GPT-5.4

GPT-5.4 是 OpenAI 为专业工作打造的最强大且最高效的前沿模型。它将推理、编程和智能体工作流(agentic workflows)方面的最新进展集成于单一模型之中。该模型具备源自 GPT-5.3-Codex 的行业领先编程能力、原生的顶尖计算机操作能力,以及在大型生态系统中经过改进的工具调用能力。它擅长处理涉及电子表格、演示文稿和文档的专业任务。在基准测试中,它在 GDPval 上达到了 83.0%,OSWorld-Verified 为 75.0%,BrowseComp 为 82.7%,SWE-Bench Pro 为 57.7%,MMMU Pro 为 81.2%。GPT-5.4 支持高达 272K 的上下文长度(实验性支持 1M),并实现了迄今为止 Token 效率最高的推理。

技术规格

专有模型的架构规格未公开。

关于 GPT-5.4

GPT-5.4 is OpenAI's most capable and efficient frontier model for professional work, combining the industry-leading coding capabilities of GPT-5.3-Codex with major advances in reasoning, computer use, and agentic workflows. It introduces native computer-use capabilities, tool search for large tool ecosystems, substantially improved knowledge work (spreadsheets, presentations, documents), and is OpenAI's most factual and token-efficient reasoning model. Supports up to 1M context tokens in Codex. Released March 5, 2026.


其他 GPT-5.4 模型