趋近智
参数
未公开
上下文长度
272K
模态
Multimodal
架构
未公开
许可证
Proprietary
发布日期
5 Mar 2026
训练数据截止日期
-
API 价格 (每 1M)
输入: $2.50 · 输出: $15.00
排名
#30
| 基准 | 分数 | 排名 |
|---|---|---|
研究生级问答 | 0.928 | 10 |
max 0.79 | 12 | |
max 0.94 | 16 | |
max 0.78 | 17 | |
LiveBench 平均分 | max 0.78 | 17 |
max 0.88 | 21 | |
max 0.54 | 30 | |
智能体指数 | max 0.44 | 30 |
通用文本 | high 1475 标准 1465 | 31 42 |
max 0.78 | 32 | |
Agent Arena | high 0.31 | 33 |
max 0.71 | 44 | |
Web 开发 | high 1465 medium 1443 标准 1395 | 56 61 76 |
max 0.39 low 0.28 标准 0.18 | 56 109 170 |
排名
#30
编程排名
#46
GPT-5.4 是 OpenAI 为专业工作打造的最强大且最高效的前沿模型。它将推理、编程和智能体工作流(agentic workflows)方面的最新进展集成于单一模型之中。该模型具备源自 GPT-5.3-Codex 的行业领先编程能力、原生的顶尖计算机操作能力,以及在大型生态系统中经过改进的工具调用能力。它擅长处理涉及电子表格、演示文稿和文档的专业任务。在基准测试中,它在 GDPval 上达到了 83.0%,OSWorld-Verified 为 75.0%,BrowseComp 为 82.7%,SWE-Bench Pro 为 57.7%,MMMU Pro 为 81.2%。GPT-5.4 支持高达 272K 的上下文长度(实验性支持 1M),并实现了迄今为止 Token 效率最高的推理。
专有模型的架构规格未公开。
GPT-5.4 is OpenAI's most capable and efficient frontier model for professional work, combining the industry-leading coding capabilities of GPT-5.3-Codex with major advances in reasoning, computer use, and agentic workflows. It introduces native computer-use capabilities, tool search for large tool ecosystems, substantially improved knowledge work (spreadsheets, presentations, documents), and is OpenAI's most factual and token-efficient reasoning model. Supports up to 1M context tokens in Codex. Released March 5, 2026.
Assistant
在线