趋近智
排名
#5
| 基准 | 分数 | 排名 |
|---|---|---|
数据分析 high | 0.82 | 🥇 1 |
数据分析 max | 0.82 | 🥇 1 |
通用 high | 0.80 | 🥉 3 |
数学 high | 0.96 | 🥉 3 |
LiveBench 平均分 high | 0.80 | 🥉 3 |
LiveBench 平均分 max | 0.80 | 🥉 3 |
编程 high | 0.82 | ⭐ 5 |
编程 max | 0.82 | ⭐ 5 |
推理 high | 0.90 | 5 |
推理 max | 0.90 | 5 |
Agent Arena high | 0.09 | 9 |
通用文本 high | 1482 | ⭐ 11 |
通用文本 标准 | 1476 | ⭐ 18 |
智能编程 high | 0.54 | 12 |
智能编程 max | 0.54 | 12 |
Web 开发 max | 1508 | 28 |
Web 开发 标准 | 1508 | 28 |
Web 开发 high | 1507 | 30 |
排名
#5
编程排名
#23
GPT-5.5 是 OpenAI 针对智能体任务(agentic work)推出的性能最强、效率最高的前沿模型,在 Terminal-Bench 2.0 (82.7%)、SWE-Bench Pro (58.6%)、OSWorld-Verified (78.7%)、BrowseComp (84.4%)、ARC-AGI-2 (85.0%) 和 FrontierMath Tier 4 (35.4%) 上均取得了行业领先(state-of-the-art)的性能表现。在生产环境中,其单 token 延迟与 GPT-5.4 持平,但 token 消耗显著减少。该模型支持文本、图像及工具输入,并具备 1M 上下文窗口。GPT-5.5 于 2026 年 4 月 23 日在 ChatGPT 和 Codex 中发布;API 以 gpt-5.5 的名称提供,定价为每百万输入 token 5 美元,每百万输出 token 30 美元。
专有模型的详细架构规格有限。
GPT-5.5 is OpenAI's smartest and most intuitive model yet, and the next step toward a new way of getting work done on a computer. It excels at agentic coding, computer use, knowledge work, and scientific research — areas where progress depends on reasoning across context and taking action over time. It matches GPT-5.4 per-token latency in real-world serving while performing at a significantly higher level of intelligence, and uses fewer tokens to complete the same tasks. State-of-the-art on Terminal-Bench 2.0 (82.7%), SWE-Bench Pro (58.6%), OSWorld-Verified (78.7%), and ARC-AGI-2 (85.0%). Features a 1M context window and costs $5/M input tokens and $30/M output tokens. Released April 23, 2026.
APX AI
在线