AI RACE— 每日追蹤 AI 競爭賽局
AI 模型

GPT-5.1 Codex

OpenAI

立即試用 ↗
發布日期2025/11
推論速度0 tokens/s
首字延遲時間0.00s
輸入價格$1.25/1M
輸出價格$10/1M

排名表現

#106
AIM 40.7最高排名 #54
#102
AIM 83.3最高排名 #58
#8
AIM 90.6最高排名 #8
#89
AIM 79.7最高排名 #70

評測分析

GPT-5.1 Codex is a specialized model for software engineers, optimized for development workflows rather than conversational AI or general reasoning. It suits dev teams needing a powerful coding copilot, but should not be used for chatbots, customer support, or general reasoning tasks.

強項

  • Specialized coding and code generation: the 'Codex' name reflects a distinct strength in processing code and solving coding problems
  • Solid mathematics (ranked #8): handles formulas, precise calculations, and algorithms
  • Strong debugging and code refactoring capabilities, with a deep understanding of programming logic context

弱項

  • Limited general reasoning (ranked #67): weak in complex multi-step reasoning problems outside the programming domain
  • Weak conversational interaction (ranked #63): not a good choice for chatbots, customer support, or concept explanations
  • Low AI Agent capabilities (ranked #71): limited in automating complex tasks and workflow orchestration

適用情境

Programming assistance: code generation, bug fixing, code review, algorithm optimizationSolving coding challenges: LeetCode, technical interview preparation, algorithmic calculationsCode refactoring and migration: translating code between languages, cleaning up legacy code

指南與影片

Access Codex via a ChatGPT account (Free or Go plan) or the OpenAI API. The model is used for software development, code review, PR generation, and large project debugging. Use it via CLI, IDE extensions, ChatGPT web, or GitHub bots. Optimized for multi-file tasks and long sessions thanks to an automatic compaction feature that refreshes the context window.

相關評測