2024年9月o1 推理模型
#1 o1-preview时代分 99 · Elo≈1330
上新定价推理模型范式
Monthly public archive

2024-09 · o1 reasoning models

Launch: o1-preview / o1-mini

o1-preview / o1-mini open the reasoning-model era—test-time compute becomes a product axis. Scores are era-relative.

Score methodScores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Month landscape board

#ModelVendorEra scoreElo≈PricingHighlight
🥇o1-previewNEWOpenAI991330推理溢价数学/竞赛跃迁
🥈Claude 3.5 SonnetAnthropic951280日常编码仍强
🥉GPT-4oOpenAI931270默认体验
4o1-miniNEWOpenAI901290中低便宜推理

Elo≈ is a LMSYS-style discussion-scale estimate for recap — not an official historical snapshot.

Era long-read for this month: Reasoning models rise: o1 to DeepSeek-R1

Launches this month

  • o1-preview / o1-miniOpenAIapi

    推理模型时代

Pricing changes

  • o1-preview · 推理溢价 · approx. $15/$60 /1M

    按推理 token 计费叙事

Benchmark context

Benchmark context: public discussion of Arena Elo, AA Intelligence, and open suites (GPQA / SWE-bench, etc.). This page is a narrative archive, not a live probe.

Some historical narratives remain bilingual; model names stay in original form.

2024-09 LLM public board · o1 reasoning models (launches / pricing / … · GrokCode 倍率榜