2024年10月o1 推理模型
#1 o1-preview时代分 99 · Elo≈1330
上新定价
Monthly public archive

2024-10 · o1 reasoning models

Launch: Claude 3.5 Sonnet (new)

o1 reasoning models · Launch: Claude 3.5 Sonnet (new) · Scores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Score methodScores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Month landscape board

#ModelVendorEra scoreElo≈PricingHighlight
🥇o1-previewOpenAI991330推理溢价数学/竞赛跃迁
🥈Claude 3.5 SonnetNEWAnthropic951280日常编码仍强
🥉GPT-4oOpenAI931270默认体验
4o1-miniOpenAI901290中低便宜推理

Elo≈ is a LMSYS-style discussion-scale estimate for recap — not an official historical snapshot.

Launches this month

  • Claude 3.5 Sonnet (new)Anthropicapi

    Computer use 等

Pricing changes

  • Claude 3.5 Sonnet · 新版同价档 · approx. $3/$15 /1M

    能力提升、价稳

Benchmark context

Benchmark context: public discussion of Arena Elo, AA Intelligence, and open suites (GPQA / SWE-bench, etc.). This page is a narrative archive, not a live probe.

Some historical narratives remain bilingual; model names stay in original form.

2024-10 LLM public board · o1 reasoning models (launches / pricing / … · GrokCode 倍率榜