2024年12月推理普及
#1 o1时代分 99 · Elo≈1340
上新定价
Monthly public archive

2024-12 · Reasoning goes mainstream

Launch: o1 正式版 / o1-pro

Reasoning goes mainstream · Launch: o1 正式版 / o1-pro · Scores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Score methodScores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Month landscape board

#ModelVendorEra scoreElo≈PricingHighlight
🥇o1NEWOpenAI991340正式版
🥈Claude 3.5 Sonnet (new)Anthropic961285Computer use
🥉Gemini 2.0 Flash 叙事Google901260速度
4DeepSeek-V3 等DeepSeek881240极低开源性价比

Elo≈ is a LMSYS-style discussion-scale estimate for recap — not an official historical snapshot.

Launches this month

  • o1 正式版 / o1-proOpenAIapi

    推理定价分层

Pricing changes

  • o1 · 正式版推理价 · approx. $15/$60 /1M

    含推理 token

Benchmark context

Benchmark context: public discussion of Arena Elo, AA Intelligence, and open suites (GPQA / SWE-bench, etc.). This page is a narrative archive, not a live probe.

Some historical narratives remain bilingual; model names stay in original form.

2024-12 LLM public board · Reasoning goes mainstream (launches / pric… · GrokCode 倍率榜