2024年11月o1 推理模型
#1 o1-preview时代分 99 · Elo≈1330
Monthly public archive

2024-11 · o1 reasoning models

o1 reasoning models · landscape continues

o1 reasoning models · Scores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Score methodScores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Month landscape board

#ModelVendorEra scoreElo≈PricingHighlight
🥇o1-previewOpenAI991330推理溢价数学/竞赛跃迁
🥈Claude 3.5 SonnetAnthropic951280日常编码仍强
🥉GPT-4oOpenAI931270默认体验
4o1-miniOpenAI901290中低便宜推理

Elo≈ is a LMSYS-style discussion-scale estimate for recap — not an official historical snapshot.

Launches this month

No separately tagged blockbuster launches this month (landscape continues).

Pricing changes

No separately tagged pricing events (prior month scale carries over).

Benchmark context

Benchmark context: public discussion of Arena Elo, AA Intelligence, and open suites (GPQA / SWE-bench, etc.). This page is a narrative archive, not a live probe.

Some historical narratives remain bilingual; model names stay in original form.

2024-11 LLM public board · o1 reasoning models (launches / pricing / … · GrokCode 倍率榜