2025年1月R1 冲击波
#1 o1 / o1-pro时代分 98 · Elo≈1345
上新定价开源推理冲击定价
Monthly public archive

2025-01 · R1 shockwave

Launch: DeepSeek-R1

DeepSeek-R1 shocks open-weight reasoning cost curves; public boards and free-pool probes diverge more. Era-relative scoring.

Score methodScores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Month landscape board

#ModelVendorEra scoreElo≈PricingHighlight
🥇o1 / o1-proOpenAI981345闭源推理顶
🥈DeepSeek-R1NEWDeepSeek961320极低/开源开源推理逼近
🥉Claude 3.5 SonnetAnthropic931290
4GPT-4oOpenAI901275

Elo≈ is a LMSYS-style discussion-scale estimate for recap — not an official historical snapshot.

Launches this month

  • DeepSeek-R1DeepSeekopen

    推理开源冲击定价

Pricing changes

  • DeepSeek-R1 · 开源冲击 · approx. $0.55/$2.19 /1M

    API/开源双轨压价

  • DeepSeek-R1 · 极低 API/开源 · approx. $0.55/$2.19 /1M

    冲击全球定价预期

Benchmark context

Benchmark context: public discussion of Arena Elo, AA Intelligence, and open suites (GPQA / SWE-bench, etc.). This page is a narrative archive, not a live probe.

Some historical narratives remain bilingual; model names stay in original form.

2025-01 LLM public board · R1 shockwave (launches / pricing / scores) · GrokCode 倍率榜