2025年2月R1 冲击波
#1 o1 / o1-pro时代分 98 · Elo≈1345
上新定价
Monthly public archive

2025-02 · R1 shockwave

Launch: Claude 3.7 Sonnet

R1 shockwave · Launch: Claude 3.7 Sonnet · Scores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Score methodScores are era-relative (month top ≈95–100) for landscape review — not official historical Artificial Analysis replays. From 2026-08, the latest month aligns with the live AA-style snapshot.

Month landscape board

#ModelVendorEra scoreElo≈PricingHighlight
🥇o1 / o1-proOpenAI981345闭源推理顶
🥈DeepSeek-R1DeepSeek961320极低/开源开源推理逼近
🥉Claude 3.5 SonnetNEWAnthropic931290
4GPT-4oOpenAI901275

Elo≈ is a LMSYS-style discussion-scale estimate for recap — not an official historical snapshot.

Launches this month

  • Claude 3.7 SonnetAnthropicapi

    混合推理叙事

Pricing changes

  • o3-mini · 推理下探 · approx. $1.1/$4.4 /1M

    量级示意

Benchmark context

Benchmark context: public discussion of Arena Elo, AA Intelligence, and open suites (GPQA / SWE-bench, etc.). This page is a narrative archive, not a live probe.

Some historical narratives remain bilingual; model names stay in original form.

2025-02 LLM public board · R1 shockwave (launches / pricing / scores) · GrokCode 倍率榜