Claude 3.5 Sonnet
LatestAnthropic
•
Proprietary# 111
Released
Oct 22, 2024
# 16
Knowledge Cutoff
Apr 24
# 9
Context Length
200K
Benchmarks
# 132
Code RankedAGI
41.2%
# 48
SWEBench Verified
49.0%
# 153
Agentic RankedAGI
32.8%
# 33
LiveCodeBench v6
36.4%
# 27
LiveCodeBench v5
39.8%
# 12
Code LMArena
1313
# 27
Codeforces ELO
717
# 19
Aider Polyglot
51.6%
# 9
Code LiveBench (old)
67.1%
# 162
Reason RankedAGI
36.4%
# 67
HLE
4.8%
# 69
GPQA Diamond
65.0%
# 55
Text Arena
1355
# 55
AIME 2025 I & II
3.0%
# 37
AIME 2024
16.0%
# 1
Human Eval
93.7%
# 6
Human Eval+
86.2%
# 26
NYT Connections
17.7%
# 24
MMLU Pro
78.0%
# 14
MMLU
88.0%
# 21
MMMU
70.4%
# 27
Halluc. Hughes
4.6%
# 3
Aidan Bench
2691
# 17
Avg LiveBench (old)
60.7%
# 6
IF Evaluation
89.3%
# 27
Coding LiveBench 25.4
32.3%
# 28
Data LiveBench
52.8%
# 8
Language LiveBench
53.8%
# 5
Quality Artificial Analysis
80
# 199
Math RankedAGI
37.2%
# 148
RAGI RankedAGI
41.1%
Pricing
# 39
Input Cost /M
$3
# 46
Output Cost /M
$15
# 26
Cached Cost /M
$0.3