Claude 3.5 Sonnet
LatestAnthropic
•
Proprietary# 55
Released
Oct 22, 2024
# 10
Knowledge Cutoff
Apr 24
# 6
Context Length
200K
Benchmarks
# 51
Code RankedAGI
46.6%
# 20
Aider Polyglot
51.6%
# 17
SWEBench Verified
49.0%
# 8
WebDev Arena
1245.4
# 17
LiveCodeBench v6
36.4%
# 28
LiveCodeBench v5
39.8%
# 23
Codeforces ELO
717
# 18
Code LMArena
1313
# 10
Code LiveBench (old)
67.1%
# 34
GPQA Diamond
65.0%
# 20
Reason LiveBench (old)
58.7%
# 24
ELO LMArena
1355
# 43
AIME 2025 I & II
3.0%
# 27
Math LiveBench (old)
51.3%
# 17
MATH
78.3%
# 23
MATH 500
78.3%
# 23
Humanity Last Exam
4.8%
# 1
Human Eval
93.7%
# 6
Human Eval+
86.2%
# 19
NYT Connections
17.7%
# 10
MMLU Pro
78.0%
# 12
MMLU
88.0%
# 18
MMMU
70.4%
# 26
Halluc. Hughes
4.6%
# 3
Aidan Bench
2691
# 41
AIME 2024
16.0%
# 31
IF LiveBench (old)
69.3%
# 20
Avg LiveBench (old)
60.7%
# 6
IF Evaluation
89.3%
Pricing
# 28
Input Cost /M
$3
# 33
Output Cost /M
$15
# 11
Cached Cost /M
$0.3