Papers › Claude 3.5 Sonnet Model Card Addendum
Claude 3.5 Sonnet Model Card Addendum
Anthropic
This addendum to our Claude 3 Model Card describes Claude 3.5 Sonnet, a new model which outperforms our previous most capable model, Claude 3 Opus, while operating faster and at a lower cost. Claude 3.5 Sonnet offers improved capabilities, including better coding and visual processing. Since it is an evolution of the Claude 3 model family, we are providing an addendum rather than a new model card. We provide updated key evaluations and results from our safety testing.
Code
No code repository is listed for this paper in the archive or in Syntology's graph.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| MMR total | MRR-Benchmark | Claude 3.5 Sonnet | Total Column Score | 463 | #1 of 14 | Archive leaderboard | report |
| Question Answering | NewsQA | Anthropic/claude-3-7-sonnet | EM | 74.23 | #6 of 18 | Archive leaderboard | report |
| Question Answering | NewsQA | Anthropic/claude-3-7-sonnet | F1 | 82.3 | #6 of 18 | Archive leaderboard | report |
| Visual Question Answering | MM-Vet | Claude 3.5 Sonnet (claude-3-5-sonnet-20240620) | GPT-4 score | 74.2±0.2 | #5 of 231 | Archive leaderboard | report |
| Visual Question Answering | MM-Vet v2 | Claude 3.5 Sonnet (claude-3-5-sonnet-20240620) | GPT-4 score | 71.8±0.2 | #3 of 24 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections