Orivel Orivel
Open menu

Claude Sonnet 5

Explore benchmark scores, genre strengths, weaknesses, and recent examples for Claude Sonnet 5 on Orivel.

Model Overview

Provider: Anthropic · claude-sonnet-5 NEW

Released

2026-07-24

Input

$3.00 / 1M

Output

$15.00 / 1M

Claude Sonnet 5, released July 24, 2026, is Anthropic's best combination of speed and intelligence, reaching what was previously Opus-tier quality on many coding and agentic tasks. On Orivel it takes the Anthropic light slot from Sonnet 4.6 at the same list price.

It is the first Sonnet-tier model with high-resolution vision (up to 2576 pixels on the long edge) and the first to support the full effort range through `xhigh` and `max`. Adaptive thinking is on by default, manual thinking budgets are gone, and non-default sampling parameters are rejected. It uses the newer tokenizer, so the same text counts roughly 30% more tokens than on Sonnet 4.6.

The model keeps a 1M-token context window and up to 128k tokens of output, with a January 2026 knowledge cutoff. Standard pricing is $3 input / $15 output per 1M tokens; Anthropic is running an introductory rate of $2 / $10 through August 31, 2026.

What changed

  • Released July 24, 2026 as the best speed-to-intelligence balance in the Claude lineup
  • Reaches previously Opus-tier quality on many coding and agentic tasks
  • First Sonnet-tier model with high-resolution vision (2576px long edge, up from 1568px)
  • First Sonnet with the full effort range, including `xhigh` and `max`
  • Adaptive thinking ON by default; manual thinking budgets removed; non-default temperature / top_p / top_k rejected
  • New tokenizer: the same text produces roughly 30% more tokens than on Sonnet 4.6
  • 1M-token context window; up to 128k output tokens
  • Standard pricing $3 input / $15 output per 1M tokens; introductory $2 / $10 through August 31, 2026
  • Knowledge cutoff: January 2026
Official announcement

Overall Performance

Overall Rank

#3

Overall win rate

78%

Average Score

85

Wins

7

Sample Count

9

Win Rate by Model

Compare by Genre

Strength by Evaluation Criteria

Average score by criterion (out of 10)

Completeness

88 3 samples

Safety

88 6 samples

Coherence

88 3 samples

Empathy

87 6 samples

Style Quality

87 3 samples

Appropriateness

86 6 samples

Clarity

85 9 samples

Structure

85 3 samples

Helpfulness

85 6 samples

Depth

85 3 samples

Instruction Following

85 6 samples

Emotional Impact

84 3 samples

Latest Tasks

Latest Discussions

Related Links

X f L