← Back to Leaderboard

anthropic logoAnthropic | Claude Haiku 5.5

Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and knowledge work, and accepts text and image input with a 1M-token context window. It is the first Haiku model with adjustable effort. Thinking is adaptive and on by default, with effort as the main lever for trading off depth, latency, and cost; it can be turned off at low, medium, and high effort.

Compare

Good

68.3
Most Recent Test

Usable with some limitations.

Strengths

  • Maintains doctrinal fidelity
  • Affirms Christian worldview
  • Excellent at 1.3. Apologetics
  • Excellent at 2.1. Exclusivity of Jesus
  • Excellent at 2.2. Universality of Sin

Weaknesses

  • Struggles with 1.1. Missiological Research
Overall Score
68.3
Tier 1 (Task) 70%
61.9
Tier 2 (Doctrine) 20%
76.7
Tier 3 (Worldview) 10%
96.7
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and knowledge work, and accepts text and image input with a 1M-token context window. It is the first Haiku model with adjustable effort. Thinking is adaptive and on by default, with effort as the main lever for trading off depth, latency, and cost; it can be turned off at low, medium, and high effort.

Provider

Anthropic

Model ID

anthropic/claude-haiku-5.5

Tests Run

1

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
33
1.1
73
1.2
87
1.3
57
1.4
67
1.5
73
1.6
43
1.7
100
2.1
90
2.2
60
2.3
80
2.4
60
2.5
70
2.6
75
3.1
100
3.2
100
3.3
100
3.4
100
3.5
100
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
10/9/202668.31.0.061.976.796.7automated