← Back to Leaderboard

M
Mistral | Mistral Large 4 0

Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 1M-token context window with up to 256K output tokens, and supports tool calling and structured outputs.

Compare

Fair

44.3
Most Recent Test

Significant guardrail issues may impede work.

Strengths

  • Excellent at 3.2. Historical Jesus

Weaknesses

  • Limited task completion capability
  • Weak doctrinal alignment
  • Struggles with 1.1. Missiological Research
  • Struggles with 1.7. Difficult Passages
  • Struggles with 2.3. Reality of Judgment
Overall Score
44.3
Tier 1 (Task) 70%
46.7
Tier 2 (Doctrine) 20%
26.7
Tier 3 (Worldview) 10%
63.3
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 1M-token context window with up to 256K output tokens, and supports tool calling and structured outputs.

Provider

Mistral

Model ID

mistralai/mistral-large-4-0

Tests Run

1

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
37
1.1
50
1.2
50
1.3
60
1.4
43
1.5
70
1.6
17
1.7
60
2.1
40
2.2
10
2.3
0
2.4
40
2.5
10
2.6
25
3.1
100
3.2
63
3.3
75
3.4
50
3.5
67
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
10/10/202644.31.0.046.726.763.3automated