← Back to Leaderboard
MMistral | Mistral Large 4 0
Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 1M-token context window with up to 256K output tokens, and supports tool calling and structured outputs.
Fair
44.3
Most Recent TestSignificant guardrail issues may impede work.
Strengths
- Excellent at 3.2. Historical Jesus
Weaknesses
- Limited task completion capability
- Weak doctrinal alignment
- Struggles with 1.1. Missiological Research
- Struggles with 1.7. Difficult Passages
- Struggles with 2.3. Reality of Judgment
Overall Score
44.3
Tier 1 (Task) 70%
46.7
Tier 2 (Doctrine) 20%
26.7
Tier 3 (Worldview) 10%
63.3
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description
Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 1M-token context window with up to 256K output tokens, and supports tool calling and structured outputs.
Provider
Mistral
Model ID
mistralai/mistral-large-4-0
Tests Run
1
Categories
Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
37
1.1
50
1.2
50
1.3
60
1.4
43
1.5
70
1.6
17
1.7
60
2.1
40
2.2
10
2.3
0
2.4
40
2.5
10
2.6
25
3.1
100
3.2
63
3.3
75
3.4
50
3.5
67
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession
Recent Tests
Recent Test Runs
| Date | Score | Version | Tier 1 (Task) | Tier 2 (Gospel) | Tier 3 (Worldview) | Trust Tier |
|---|---|---|---|---|---|---|
| 10/10/2026 | 44.3 | 1.0.0 | 46.7 | 26.7 | 63.3 | automated |