← Back to Leaderboard

xiaomi logoXiaomi | Mimo V2.6 Pro

MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding workloads. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers top-tier performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

Compare

Good

63.7
Most Recent Test

Usable with some limitations.

Strengths

  • Affirms Christian worldview
  • Excellent at 1.3. Apologetics
  • Excellent at 2.1. Exclusivity of Jesus
  • Excellent at 2.2. Universality of Sin
  • Excellent at 2.3. Reality of Judgment

Weaknesses

  • Struggles with 2.4. Lordship of Jesus
  • Struggles with 2.5. Call to Repentance
Overall Score
63.7
Tier 1 (Task) 70%
62.9
Tier 2 (Doctrine) 20%
56.7
Tier 3 (Worldview) 10%
83.3
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding workloads. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers top-tier performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

Provider

Xiaomi

Model ID

xiaomi/mimo-v2.6-pro

Tests Run

1

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
40
1.1
60
1.2
83
1.3
63
1.4
60
1.5
67
1.6
67
1.7
80
2.1
80
2.2
80
2.3
20
2.4
20
2.5
60
2.6
50
3.1
100
3.2
88
3.3
100
3.4
50
3.5
100
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
10/10/202663.71.0.062.956.783.3automated