← Back to Leaderboard

xiaomi logoXiaomi | Mimo V2.6 Flash

MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for greater computational efficiency. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers strong performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

Compare

Good

63.0
Most Recent Test

Usable with some limitations.

Strengths

  • Excellent at 1.2. Evangelistic Material
  • Excellent at 1.3. Apologetics
  • Excellent at 2.2. Universality of Sin
  • Excellent at 2.3. Reality of Judgment
  • Excellent at 3.3. The Crucifixion

Weaknesses

  • Struggles with 1.1. Missiological Research
  • Struggles with 3.1. Existence of God
Overall Score
63.0
Tier 1 (Task) 70%
63.3
Tier 2 (Doctrine) 20%
60.0
Tier 3 (Worldview) 10%
66.7
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for greater computational efficiency. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers strong performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

Provider

Xiaomi

Model ID

xiaomi/mimo-v2.6-flash

Tests Run

1

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
37
1.1
83
1.2
83
1.3
63
1.4
57
1.5
70
1.6
50
1.7
40
2.1
80
2.2
80
2.3
70
2.4
40
2.5
50
2.6
0
3.1
75
3.2
88
3.3
50
3.4
75
3.5
83
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
10/10/202663.01.0.063.360.066.7automated