← Back to Leaderboard

Q
Qwen | Qwen3.8 Omni Flash

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization, video editing and production workflows, audio-video dialogue, coding, knowledge work, and GUI interaction, and it is particularly strong at long-form multimedia tasks that combine speech, sound, and visual context. It also supports two-channel and four-channel spatial audio understanding.

Compare

Good

72.7
Most Recent Test

Usable with some limitations.

Strengths

  • Maintains doctrinal fidelity
  • Affirms Christian worldview
  • Excellent at 2.1. Exclusivity of Jesus
  • Excellent at 2.2. Universality of Sin
  • Excellent at 2.3. Reality of Judgment

Weaknesses

No notable weaknesses identified

Overall Score
72.7
Tier 1 (Task) 70%
66.7
Tier 2 (Doctrine) 20%
80.0
Tier 3 (Worldview) 10%
100.0
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization, video editing and production workflows, audio-video dialogue, coding, knowledge work, and GUI interaction, and it is particularly strong at long-form multimedia tasks that combine speech, sound, and visual context. It also supports two-channel and four-channel spatial audio understanding.

Provider

Qwen

Model ID

qwen/qwen3.8-omni-flash

Tests Run

1

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
43
1.1
67
1.2
77
1.3
77
1.4
67
1.5
67
1.6
70
1.7
90
2.1
100
2.2
90
2.3
70
2.4
40
2.5
90
2.6
100
3.1
100
3.2
100
3.3
100
3.4
100
3.5
100
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
10/10/202672.71.0.066.780.0100.0automated