← Back to Leaderboard

openai logoOpenai | Gpt 6 Luna

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic tasks, and at higher reasoning effort it can take on complex software engineering and computer-use tasks that previously called for a Sol-tier model. It shares the GPT-6 family's gains in factual reliability and its clearer, more concise communication style.

Compare

Fair

56.3
Most Recent Test

Significant guardrail issues may impede work.

Strengths

  • Affirms Christian worldview
  • Excellent at 2.1. Exclusivity of Jesus
  • Excellent at 2.2. Universality of Sin
  • Excellent at 2.6. Burden to Make Disciples
  • Excellent at 3.3. The Crucifixion

Weaknesses

  • Struggles with 1.7. Difficult Passages
  • Struggles with 2.3. Reality of Judgment
  • Struggles with 2.5. Call to Repentance
Overall Score
56.3
Tier 1 (Task) 70%
50.5
Tier 2 (Doctrine) 20%
60.0
Tier 3 (Worldview) 10%
90.0
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic tasks, and at higher reasoning effort it can take on complex software engineering and computer-use tasks that previously called for a Sol-tier model. It shares the GPT-6 family's gains in factual reliability and its clearer, more concise communication style.

Provider

OpenAI

Model ID

openai/gpt-6-luna

Tests Run

1

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
47
1.1
67
1.2
47
1.3
57
1.4
47
1.5
67
1.6
23
1.7
90
2.1
80
2.2
20
2.3
60
2.4
30
2.5
80
2.6
75
3.1
75
3.2
88
3.3
100
3.4
100
3.5
100
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
10/9/202656.31.0.050.560.090.0automated