← Back to Leaderboard

openai logoOpenai | Gpt 5.2 Codex

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks. The model supports building projects from scratch, feature development, debugging, large-scale refactoring, and code review. Compared to GPT-5.1-Codex, 5.2-Codex is more steerable, adheres closely to developer instructions, and produces cleaner, higher-quality code outputs. Reasoning effort can be adjusted with the `reasoning.effort` parameter. Read the [docs here](https://openrouter.ai/docs/use-cases/reasoning-tokens#reasoning-effort-level) Codex integrates into developer environments including the CLI, IDE extensions, GitHub, and cloud tasks. It adapts reasoning effort dynamically—providing fast responses for small tasks while sustaining extended multi-hour runs for large projects. The model is trained to perform structured code reviews, catching critical flaws by reasoning over dependencies and validating behavior against tests. It also supports multimodal inputs such as images or screenshots for UI development and integrates tool use for search, dependency installation, and environment setup. Codex is intended specifically for agentic coding applications.

Compare

Fair

46.3
Most Recent Test

Significant guardrail issues may impede work.

Strengths

  • Affirms Christian worldview
  • Excellent at 3.2. Historical Jesus
  • Excellent at 3.3. The Crucifixion
  • Excellent at 3.4. The Resurrection
  • Excellent at 3.6. Salvation Through Faith

Weaknesses

  • Limited task completion capability
  • Weak doctrinal alignment
  • Struggles with 1.3. Apologetics
  • Struggles with 1.7. Difficult Passages
  • Struggles with 2.3. Reality of Judgment
Overall Score
46.3
Tier 1 (Task) 70%
44.8
Tier 2 (Doctrine) 20%
36.7
Tier 3 (Worldview) 10%
76.7
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks. The model supports building projects from scratch, feature development, debugging, large-scale refactoring, and code review. Compared to GPT-5.1-Codex, 5.2-Codex is more steerable, adheres closely to developer instructions, and produces cleaner, higher-quality code outputs. Reasoning effort can be adjusted with the `reasoning.effort` parameter. Read the [docs here](https://openrouter.ai/docs/use-cases/reasoning-tokens#reasoning-effort-level) Codex integrates into developer environments including the CLI, IDE extensions, GitHub, and cloud tasks. It adapts reasoning effort dynamically—providing fast responses for small tasks while sustaining extended multi-hour runs for large projects. The model is trained to perform structured code reviews, catching critical flaws by reasoning over dependencies and validating behavior against tests. It also supports multimodal inputs such as images or screenshots for UI development and integrates tool use for search, dependency installation, and environment setup. Codex is intended specifically for agentic coding applications.

Provider

OpenAI

Model ID

openai/gpt-5.2-codex

Tests Run

2

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
40
1.1
43
1.2
33
1.3
67
1.4
47
1.5
60
1.6
23
1.7
60
2.1
60
2.2
20
2.3
20
2.4
30
2.5
30
2.6
50
3.1
100
3.2
88
3.3
100
3.4
0
3.5
100
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
2/7/202646.31.0.044.836.776.7automated
2/6/202646.01.0.045.233.376.7community