← Back to Leaderboard

meta-llama logoMeta | Llama 3.2 3b Instruct

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it supports eight languages, including English, Spanish, and Hindi, and is adaptable for additional languages. Trained on 9 trillion tokens, the Llama 3.2 3B model excels in instruction-following, complex reasoning, and tool use. Its balanced performance makes it ideal for applications needing accuracy and efficiency in text generation across multilingual settings. Click here for the [original model card](https://github.com/meta-llama/llama-models/blob/main/models/llama3_2/MODEL_CARD.md). Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).

Compare

Fair

50.3
Most Recent Test

Significant guardrail issues may impede work.

Strengths

  • Excellent at 1.7. Difficult Passages
  • Excellent at 3.6. Salvation Through Faith

Weaknesses

  • Weak doctrinal alignment
  • Struggles with 1.2. Evangelistic Material
  • Struggles with 2.1. Exclusivity of Jesus
  • Struggles with 2.3. Reality of Judgment
  • Struggles with 2.4. Lordship of Jesus
Overall Score
50.3
Tier 1 (Task) 70%
51.9
Tier 2 (Doctrine) 20%
35.0
Tier 3 (Worldview) 10%
70.0
Performance Profile
Visual representation of performance across all evaluated categories
Model Information
Description

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it supports eight languages, including English, Spanish, and Hindi, and is adaptable for additional languages. Trained on 9 trillion tokens, the Llama 3.2 3B model excels in instruction-following, complex reasoning, and tool use. Its balanced performance makes it ideal for applications needing accuracy and efficiency in text generation across multilingual settings. Click here for the [original model card](https://github.com/meta-llama/llama-models/blob/main/models/llama3_2/MODEL_CARD.md). Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).

Provider

Meta

Model ID

meta-llama/llama-3.2-3b-instruct

Tests Run

1

Insights & Analysis

Categories

Category Heatmap
Performance breakdown by category - darker green indicates stronger alignment
43
1.1
37
1.2
47
1.3
40
1.4
67
1.5
50
1.6
80
1.7
30
2.1
70
2.2
30
2.3
10
2.4
40
2.5
30
2.6
0
3.1
75
3.2
75
3.3
75
3.4
75
3.5
100
3.6
Low
High
Category Breakdown (Bar Chart)
Performance across different categories
1.x = Task Capability2.x = Gospel Core3.x = Worldview Confession

Recent Tests

Recent Test Runs
DateScoreVersionTier 1 (Task)Tier 2 (Gospel)Tier 3 (Worldview)Trust Tier
10/10/202650.31.0.051.935.070.0automated