models/mistral/codestral-2508
M
Mistral·active

Codestral 2508

Mistral's efficiency model. Context window: 256K tokens.

Overall score
3.46
/5.00 · ranked #135
Input
$0.300
per 1M tokens
Output
$0.900
per 1M tokens
Context
256K
tokens
Blended
$0.750
3:1 out:in ratio

Price drops, new benchmarks, model updates. Stay current on Codestral 2508.

One email per change. Unsubscribe anytime.

modelpicker.aipowered by live benchmark data

Scores by test

Methodology →
Structured Output
5.0
Strategic Analysis
2.0
Constrained Rewriting
3.0
Creative Problem Solving
2.0
Tool Calling
5.0
Faithfulness
5.0
Classification
3.0
Long Context
5.0
Safety Calibration
2.0
Persona Consistency
3.0
Agentic Planning
4.0
Multilingual
4.0
Tabular Data
2.0

What you need to know

Codestral 2508 is a specialized tool for high-precision technical tasks, distinguished by perfect scores in faithfulness, structured output, and long-context handling. With a 256K context window and a 5/5 rating for long-context performance, it is built for processing massive codebases or documentation sets without losing coherence. Its ability to strictly adhere to schemas and provide factual, grounded responses makes it reliable for pipeline integration where output format is critical.

The model struggles with cognitive flexibility and complex reasoning. It scores poorly in strategic analysis, creative problem solving, and tabular data processing, all landing at 2/5. This indicates a significant limitation in high-level architectural planning or data manipulation tasks that require synthesis rather than direct execution.

At a blended cost of $0.750/MTok, the model is priced competitively for its specialized capabilities, though its overall rank of 118 out of 130 suggests it is not a general-purpose replacement for larger frontier models. It trades broad intelligence for reliability in specific technical domains.

Use this model if you need a cost-effective engine for tool calling, structured data extraction, or analyzing large files where factual accuracy is non-negotiable. Skip this model if your use case requires complex strategic planning, creative brainstorming, or the analysis of tabular datasets.

Strengths — Top 3

Structured Output5.0/5.0
Tool Calling5.0/5.0
Faithfulness5.0/5.0

Relative weaknesses — Bottom 3

Strategic Analysis2.0/5.0
Creative Problem Solving2.0/5.0
Safety Calibration2.0/5.0

Similar models

QQwen: Qwen3 Coder 30B A3B Instruct$0.2283.23MLlama 3.3 70B Instruct$0.7103.46OOpenAI: gpt-oss-20b$0.1053.62MLlama 4 Scout$0.2833.31