models/qwen/qwen3-8-max
Q
Qwen·active

Qwen: Qwen3.8 Max

Qwen's flagship model. Long-context specialist with 1M window.

Overall score
4.69
/5.00 · ranked #21
Input
$2.00
per 1M tokens
Output
$6.00
per 1M tokens
Context
1M
tokens
Blended
$5.00
3:1 out:in ratio

Price drops, new benchmarks, model updates. Stay current on Qwen: Qwen3.8 Max.

One email per change. Unsubscribe anytime.

modelpicker.aipowered by live benchmark data

Scores by test

Methodology →
Structured Output
5.0
Strategic Analysis
5.0
Constrained Rewriting
4.0
Creative Problem Solving
4.0
Tool Calling
5.0
Faithfulness
5.0
Classification
4.0
Long Context
5.0
Safety Calibration
5.0
Persona Consistency
5.0
Agentic Planning
4.0
Multilingual
5.0
Tabular Data
5.0
AIME 2025
99.4
GPQA Diamond
92.7
SciCode
53.2
APEX Agents
63.3
Epoch Capabilities Index (ECI)
156.6

What you need to know

Qwen3.8 Max is a high-performance frontier model characterized by exceptional reliability in structured tasks and complex reasoning. It achieves perfect internal scores in structured output, tool calling, and strategic analysis, making it a primary candidate for automation pipelines that require strict schema adherence. Its external performance is equally strong, particularly in high-difficulty reasoning, as evidenced by a 99.4% score on AIME 2025 and a 92.7% score on GPQA Diamond.

The model is designed for massive data ingestion, supporting a 1M token context window with a perfect internal score for long-context retrieval. While it maintains a high average quality score of 4.69/5.0, it shows slight relative weaknesses in classification, agentic planning, and constrained rewriting, though these remain strong at 4/5. Its Epoch Capabilities Index of 156.42 places it firmly within the range of top-tier frontier models.

At a blended cost of $5.00 per million tokens, this model is positioned as a premium offering. The pricing is consistent with other high-reasoning frontier models, meaning you are paying for accuracy and context capacity rather than cost-efficiency. It does not offer open weights, requiring a provider-based API integration.

Use this model if your project requires a massive context window, high-precision mathematical reasoning, or reliable tool calling for complex workflows. Skip this model if you are optimizing for low-cost inference or if your primary need is simple text classification.

Strengths — Top 3

Structured Output5.0/5.0
Strategic Analysis5.0/5.0
Tool Calling5.0/5.0

Relative weaknesses — Bottom 3

Constrained Rewriting4.0/5.0
Creative Problem Solving4.0/5.0
Classification4.0/5.0

Similar models

OOpenAI: GPT-5.6 Terra Pro$11.884.77ZZ.ai: GLM 5.3 Flash$0.4124.85ZGLM-4.7$1.804.69GGoogle: Gemini 3.8 Flash$3.004.69