models/qwen/qwen3-6-plus
Q
Qwen·active

Qwen: Qwen3.6 Plus

Qwen's mid-tier model. Long-context specialist with 1M window.

Overall score
4.54
/5.00 · ranked #31
Input
$0.325
per 1M tokens
Output
$1.95
per 1M tokens
Context
1M
tokens
Blended
$1.54
3:1 out:in ratio

Price drops, new benchmarks, model updates. Stay current on Qwen: Qwen3.6 Plus.

One email per change. Unsubscribe anytime.

modelpicker.aipowered by live benchmark data

Scores by test

Methodology →
Structured Output
5.0
Strategic Analysis
5.0
Constrained Rewriting
5.0
Creative Problem Solving
4.0
Tool Calling
4.0
Faithfulness
5.0
Classification
4.0
Long Context
5.0
Safety Calibration
2.0
Persona Consistency
5.0
Agentic Planning
5.0
Multilingual
5.0
Tabular Data
5.0
SWE-bench Verified
57.9
AIME 2025
90.6
GPQA Diamond
87.4
SciCode
40.7
Epoch Capabilities Index (ECI)
149.0

What you need to know

Qwen3.6 Plus is engineered for high-precision structural tasks and complex logic, achieving perfect internal scores in strategic analysis, agentic planning, and structured output. Its primary differentiator is a massive 1M token context window paired with a 5/5 rating for long-context faithfulness, making it highly reliable for processing extensive datasets or large codebases without losing coherence.

The model is priced at $0.325 per million input tokens and $1.95 per million output tokens. Given its rank as #9 out of 71 models and an average internal score of 4.54, it offers a high performance-to-cost ratio for developers needing frontier-level reasoning without the premium pricing of the top three models.

The most significant technical risk is safety calibration, which scored a 2/5. This indicates a higher likelihood of producing unfiltered or non-compliant responses compared to more heavily aligned models. While it excels in multilingual tasks and persona consistency, it is slightly less effective at creative problem solving and tool calling.

Use this model if you require a large context window for complex agentic workflows, structured data extraction, or strategic analysis. Skip this model if your application requires strict safety guardrails or primary reliance on autonomous tool calling.

Strengths — Top 3

Structured Output5.0/5.0
Strategic Analysis5.0/5.0
Constrained Rewriting5.0/5.0

Relative weaknesses — Bottom 3

Safety Calibration2.0/5.0
Creative Problem Solving4.0/5.0
Tool Calling4.0/5.0

Similar models

NNVIDIA: Nemotron 3 Super$0.3214.46OGPT-5$7.814.54DDeepSeek V3.2$0.3674.31DDeepSeek V4 Pro$0.7614.46