models/qwen/qwen3-30b-a3b-instruct-2507
Q
Qwen·active

Qwen: Qwen3 30B A3B Instruct 2507

Qwen's efficiency model. Context window: 262K tokens.

Overall score
3.85
/5.00 · ranked #104
Input
$0.048
per 1M tokens
Output
$0.193
per 1M tokens
Context
262K
tokens
Blended
$0.157
3:1 out:in ratio

Price drops, new benchmarks, model updates. Stay current on Qwen: Qwen3 30B A3B Instruct 2507.

One email per change. Unsubscribe anytime.

modelpicker.aipowered by live benchmark data

Scores by test

Methodology →
Structured Output
5.0
Strategic Analysis
5.0
Constrained Rewriting
4.0
Creative Problem Solving
4.0
Tool Calling
4.0
Faithfulness
5.0
Classification
3.0
Long Context
3.0
Safety Calibration
1.0
Persona Consistency
4.0
Agentic Planning
4.0
Multilingual
5.0
Tabular Data
3.0

What you need to know

Qwen3 30B A3B Instruct 2507 is optimized for precision and reliability in structured tasks. It achieves perfect scores in structured output, faithfulness, and multilingual capabilities, making it a strong candidate for pipelines requiring strict adherence to formats and high factual accuracy across different languages.

Despite a massive 262K context window, the model underperforms in long-context retrieval and tabular data processing, both scoring 3/5. This suggests that while the model can accept large inputs, its ability to effectively reason over that data is limited. Additionally, a critical failure in safety calibration (1/5) means this model lacks the internal guardrails necessary for user-facing applications without external filtering.

At a blended cost of $0.157/MTok, the model is priced competitively for its size. It provides high-tier strategic analysis and agentic planning at a fraction of the cost of frontier models, offering a favorable trade-off for backend automation where safety is managed at the application layer.

Use this model for multilingual automation, complex strategic analysis, or any workflow requiring rigid structured outputs. Skip this model for customer-facing chatbots due to poor safety calibration or for tasks requiring deep analysis of very long documents.

Strengths — Top 3

Structured Output5.0/5.0
Strategic Analysis5.0/5.0
Faithfulness5.0/5.0

Relative weaknesses — Bottom 3

Safety Calibration1.0/5.0
Classification3.0/5.0
Long Context3.0/5.0

Similar models

AArcee AI: Trinity Large Thinking$0.6934.23MMistral Large 3 2512$1.253.69MMistral Medium 3.5$6.004.15IinclusionAI: Ling-2.6-flash$0.0254.08