Qwen logo

Qwen AI models

3 Qwen models on FastRouter, all behind one OpenAI-compatible API. Compare pricing, context windows and benchmarks, then open any model for its providers and code samples.

Filter in catalog
Models
3
Largest context
41K
Qwen3 30B A3B
Lowest input /1M
$0.12
Qwen3 30B A3B
Top intelligence
7.7
Qwen2.5 72B Instruct

All Qwen models

Qwen logo

A Mixture-of-Experts model that activates only 3.3B of its 30.5B parameters per forward pass. Offers dual thinking modes: a detailed step-by-step reasoning mode for complex problems and a faster non-thinking mode for simpler queries. Despite its small active parameter count, it outperforms QwQ-32B that has 10 times more activated parameters.

Qwen/Qwen3-30B-A3BApr 28, 20252.8s latency
Context
41K
Price /1M
$0.12 in$0.50 out
Intel
6.6
Qwen logo

A versatile dense model offering powerful reasoning abilities paired with efficient dialogue processing. Developed by Alibaba, it supports extensive 32K context windows (extendable to 131K), features both thinking and non-thinking modes, and excels in coding, mathematics, and multilingual support across 119 languages and dialects.

Qwen/Qwen3-14BApr 28, 20252.3s latency
Context
41K
Price /1M
$0.12 in$0.24 out
Intel
—
Qwen logo

Alibaba's largest Qwen2.5 model featuring improved capabilities in coding, mathematics, and instruction following across more than 29 languages. With 72B parameters and 128K context support, it delivers top-tier performance for complex tasks requiring deep context understanding and sophisticated reasoning.

Qwen/Qwen2.5-72B-InstructSep 19, 202415s latency
Context
33K
Price /1M
$0.36 in$0.40 out
Intel
7.7

Frequently asked questions

Model scores sourced from ArtificialAnalysis

Qwen AI Models: API Pricing & Benchmarks | FastRouter.ai