Alibaba Qwen price tracker
Qwen API Pricing 2026
Compare first-party Alibaba Model Studio aliases with hosted Qwen routes from Novita, Together, Castform, and Groq. The tables use the same daily-refreshed structured data as our main API comparison, including cache and long-context fields where the provider publishes them.
28 current routes across 4 providers · Rates refreshed · Qwen Studio plan status checked
First-party Alibaba aliases
3
Managed Qwen routes
25
Lowest tracked input rate
$0.03
Qwen3.5 4B
Alibaba Model Studio Qwen prices
These are direct first-party international Model Studio aliases in USD per 1 million tokens. Current promotions are temporary and have no published end date, so the acquisition gate checks the alias, list rate, promotional label, tier boundary, and rate every day.
Qwen3.7-Max | Alibaba | Mid | 1M | $1.25 | $0.125 | $3.75 |
Qwen3.6-Flash | Alibaba | Low | 1M | $0.25 | $0.025 | $1.50 |
Qwen3.7-Plus | Alibaba | Low | 1M | $0.32 | $0.032 | $1.28 |
- Qwen3.7-MaxAlibabaMid
- Input
- $1.25
- Cached
- $0.125
- Output
- $3.75
- Qwen3.6-FlashAlibabaLow
- Input
- $0.25
- Cached
- $0.025
- Output
- $1.50
- Qwen3.7-PlusAlibabaLow
- Input
- $0.32
- Cached
- $0.032
- Output
- $1.28
Sponsored links may earn us a commission at no extra cost to you. Affiliate status never changes model ordering.
All product names, logos, and brands are property of their respective owners and are used for identification purposes only.
Request-size pricing matters
Plus and Flash become more expensive above 256K input tokens
The headline table shows the standard tier. The structured rows for Qwen3.7-Plus and Qwen3.6-Flash also include the provider's higher tier for requests above 256K input tokens. The Qwen3.7-Max alias currently uses one promotional tier through its 1M context window.
Model your token mix in the calculator →Managed Qwen API prices
A managed route is a separate product and bill. It may expose open Qwen weights, an Alibaba closed-model route, or a provider-specific snapshot with different caching, context, throughput, and support. Compare exact IDs rather than assuming two rows with similar names are interchangeable.
Qwen3.5 9B | Together | Low | - | $0.17 | - | $0.25 |
Qwen3 235B A22B Instruct 2507 | Together | Low | - | $0.20 | - | $0.60 |
Qwen2.5 7B Instruct Turbo | Together | Low | - | $0.30 | - | $0.30 |
Qwen3.5 397B A17B | Together | Low | - | $0.60 | $0.35 | $3.60 |
Qwen3.7-Max | Novita | Mid | - | $1.25 | $0.25 | $3.75 |
Qwen3.8 Max | Novita | Mid | - | $2.00 | $0.25 | $6.00 |
CQwen3.5 4B | Ccastform | Low | - | $0.03 | - | $0.15 |
Qwen3 Coder 30B A3B Instruct | Novita | Low | - | $0.07 | - | $0.27 |
Qwen3 235B A22B Instruct 2507 | Novita | Low | - | $0.09 | - | $0.58 |
Qwen3 Next 80B A3B Instruct | Novita | Low | - | $0.15 | - | $1.50 |
Qwen3 235B A22B | Novita | Low | - | $0.20 | - | $0.80 |
Qwen3 Coder Next | Novita | Low | - | $0.20 | - | $1.50 |
Qwen3 VL 30B A3B Instruct | Novita | Low | - | $0.20 | - | $0.70 |
Qwen3.6-35B-A3B | Novita | Low | - | $0.248 | - | $1.49 |
Qwen MT Plus | Novita | Low | - | $0.25 | - | $0.75 |
Qwen3.5-35B-A3B | Novita | Low | - | $0.25 | - | $2.00 |
Qwen3 235B A22B Thinking 2507 | Novita | Low | - | $0.30 | - | $3.00 |
Qwen3 VL 235B A22B Instruct | Novita | Low | - | $0.30 | - | $1.50 |
Qwen3.5-27B | Novita | Low | - | $0.30 | - | $2.40 |
Qwen 2.5 72B Instruct | Novita | Low | - | $0.38 | - | $0.40 |
Qwen3 Coder 480B A35B Instruct | Novita | Low | - | $0.38 | - | $1.55 |
Qwen3.5-122B-A10B | Novita | Low | - | $0.40 | - | $3.20 |
Qwen3.5-397B-A17B | Novita | Low | - | $0.60 | - | $3.60 |
Qwen3.6-27B | Novita | Low | - | $0.60 | - | $3.60 |
Qwen3 VL 235B A22B Thinking | Novita | Low | - | $0.98 | - | $3.95 |
- Qwen3.5 9BTogetherLow
- Input
- $0.17
- Cached
- -
- Output
- $0.25
- Qwen3 235B A22B Instruct 2507TogetherLow
- Input
- $0.20
- Cached
- -
- Output
- $0.60
- Qwen2.5 7B Instruct TurboTogetherLow
- Input
- $0.30
- Cached
- -
- Output
- $0.30
- Qwen3.5 397B A17BTogetherLow
- Input
- $0.60
- Cached
- $0.35
- Output
- $3.60
- Qwen3.7-MaxNovitaMid
- Input
- $1.25
- Cached
- $0.25
- Output
- $3.75
- Qwen3.8 MaxNovitaMid
- Input
- $2.00
- Cached
- $0.25
- Output
- $6.00
- CQwen3.5 4BCcastformLow
- Input
- $0.03
- Cached
- -
- Output
- $0.15
- Qwen3 Coder 30B A3B InstructNovitaLow
- Input
- $0.07
- Cached
- -
- Output
- $0.27
- Qwen3 235B A22B Instruct 2507NovitaLow
- Input
- $0.09
- Cached
- -
- Output
- $0.58
- Qwen3 Next 80B A3B InstructNovitaLow
- Input
- $0.15
- Cached
- -
- Output
- $1.50
- Qwen3 235B A22BNovitaLow
- Input
- $0.20
- Cached
- -
- Output
- $0.80
- Qwen3 Coder NextNovitaLow
- Input
- $0.20
- Cached
- -
- Output
- $1.50
- Qwen3 VL 30B A3B InstructNovitaLow
- Input
- $0.20
- Cached
- -
- Output
- $0.70
- Qwen3.6-35B-A3BNovitaLow
- Input
- $0.248
- Cached
- -
- Output
- $1.49
- Qwen MT PlusNovitaLow
- Input
- $0.25
- Cached
- -
- Output
- $0.75
- Qwen3.5-35B-A3BNovitaLow
- Input
- $0.25
- Cached
- -
- Output
- $2.00
- Qwen3 235B A22B Thinking 2507NovitaLow
- Input
- $0.30
- Cached
- -
- Output
- $3.00
- Qwen3 VL 235B A22B InstructNovitaLow
- Input
- $0.30
- Cached
- -
- Output
- $1.50
- Qwen3.5-27BNovitaLow
- Input
- $0.30
- Cached
- -
- Output
- $2.40
- Qwen 2.5 72B InstructNovitaLow
- Input
- $0.38
- Cached
- -
- Output
- $0.40
- Qwen3 Coder 480B A35B InstructNovitaLow
- Input
- $0.38
- Cached
- -
- Output
- $1.55
- Qwen3.5-122B-A10BNovitaLow
- Input
- $0.40
- Cached
- -
- Output
- $3.20
- Qwen3.5-397B-A17BNovitaLow
- Input
- $0.60
- Cached
- -
- Output
- $3.60
- Qwen3.6-27BNovitaLow
- Input
- $0.60
- Cached
- -
- Output
- $3.60
- Qwen3 VL 235B A22B ThinkingNovitaLow
- Input
- $0.98
- Cached
- -
- Output
- $3.95
Sponsored links may earn us a commission at no extra cost to you. Affiliate status never changes model ordering.
All product names, logos, and brands are property of their respective owners and are used for identification purposes only.
Is Qwen Studio free?
Qwen's public Studio experience does not currently expose a stable monthly subscription rate card or fixed public quota that we can maintain as a plan. That is why Qwen is not inserted into the subscription matrix as an invented zero-dollar plan. The separate API has published token billing and selected activation quotas.
See Qwen Studio plan and limit status →Want one endpoint for Qwen and DeepSeek?
Novita is an active AI Pricing Guru partner and carries multiple Qwen routes alongside DeepSeek and other open-model families. Verify the exact model ID and run the same acceptance test before choosing a host.
Check Novita's current catalog →Affiliate disclosure: we may earn a commission at no extra cost to you. Partner status does not change the table or ordering.
Which Qwen model should you choose?
Start with Flash for inexpensive classification, extraction, retrieval, and short structured responses. Move to Plus when the task needs stronger reasoning, coding, or long-context work. Use Max as an escalation route only when it produces enough fewer failures or revisions to justify the higher accepted-task cost.
For coding-specific open models, compare the hosted Qwen Coder rows in the managed table and read the Qwen Code pricing guide. For buyer-oriented price and capability tradeoffs, see Qwen vs DeepSeek API pricing.
Open weights are not the same as a free API
Many Qwen models can be downloaded under permissive model licenses, but production self-hosting still pays for accelerators, memory, electricity or rental, redundancy, monitoring, and engineering time. Compare those costs in the local AI versus API calculator before treating a downloadable checkpoint as free inference.
Frequently asked questions
How much does the Qwen API cost?
Qwen pricing depends on the exact model, request size, cache use, region, and host. This page currently tracks 28 active Qwen routes across 4 providers. The live tables show the current input, cached-input, and output rates per 1 million tokens.
Is Qwen free?
Qwen Studio provides a consumer chat experience, but its public site does not publish a stable paid-plan rate card or fixed public usage allowance. Alibaba Model Studio separately offers activation quotas for selected international API models; after the applicable quota, API usage is billed per token. Do not treat free chat access as a free production API.
What is the cheapest hosted Qwen model?
Qwen3.5 4B is the lowest current tracked Qwen route by standard input rate at $0.03 per 1 million input tokens. Output price, retries, context tier, and host reliability can change the cheapest accepted-task result.
Should I use Alibaba Model Studio or a managed Qwen host?
Use Alibaba Model Studio when you want first-party aliases, regional controls, and Qwen-specific features. Use a managed host when one compatible endpoint, cross-family routing, or an existing cloud relationship matters more. Benchmark the exact snapshot because model names, cache rules, and prices can differ by host.
Does Qwen long-context pricing cost more?
Some first-party Qwen aliases change rates when the request crosses a documented input-token threshold. The Qwen3.7-Plus and Qwen3.6-Flash rows on this page carry their higher long-context tiers in the live structured dataset rather than hiding the higher rate in prose.
Sources and maintenance
First-party models are checked against Alibaba Cloud's model catalog, international pricing page, and cache rules. Managed rows use each host's own rate card. Qwen Studio status is checked against Qwen's official product. All are supervised on a daily cadence.