Qwen · Model
Qwen3.5-122B-A10B pricing, context window and specs
Qwen3.5-122B-A10B costs $0.26 per million input tokens and $2.08 per million output tokens, and reads up to 262,144 tokens per request. By blended price it is number 8 of Qwen’s 13 current models here, counting from the cheapest.
- Input / 1M
- $0.26
- Output / 1M
- $2.08
- Cached / 1M
- –
- Context
- 262,144
- Max output
- 65,536
- Released
- 2026-02-25
Estimate your cost with Qwen3.5-122B-A10BCompare all Qwen modelsCount tokens for Qwen3.5-122B-A10B
Pricing
Qwen3.5-122B-A10B API pricing
| Price per 1M tokens | Input | Output |
|---|---|---|
| Standard | $0.26 | $2.08 |
Output costs 8.0× the input price, so long replies drive the bill more than long prompts. No cached-input price is published, so repeated prompt prefixes are billed at the full input price.
Examples
What Qwen3.5-122B-A10B costs in practice
| Workload | Tokens in / out | Requests/day | Per request | Per month |
|---|---|---|---|---|
| Support chatbot | 1,500 / 400 | 1,000 | $0.00122 | $37.17 |
| RAG app | 6,000 / 500 | 500 | $0.0026 | $39.54 |
| Coding agent | 40,000 / 2,000 | 200 | $0.0146 | $88.57 |
| Document summariser | 8,000 / 600 | 300 | $0.00333 | $30.37 |
Change any number in the cost calculator.
Context
Context window
Qwen3.5-122B-A10B accepts up to 262,144 tokens per request, roughly 196,608 English words or 393 pages. That budget is shared between your prompt and the reply, and a single reply is capped at 65,536 tokens.
Filling the whole window with one prompt costs about $0.0682 at list price, so for repeated long-document work, retrieval (sending only the relevant parts) or prompt caching is usually much cheaper.
Features
Capabilities
- Image input
- Yes
- Tool / function calling
- Yes
- Reasoning mode
- Yes
- Prompt caching
- No
- Structured output
- Yes
- Open weights
- Yes
- Accepts
- text, image, video
- Knowledge cutoff
- Not published
As declared by the provider’s API. It shows a feature exists, not how well it works.
API
Qwen3.5-122B-A10B model ID
- OpenRouter
qwen/qwen3.5-122b-a10b- Open weights
- Qwen/Qwen3.5-122B-A10B
Use the OpenRouter ID when calling it through OpenRouter; Qwen’s own API may name it differently. See the model ID reference for the providers we document.
Qwen
Where Qwen3.5-122B-A10B sits in Qwen’s line-up
Ranked by blended price (three parts input to one part output, $0.715 per 1M for Qwen3.5-122B-A10B), it is number 8 of Qwen’s 13 current models here, counting from the cheapest. One step down, Qwen3.7 Plus costs $0.56 blended (22% less). Unlike Qwen3.5-122B-A10B, it doesn’t accept video, reads up to 1,000,000 tokens (262,144 here) and caps a reply at 131,072 tokens (65,536 here). One step up, Qwen3.6 27B costs $1.04 (1.5× as much). Unlike Qwen3.5-122B-A10B, it caps a reply at 81,920 tokens (65,536 here). See every Qwen model and price
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Qwen3.5-9B | $0.10 | $0.15 | 256,000 | 2026-03-10 |
| Qwen3.8 Flash | $0.15 | $0.47 | 1,000,000 | 2026-08-26 |
| Qwen3.8 Omni Flash | $0.15 | $0.47 | 1,000,000 | 2026-09-21 |
| Qwen3 Coder Next | $0.12 | $0.80 | 262,144 | 2026-02-04 |
| Qwen3.6 35B A3B | $0.10 | $0.95 | 262,144 | 2026-04-27 |
| Qwen3 VL 8B Instruct | $0.25 | $0.75 | 262,144 | 2025-10-14 |
| Qwen3.7 Plus | $0.32 | $1.28 | 1,000,000 | 2026-06-03 |
| Qwen3.5-122B-A10B (this page) | $0.26 | $2.08 | 262,144 | 2026-02-25 |
| Qwen3.6 27B | $0.32 | $3.20 | 262,144 | 2026-04-27 |
| Qwen3.5 397B A17B | $0.55 | $3.50 | 262,144 | 2026-02-16 |
| Qwen3.8 2.4T A95B | $2 | $6 | 1,048,576 | 2026-08-12 |
| Qwen3.8 Max (0902) | $2 | $6 | 1,000,000 | 2026-09-03 |
| Qwen3.8 Max Prime | $4 | $12 | 1,000,000 | 2026-09-23 |
Qwen’s current models with a page here, cheapest first. Older and superseded models are listed on the provider page and in the comparison table.
Save
Cheaper alternatives
- Qwen3.5-9BQwen84% cheaper
- Seed 1.6 FlashByteDance Seed82% cheaper
- GPT-6 LunaOpenAI72% cheaper
- Claude Haiku 5.5Anthropic72% cheaper
- GLM 4.6VZ.ai37% cheaper
Current models from major providers, one per provider, that keep image input, tool calling, a reasoning mode, a usable context window and output limit. Compared by blended price.
Compare
Similarly priced models
- Seed-2.0-LiteByteDance Seed$0.25 / $2
- Gemini 3.5 Flash LiteGoogle$0.30 / $2.50
- Nova 2 LiteAmazon$0.30 / $2.50
- Command A+Cohere$0.30 / $1.50
- GLM 5.3 FlashXZ.ai$0.37 / $1.25
Current models from other major providers, closest blended price.
FAQ
Qwen3.5-122B-A10B questions
How much does Qwen3.5-122B-A10B cost?
Qwen3.5-122B-A10B costs $0.26 per million input tokens and $2.08 per million output tokens. A typical chatbot reply (1,500 tokens in, 400 out) costs about $0.00122. Use the LLM cost calculator for your own workload.
What is Qwen3.5-122B-A10B’s context window?
262,144 tokens, about 196,608 English words or 393 printed pages, shared between your prompt and the reply. A single reply can be up to 65,536 tokens.
Does Qwen3.5-122B-A10B support prompt caching and batch requests?
No cached-input price is published for it. No batch price is published for it.
What are cheaper alternatives to Qwen3.5-122B-A10B?
With the same essentials (image input and tool calling), the cheapest options are Qwen3.5-9B (84% cheaper), Seed 1.6 Flash (82% cheaper), GPT-6 Luna (72% cheaper). Cheaper doesn’t mean equivalent, so test them on your own prompts.
What is the API model ID for Qwen3.5-122B-A10B?
On OpenRouter it is qwen/qwen3.5-122b-a10b. Qwen’s own API may use a different name, so check its model list before you deploy. Pass the ID as the model value in each request.
Is Qwen3.5-122B-A10B open source?
Yes, Qwen3.5-122B-A10B has open weights, so you can run it yourself or choose between several hosting providers. The prices here are a typical hosted price.