Qwen: Qwen3.5-35B-A3B
qwen/qwen3.5-35b-a3b- Context window
- 262,144 tokens
- Maximum output
- 16,384 tokens
- Input price
- $0.3125 per 1M tokens
- Output price
- $1.25 per 1M tokens
- Modalities
- text, image, video
Last verified · Source ↗
Official description
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
Read the official model reference ↗.
Where this model fits
qwen/qwen3.5-35b-a3b is part of the OpenRouter API catalog. Its documented modalities are text, image, video. Choose it against the request format, output requirements and operational limits of your application.
The source metadata identifies support for router hosting. Confirm endpoint-specific parameters in the official reference before integrating that capability.
For a repeatable evaluation, use representative requests, inspect structured output or tool calls, and log latency and usage in your own account. Catalog metadata establishes compatibility; it does not measure answer quality on your workload. Start from the provider guide for authentication and endpoint setup.
Current pricing
| Model | Price type | USD | Unit | Tier | Region | Source |
|---|---|---|---|---|---|---|
| Qwen: Qwen3.5-35B-A3B | Cached Input | $0.1563 | per 1M tokens | Standard | See source | Official source ↗ |
| Qwen: Qwen3.5-35B-A3B | Input | $0.3125 | per 1M tokens | Standard | See source | Official source ↗ |
| Qwen: Qwen3.5-35B-A3B | Output | $1.25 | per 1M tokens | Standard | See source | Official source ↗ |
Last verified · Source ↗
Estimate request, daily and monthly usage in the API cost calculator.
Rate limits and availability
Published limits have not yet been verified for this selection. Check the official limits documentation and your account console.
Last verified · Source ↗
Account-specific capacity may differ. Check the provider rate-limit guide and the free-tier conditions.
Three alternatives to compare
These alternatives share a documented modality and have the closest combination of input price and context size in the current catalog. This is a specification comparison, not a quality benchmark.
| Model | Provider | Input / 1M tokens | Context |
|---|---|---|---|
| Kwaipilot: KAT-Coder-Pro V2 | OpenRouter | $0.3 | 262,144 |
| Qwen: Qwen3 Coder 480B A35B | OpenRouter | $0.3 | 262,144 |
| Qwen: Qwen3.6 27B | OpenRouter | $0.3 | 262,144 |
Use the exact identifier in Python
This identifier fragment belongs in the authenticated request described in the provider Python tutorial. The tutorial covers the correct SDK, endpoint, environment variable and response shape.
model = "qwen\/qwen3.5-35b-a3b"
# Pass model to the provider-specific request in the tutorial.Recent changes
- OpenRouter · Qwen: Qwen3.5-35B-A3B — Amount UsdNot previously recorded → Price category: cached_input · USD: 0.15625 · Unit: per 1M tokensSource ↗
- OpenRouter · Qwen: Qwen3.5-35B-A3B — Amount UsdNot previously recorded → Price category: output · USD: 1.25 · Unit: per 1M tokensSource ↗
- OpenRouter · Qwen: Qwen3.5-35B-A3B — Amount UsdNot previously recorded → Price category: input · USD: 0.3125 · Unit: per 1M tokensSource ↗
- OpenRouter · Qwen: Qwen3.5-35B-A3B — First SeenNot previously recorded → Model: Qwen: Qwen3.5-35B-A3BSource ↗
Subscribe to the changelog RSS feed
Frequently asked questions
What is the exact API model identifier?
qwen/qwen3.5-35b-a3b exactly as shown in the provider model reference. Display names are not interchangeable with API identifiers.How much does this model cost?
How much input can the model accept?
Can I use a free tier?
Where do these model facts come from?
Sources
Last verified · Source ↗