Qwen: Qwen3.5-Flash
QWEN Developer Architecture Profile
Intelligence (ELO)1429Aider: 40.6%
Max Context Limit1,000,000Max Out: 65,536 tokens
Prompt CachingStandardNo Cache Discount
Standard API / 1M$0.33Blended Standard Rate
Model Capabilities & Contract SLAs
- Drafting
- Classification
- Vision
- Agentic
- Top Tier
- Coding & Logic
- Fictional
- 🛠️ Native Tool Calling
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...
Granular Pricing Matrix
Input Tokens (Standard Prompt)$0.07 / 1M
Output Tokens (Completion)$0.26 / 1M
Pricing data via OpenRouter. Sync: 10/2/2026
Evaluate Competitors
VS Engine MatchupQwen: Qwen3.5-Flash vs inclusionAI: Ling 3.0 Flash VLVS Engine MatchupQwen: Qwen3.5-Flash vs Nex AGI: Nex-N2.5-MiniVS Engine MatchupQwen: Qwen3.5-Flash vs SpaceXAI: Grok 4.3VS Engine MatchupQwen: Qwen3.5-Flash vs Z.ai: GLM 4.6VVS Engine MatchupQwen: Qwen3.5-Flash vs Mistral: Ministral 3 8B 2512VS Engine MatchupQwen: Qwen3.5-Flash vs Mistral: Mistral Large 3 2512 (batch)