Qwen: Qwen3.8 Flash
QWEN Developer Architecture Profile
Intelligence (ELO)1421Chatbot Arena Verified
Max Context Limit1,000,000Max Out: 131,072 tokens
Prompt Caching89% OFF$0.02 / 1M cached
Standard API / 1M$0.62🧠 Deep Reasoner
Model Capabilities & Contract SLAs
- Drafting
- Classification
- Reasoning
- Agentic
- Coding & Logic
- Fictional
- 🛠️ Native Tool Calling
- 🧠 Extended Test-Time Reasoning
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Granular Pricing Matrix
Input Tokens (Standard Prompt)$0.15 / 1M
⚡ Cached Input (Prompt Cache Read)89% OFF$0.02 / 1M
Output Tokens (Completion)$0.47 / 1M
Pricing data via OpenRouter. Sync: 10/2/2026
Evaluate Competitors
VS Engine MatchupQwen: Qwen3.8 Flash vs Z.ai: GLM 5.2VS Engine MatchupQwen: Qwen3.8 Flash vs Qwen: Qwen3 Coder PlusVS Engine MatchupQwen: Qwen3.8 Flash vs Mistral: Mistral Medium 3.1 (batch)VS Engine MatchupQwen: Qwen3.8 Flash vs Qwen: Qwen3 Coder 480B A35BVS Engine MatchupQwen: Qwen3.8 Flash vs MythoMax 13BVS Engine MatchupQwen: Qwen3.8 Flash vs PrismML: Ternary Bonsai 2 27B