Qwen: Qwen3.8 Omni Flash
QWEN Developer Architecture Profile
Intelligence (ELO)1424Chatbot Arena Verified
Max Context Limit1,000,000Max Out: 131,072 tokens
Prompt Caching89% OFF$0.02 / 1M cached
Standard API / 1M$0.62🧠 Deep Reasoner
Model Capabilities & Contract SLAs
- Drafting
- Classification
- Audio Gen
- Reasoning
- Agentic
- Coding & Logic
- Fictional
- 🛠️ Native Tool Calling
- 🧠 Extended Test-Time Reasoning
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
Granular Pricing Matrix
Input Tokens (Standard Prompt)$0.15 / 1M
⚡ Cached Input (Prompt Cache Read)89% OFF$0.02 / 1M
Output Tokens (Completion)$0.47 / 1M
Pricing data via OpenRouter. Sync: 10/2/2026
Evaluate Competitors
VS Engine MatchupQwen: Qwen3.8 Omni Flash vs Qwen: Qwen3.6 35B A3BVS Engine MatchupQwen: Qwen3.8 Omni Flash vs Qwen: Qwen3 MaxVS Engine MatchupQwen: Qwen3.8 Omni Flash vs Qwen: Qwen3 Coder FlashVS Engine MatchupQwen: Qwen3.8 Omni Flash vs Z.ai: GLM 4.5VS Engine MatchupQwen: Qwen3.8 Omni Flash vs Meta: Llama 3.2 3B InstructVS Engine MatchupQwen: Qwen3.8 Omni Flash vs Z.ai: GLM Flash Latest