Qwen: Qwen3.8 2.4T A95B
QWEN Developer Architecture Profile
Intelligence (ELO)1433Chatbot Arena Verified
Max Context Limit1,048,576Max Out: 131,072 tokens
Prompt Caching88% OFF$0.25 / 1M cached
Standard API / 1M$8.00Blended Standard Rate
Model Capabilities & Contract SLAs
- Agentic
- Coding & Logic
- Fictional
- 🛠️ Native Tool Calling
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Granular Pricing Matrix
Input Tokens (Standard Prompt)$2.00 / 1M
⚡ Cached Input (Prompt Cache Read)88% OFF$0.25 / 1M
Output Tokens (Completion)$6.00 / 1M
Pricing data via OpenRouter. Sync: 10/2/2026
Evaluate Competitors
VS Engine MatchupQwen: Qwen3.8 2.4T A95B vs OpenAI: GPT-6 LunaVS Engine MatchupQwen: Qwen3.8 2.4T A95B vs OpenAI: GPT-6 Sol ProVS Engine MatchupQwen: Qwen3.8 2.4T A95B vs MoonshotAI: Kimi K2.7 CodeVS Engine MatchupQwen: Qwen3.8 2.4T A95B vs Qwen: Qwen3.7 MaxVS Engine MatchupQwen: Qwen3.8 2.4T A95B vs Google: Gemma 4 31BVS Engine MatchupQwen: Qwen3.8 2.4T A95B vs Arcee AI: Trinity Large Thinking