← Back to Directory

Qwen: Qwen3.8 2.4T A95B

QWEN Developer Architecture Profile

Intelligence (ELO)1433Chatbot Arena Verified
Max Context Limit1,048,576Max Out: 131,072 tokens
Prompt Caching88% OFF$0.25 / 1M cached
Standard API / 1M$8.00Blended Standard Rate

Model Capabilities & Contract SLAs

  • Agentic
  • Coding & Logic
  • Fictional
  • 🛠️ Native Tool Calling
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Granular Pricing Matrix

Input Tokens (Standard Prompt)$2.00 / 1M
⚡ Cached Input (Prompt Cache Read)88% OFF$0.25 / 1M
Output Tokens (Completion)$6.00 / 1M

Pricing data via OpenRouter. Sync: 10/2/2026

Evaluate Competitors