← Back to Directory

inclusionAI: Ling 3.0 Flash

INCLUSIONAI Developer Architecture Profile

Intelligence (ELO)1435Chatbot Arena Verified
Max Context Limit262,144Max Out: 32,768 tokens
Prompt Caching80% OFF$0.0042 / 1M cached
Standard API / 1M$0.08Blended Standard Rate

Model Capabilities & Contract SLAs

  • Drafting
  • Classification
  • Coding & Logic
  • Fictional
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Granular Pricing Matrix

Input Tokens (Standard Prompt)$0.02 / 1M
⚡ Cached Input (Prompt Cache Read)80% OFF$0.0042 / 1M
Output Tokens (Completion)$0.06 / 1M

Pricing data via OpenRouter. Sync: 10/2/2026

Evaluate Competitors