Gemini 3.1 Flash Lite
GOOGLE Developer Architecture Profile
Intelligence (ELO)1493Chatbot Arena Verified
Max Context Limit1,048,576Max Out: 65,536 tokens
Prompt Caching90% OFF$0.02 / 1M cached
Standard API / 1M$1.75Blended Standard Rate
Model Capabilities & Contract SLAs
- Fictional
- Drafting
- Classification
- Conversational
- Audio Gen
- Agentic
- Coding & Logic
- 📋 Strict JSON Schema
- 🛠️ Native Tool Calling
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
Granular Pricing Matrix
Input Tokens (Standard Prompt)$0.25 / 1M
⚡ Cached Input (Prompt Cache Read)90% OFF$0.02 / 1M
Output Tokens (Completion)$1.50 / 1M
Internal Reasoning Tokens$1.50 / 1M
Pricing data via OpenRouter. Sync: 10/2/2026
Evaluate Competitors
VS Engine MatchupGemini 3.1 Flash Lite vs Google: Gemini 3.5 FlashVS Engine MatchupGemini 3.1 Flash Lite vs Google: Gemini 3.5 Flash (batch)VS Engine MatchupGemini 3.1 Flash Lite vs Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)VS Engine MatchupGemini 3.1 Flash Lite vs Google: Gemini 3.1 Pro Preview Custom ToolsVS Engine MatchupGemini 3.1 Flash Lite vs OpenAI: GPT-4.1 MiniVS Engine MatchupGemini 3.1 Flash Lite vs Qwen: Qwen3.8 Max Prime