GPT-4o (2024-08-06)
OPENAI Developer Architecture Profile
Intelligence (ELO)1507Aider: 57.1%
Max Context Limit128,000Max Out: 16,384 tokens
Prompt Caching50% OFF$1.25 / 1M cached
Standard API / 1M$12.50Blended Standard Rate
Model Capabilities & Contract SLAs
- Coding & Logic
- Fictional
- Conversational
- Agentic
- 📋 Strict JSON Schema
- 🛠️ Native Tool Calling
The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...
Granular Pricing Matrix
Input Tokens (Standard Prompt)$2.50 / 1M
⚡ Cached Input (Prompt Cache Read)50% OFF$1.25 / 1M
Output Tokens (Completion)$10.00 / 1M
Pricing data via OpenRouter. Sync: 10/2/2026
Evaluate Competitors
VS Engine MatchupGPT-4o (2024-08-06) vs OpenAI: GPT Astra LatestVS Engine MatchupGPT-4o (2024-08-06) vs Sakana: Fugu Ultra v2VS Engine MatchupGPT-4o (2024-08-06) vs Fireworks: Ember-1VS Engine MatchupGPT-4o (2024-08-06) vs Google: Gemini 3.8 Flash (batch)VS Engine MatchupGPT-4o (2024-08-06) vs Google: Gemini 3.5 Flash Lite (batch)VS Engine MatchupGPT-4o (2024-08-06) vs OpenAI: GPT-5.6 Terra Pro (batch)