Updated for March 2026 · Standardized per 1 Million Tokens
LLM API Pricing Benchmark
Compare input and output token rates, prompt caching discounts, and context limits across the frontier AI providers: OpenAI, Anthropic, Google DeepMind, DeepSeek, and Mistral.
Verified DailyDirect Official Provider PricingReasoning & Vision Benchmarks Included
Sort:
Provider:
| Model & Provider | Input / 1M | Output / 1M | Context Window | Capabilities | Docs |
|---|---|---|---|---|---|
Gemini 1.5 Flash Google | $0.075 Cache: $0.018 | $0.30 | 1M tokens | Multimodal (1M Context) Budget-conscious high-context applications | Official |
Gemini 2.0 Flash Google | $0.10 Cache: $0.025 | $0.40 | 1M tokens | Multimodal (Text/Vision/Audio) Real-time multimodal agents & high volume workflows | Official |
DeepSeek-V3Value DeepSeek | $0.14 Cache: $0.014 | $0.28 | 64k tokens | Text & Code Ultra low-cost general generation & high throughput | Official |
GPT-4o mini OpenAI | $0.15 Cache: $0.075 | $0.60 | 128k tokens | Multimodal (Text + Vision) Fast, affordable multimodal tasks & automation | Official |
Codestral Mistral | $0.30 | $0.90 | 256k tokens | Code & FIM Fill-in-the-middle code completion & repo understanding | Official |
DeepSeek-R1Value DeepSeek | $0.55 Cache: $0.140 | $2.19 | 64k tokens | Reasoning & Math Open reasoning, complex STEM logic & code synthesis | Official |
Llama 3.3 70B Meta | $0.59 | $0.79 | 128k tokens | Text & Reasoning Open weights frontier parity via hosted cloud inference | Official |
Claude 3.5 Haiku Anthropic | $0.80 Cache: $0.080 | $4.00 | 200k tokens | Text & Code Low-latency agents and high-precision parsing | Official |
o3-mini OpenAI | $1.10 Cache: $0.550 | $4.40 | 200k tokens | Reasoning & STEM Fast competitive coding & structured reasoning | Official |
Gemini 1.5 Pro Google | $1.25 Cache: $0.310 | $5.00 | 2M tokens | Multimodal (2M Context) Massive codebases, full video ingestion & long doc retrieval | Official |
Mistral Large 2 Mistral | $2.00 | $6.00 | 128k tokens | Multilingual & Code European enterprise compliance & multilingual reasoning | Official |
GPT-4o OpenAI | $2.50 Cache: $1.250 | $10.00 | 128k tokens | Multimodal (Vision + Audio) Enterprise reasoning, complex multimodality | Official |
Claude 3.5 Sonnet Anthropic | $3.00 Cache: $0.300 | $15.00 | 200k tokens | Multimodal & Agentic Coding Benchmark-leading coding, architectural design | Official |
o1 OpenAI | $15.00 Cache: $7.500 | $60.00 | 200k tokens | Deep Reasoning Maximum reasoning power for academic research & math | Official |
Claude 3 Opus Anthropic | $15.00 Cache: $1.500 | $75.00 | 200k tokens | Multimodal Deep Writing Creative nuances, complex analysis | Official |
