Ranked by effective cost per token across 88 models and 10 providers. Prices in USD per 1 million tokens. See the most expensive →
Models are ranked by blended cost using a 1:3 input-to-output ratio (typical for chat/assistant workloads). A model costing $0.10/M input + $0.30/M output scores at $1.00/M blended. Context window and provider are noted but don't affect ranking. All prices sourced directly from provider pricing pages.
Price trends: ↓ Dropped = price decreased recently · → Stable = no change in past 30 days · ↑ Increased = price went up. Trends reflect changes since May 2026.
DeepSeek's speed-optimized model. 1M context at the lowest blended cost in the industry.
OpenAI's open-source 20B model. Cheapest input price of any model at $0.08/M.
Google's budget model. 1M context at $0.10/M input — best value for long-context tasks.
Meta's efficient 8B-parameter model. Equal input/output pricing at $0.14/M — great for high-volume tasks.
AI21's hybrid model. 256K context with low output pricing at $0.40/M.
OpenAI's nano-tier model. 1M context at budget pricing — ideal for bulk processing.
DeepSeek's general-purpose model. Excellent quality-to-price ratio.
OpenAI's larger open-source model. More capable than the 20B at still-reasonable pricing.
Mistral's efficient small model. Competitive pricing with 128K context.
Meta's next-gen efficient model. 1M context at $0.18/M input — excellent for long documents.
Use our free calculator to compare all 88 models and find the cheapest option for your specific usage pattern.