📅 Week of July 4, 2026

API Pricing Digest

What changed in AI API pricing — and All tools are now free.

🚨 TL;DR — Final Week

All Tools Are Free

Price monitoring for 88 models, migration code, cost dashboard — all free, no signup required.

Free Tools →
🔒 No signup required · Instant access

🆕 New Models This Week

New

GPT-5.4 mini — OpenAI's new value champion

OpenAI launched GPT-5.4 mini at $0.75 input / $4.50 output per 1M tokens. Priced between GPT-4o mini and GPT-4o, it delivers significantly better reasoning than its predecessor while remaining OpenAI's most cost-effective new-gen model. For chatbots, content generation, and data extraction that need stronger reasoning than GPT-4o mini could offer.

New mid-budget option · View pricing & alternatives →
New

GPT-5.4 Pro — OpenAI's premium reasoning model

The new flagship from OpenAI. GPT-5.4 Pro targets complex reasoning, code generation, and agentic workflows at $30.00 input / $180.00 output per 1M tokens. Premium pricing positions it alongside Claude Opus 4.8 and GPT-5.5 Pro for the most demanding workloads.

New

Gemini 3.1 Flash-Lite — Google's ultra-cheap option

Google's answer to the price war. Gemini 3.1 Flash-Lite at $0.25 input / $1.50 output per 1M tokens is one of the cheapest multimodal models from a major provider. Perfect for high-volume tasks like classification, routing, and simple Q&A where cost matters most.

▼ Budget tier · View pricing & alternatives →
New

DeepSeek V4 Flash — cheapest reasoning-capable model

DeepSeek's latest at $0.14 input / $0.28 output per 1M tokens. Remarkably cheap for a model with strong reasoning capabilities. If you're using GPT-4o for tasks that don't need its full capability, DeepSeek V4 Flash could cut your costs by 90%+.

▼ Budget tier · View pricing & alternatives →
New

GPT-oss 120B — OpenAI's open-weight budget play

OpenAI enters the budget open-weight space. GPT-oss 120B at $0.15 input / $0.60 output per 1M tokens. Same price as GPT-4o mini but with a different architecture trade-off. Worth benchmarking against your current model for classification and extraction tasks.

▼ Budget tier · View pricing & alternatives →

📉 Price Drops

Price Drop

Grok 4.3 — xAI's massive 83% output price cut

xAI rebranded Grok 3 → Grok 4.3 and slashed pricing from $3.00/$15.00 to $1.25/$2.50 per 1M tokens. That's an 83% output price cut. At $2.50/1M output, Grok 4.3 is now cheaper than Claude Haiku ($5) and competitive with Gemini 3 Flash ($3) for tasks that need solid reasoning. If you dismissed xAI's pricing before, it's time to take another look.

▼ 83% output price cut · View pricing & alternatives →

⚠️ Deprecations & Retirements

Deprecation

7 models retired this cycle — check your endpoints

Major cleanup across providers. Anthropic: Claude 4 Opus → Opus 4.8, Sonnet 4.6 → Sonnet 5, Sonnet 4 → Sonnet 4.6. Google: Gemini 2.0 Flash → 3 Flash, Gemini 2.0 Flash Lite → 3.1 Flash-Lite. DeepSeek: V3 → V4 Flash. AI21: Jamba 1.5 → 1.7. If your code references any of these endpoints, you need to migrate. APIpulse's audit tool now warns about deprecated models and shows migration paths.

Action required if using deprecated endpoints · Audit your current model →

📊 Pricing Trends

Trend

The price floor keeps dropping

Twelve months ago, the cheapest API model from a major provider was ~$0.15/1M input tokens. Today, GPT-oss 20B is at $0.08, Mistral Small at $0.10, and Gemini 2.5 Flash-Lite at $0.10. The "good enough for most tasks" price has nearly halved in a year. If you locked in pricing assumptions 6 months ago, you're overpaying.

▼ ~50% YoY on budget models · Full trend analysis →
Trend

Premium models holding steady — for now

While budget models race to the bottom, premium-tier pricing (GPT-5.4 Pro, Claude Opus 4.8, GPT-5.5 Pro) remains at $5-30 input / $25-180 output per 1M tokens. The gap between "cheap" and "best" is now 30-600x. This creates a clear optimization opportunity: route simple tasks to cheap models, reserve premium for complex reasoning.

Are you on a deprecated model? Are you overpaying?

Run a free 30-second audit of your current API model. See exactly how much you could save — and get migration code if you need to switch.

Run My Free Audit →
Takes 30 seconds · No signup required · See savings instantly

📬 Don't miss next week's changes

Get the API Pricing Digest delivered every Friday. No spam — just pricing intelligence.

Free. Unsubscribe anytime. We respect your inbox.

📚 Past Digests

Browse full archive →

✨ All tools are free

Stop checking 88 models manually

APIpulse monitors every model across 10 providers. Get alerts when prices drop, migration code ready to paste, and a cost dashboard to track savings. No signup required — everything is free.

Free Tools →
No signup required · Instant access