Blog
Pricing updates, launch news, and explainers on GPT-6, GPT-5.6, Claude, Gemini, and DeepSeek API costs — verified against official sources, not social media leaks or guesses.
Prompt Caching Pricing Compared: OpenAI, Claude, Gemini, DeepSeek
Cache-read discounts look similar across providers. What actually moves your bill is the write fee, the default behavior, and Google's hourly storage rent.
✓ Confirmed LaunchClaude Opus 5.5 Is Live: 20% Cheaper and Now the Default Model
$4/$20 per million tokens, replacing Opus 5 as the default. The cache-read discount dropped a lot more than the headline rate did.
◆ Launch newsGPT-6 Sol and Luna Are Live, and Luna Is Now the Cheapest Model We Track
GPT-6 Luna undercuts DeepSeek V4.1-Flash on price and has a Batch API DeepSeek doesn't. Full rate card for the now-complete GPT-6 family.
◆ ExplainerHow Much Does a RAG App Actually Cost to Run?
A RAG app doesn't have one bill, it has four. Where the money actually goes, and why retrieved context is the hidden multiplier.
◆ ComparisonGPT-6 vs Claude vs Gemini vs DeepSeek: Full API Pricing Comparison
Updated Sep 24: Claude Opus 5.5 added — it now ties GPT-5.6 Sol exactly on headline rate, a first for the flagship tier.
◆ ExplainerWhat Is Peak/Off-Peak LLM API Pricing? A Practical Guide
DeepSeek now bills by time of day. How the mechanic works, whether it affects you, and how to schedule batch jobs around it, including the weekday-and-holiday detail.
◆ ExplainerClaude Fable 5.1's Real Price Cut Isn't the Headline Rate, It's Cache Reads
Input and output are unchanged. Cache reads dropped 75%. Updated Sep 24: Opus 5.5 now undercuts Fable 5.1's absolute cache-read price.
✓ Confirmed LaunchGPT-6 Astra Is Live: $10/$50 Pricing, and What You're Actually Paying For
OpenAI's new flagship costs 2.5x GPT-5.6 Sol. Updated Sep 23: the GPT-6 family is now complete with Sol and Luna below it.
⚠ Pricing AlertDeepSeek Raised API Prices Up to 1,100% on August 16: The Confirmed Rate Card
Updated Sep 22: V4-Flash was renamed V4.1-Flash and cut again, partially reversing the August hike.
⏱ Pricing UpdateOpenAI Cuts GPT-5.6 Sol Prices by Over 20%, Undercutting Claude Opus 5
Input down 20%, output down 33%. Updated Sep 24: Claude Opus 5.5 has since launched at the exact same $4/$20 rate.
◇ Beginner's GuideI Want to Build an AI App. What Will the API Actually Cost Me?
A plain-terms explanation of how AI API billing works, written for people building their first project.
◆ Explainer5 Ways to Cut Your LLM API Bill Without Changing Models
Batching, caching, context trimming, routing, and reasoning effort control: five levers that cut cost with zero model-quality risk.
✓ Confirmed LaunchGemini 3.7 Flash Launches at Half the Price of 3.6 Flash — But Only Until December 31
The leaked $0.75/$3.75 pricing was real. Google confirmed it at launch — as an introductory rate with a firm expiration date, not a permanent cut.
⏱ Pricing UpdateClaude Sonnet 5's Price Increase Is Canceled: Here's What Actually Changed
Anthropic confirmed the scheduled Sept 1 hike won't happen. The $2/$10 rate is now permanent — but that's not the whole story for your bill.
◐ Pricing UnknownGPT-5.6 Sol's New Ultrafast Mode: What We Know About Pricing (and What We Don't)
OpenAI announced Ultrafast on August 13 — up to 14x faster than Standard, powered by Cerebras. The announcement says nothing about cost.
◆ ExplainerWhat Are Reasoning Tokens, and Why Do They Bill as Output?
If a reasoning-capable model's bill looks higher than the visible response justifies, this is usually why — and how to actually estimate it.