Topic

pricing

8
Pieces
AUG 16, 2026
Last filed
Tagged pricing clear ×
AUG 16, 2026
Business deep
DeepSeek raised prices 1,100% and nobody flinched
DeepSeek's peak/off-peak billing takes effect today: V4-Pro output up 4.5x at peak, cache-hit input up as much as 12x. It lands a fortnight after OpenAI cut Luna 80%. The price war was never a race to zero — it is convergence toward the true marginal cost of intelligence, approached from both directions at once.
19 MIN 18 src
AUG 14, 2026
Models deep
Post-training is the new pre-training
Grok 4.6, Gemini 3.7 Flash, DeepSeek V4-Pro-0813, and GLM-5.3 shipped within three days of each other — and not one of them is a new base model. The frontier is advancing by post-training, harness, and serving efficiency while every next-generation base stays in the oven.
16 MIN 23 src
JUL 30, 2026
Models deep
The model that cut its own price
OpenAI cut GPT-5.6 Luna by 80% three weeks after launch and said the cut was funded by Sol rewriting OpenAI's own production kernels. The coverage called it a blink at China. The better story is who did the optimizing — and what it means when a model becomes a line item on both sides of its own P&L.
18 MIN 15 src
JUL 28, 2026
Business deep
The margin is the message
In twelve days, Chinese labs pushed the capability floor toward zero, Anthropic halved the price of the frontier, and Google cut Flash tokens again. Read together, the fortnight says one thing: the capability premium is over, and margin is the only battleground left. A ledger of who pays.
16 MIN 11 src
JUL 24, 2026
Models deep
Half the price of frontier
Anthropic shipped Claude Opus 5 as near-frontier intelligence at half the cost of Fable 5. The launch reads as good news for buyers. Read from the P&L, it is a margin admission — the clearest signal yet that the race has moved from raw capability to cost-per-task, and that nobody expects customers to pay the old premium for the top of the curve.
12 MIN 6 src
JUL 21, 2026
Models deep
Google shipped three Flash models the day its flagship went missing
Google's July 21 drop cut Flash token prices 17% and added a security-tuned Cyber variant. The tell is the model that didn't ship: Gemini 3.5 Pro, promised for June, is still partner-only while compute moves to Gemini 4.
12 MIN 12 src
JUL 08, 2026
Models deep
Three labs shipped. One asked permission.
In 72 hours in early July, SpaceXAI shipped Grok 4.5, Meta shipped Muse Image and Muse Spark 1.1, and OpenAI shipped GPT-Live — while GPT-5.6 Sol sat behind a government review. The gate that dominated the headlines applied to exactly one capability class. The real competitive frontier, agentic coding at $2 a million tokens, routed straight around it.
6 MIN 10 src
JUL 03, 2026
Models news
Google just priced the generative-media floor: $0.034 an image
Google shipped Nano Banana 2 Lite at $0.034 per image and Gemini Omni Flash at $0.10 per second of video on the same day. The news is not the models. It is that Google set a public floor price for high-volume generative media, and the floor is now low enough to change who can afford to run it.
2 MIN 5 src