Topic

coding-agents

12
Pieces
AUG 21, 2026
Last filed
Tagged coding-agents clear ×
AUG 21, 2026
Agents deep
The infra checklist went viral again. Its readers fed it to their agents
A 110-item infrastructure gatekeeping list pulled 294K views on X last week. It is a verbatim copypasta of a July post that got 47 likes, its 16 headline items grade out at 7 absorbed and 2 still-biting against dated platform changelogs, and its own audience is feeding it to coding agents as a prompt.
17 MIN 18 src
JUL 28, 2026
Business deep
The margin is the message
In twelve days, Chinese labs pushed the capability floor toward zero, Anthropic halved the price of the frontier, and Google cut Flash tokens again. Read together, the fortnight says one thing: the capability premium is over, and margin is the only battleground left. A ledger of who pays.
16 MIN 11 src
JUL 27, 2026
Business deep
Emergent is a $1.5B unicorn in six months. What are we pricing?
An Indian vibe-coding startup quintupled its valuation to $1.5B in half a year on a real $120M revenue run-rate. The revenue is genuine. The question the round leaves unanswered is what, exactly, the multiple is buying when the model underneath is rented and the interface is copyable.
13 MIN 11 src
JUL 24, 2026
Models deep
Half the price of frontier
Anthropic shipped Claude Opus 5 as near-frontier intelligence at half the cost of Fable 5. The launch reads as good news for buyers. Read from the P&L, it is a margin admission — the clearest signal yet that the race has moved from raw capability to cost-per-task, and that nobody expects customers to pay the old premium for the top of the curve.
12 MIN 6 src
JUL 08, 2026
Models deep
Three labs shipped. One asked permission.
In 72 hours in early July, SpaceXAI shipped Grok 4.5, Meta shipped Muse Image and Muse Spark 1.1, and OpenAI shipped GPT-Live — while GPT-5.6 Sol sat behind a government review. The gate that dominated the headlines applied to exactly one capability class. The real competitive frontier, agentic coding at $2 a million tokens, routed straight around it.
6 MIN 10 src
JUL 04, 2026
Models deep
The gate and the giveaway
While Washington spent June turning frontier model releases into licensed events, a Chinese food-delivery company open-sourced a 1.6-trillion-parameter agentic-coding model under an MIT license — trained start to finish on domestic chips, no Nvidia silicon involved. LongCat-2.0 is the counter-move to the export-control regime, and it's already downloadable worldwide.
8 MIN 7 src
JUL 02, 2026
Agents deep
The Cursor SDK and the year the coding agent stopped being an editor
Notion embedded coding agents into its workspace in a few weeks using the Cursor SDK, so an @mention now plans, builds, tests and opens a PR without an editor in sight. The SDK turns the harness that powers Cursor into a callable runtime, and that is a bigger shift than any model release.
5 MIN 8 src
JUN 23, 2026
Agents deep
Cursor Compile 2026: the vertical stack lands, and a warning on the same stage
At its first Compile conference, Cursor announced a from-scratch model on SpaceX compute, Origin (a GitHub rival), and Cursor Mobile. Its design lead used the stage to warn the result could be slop with working buttons.
23 MIN 11 src
JUN 22, 2026
Agents deep
How much is the harness worth? The year the number got measured
Six weeks after agent harness engineering got named, the empirical question arrived: how much of a coding agent's score is the model, and how much is the scaffold around it. Four 2026 papers put numbers on it — and the numbers are larger than almost anyone guessed.
35 MIN 15 src
JUN 10, 2026
Tools deep
Stop fighting Codex CLI's approval prompts
The Codex CLI config that ends the permission nagging — config.toml profiles, sandbox and approval modes, AGENTS.md, reasoning effort, MCP servers, and the codex exec setup for CI.
24 MIN 12 src
JUN 09, 2026
Tools deep
Stop using OpenCode like a Claude Code clone
The OpenCode config that earns its keep — opencode.json layers, the build/plan agents, permissions, AGENTS.md, MCP, and the bring-your-own-model setup that survives a vendor pulling your access.
24 MIN 12 src
JUN 01, 2026
Agents deep
Grok Build, and the model xAI didn't make
xAI shipped a terminal coding agent, then put a competitor's model inside it that out-codes its own. Read against a $60B SpaceX option on Cursor, the CLI war looks less like four rivals and more like one stack assembling itself.
12 MIN 14 src