Daily Digest - 2026-09-25

A look at how OpenRouter scaled to route ten trillion tokens across models, leading to its seven-billion-dollar acquisition by Stripe.

Daily Digest - 2026-09-24

Hamel Husain explains why teams building AI products need to hunt for actual errors before inventing metrics to measure them. Most engineering teams write quantitative benchmarks before looking at real user traces, so they end up scoring the wrong failure modes.

Daily Digest - 2026-09-23

Anthropic released Claude Opus 5. 5 as both Anthropic and OpenAI cut model prices by 40 to 50 percent, resetting baseline costs for frontier capabilities and multi-agent systems.

Daily Digest - 2026-09-22

Hamel Husain and Shreya Shankar explain why writing metrics too early causes AI product teams to fail, and share a practical workflow for finding and fixing hidden failures with coding agents.

Daily Digest - 2026-09-21

Former InstructGPT co-author Diogo Almeida is building fast, calibrated "System One" decision models designed specifically for software automation instead of standard autoregressive LLMs.

Daily Digest - 2026-09-20

Shopify founder Tobi Lütke argues that MCP-versus-CLI is the wrong frame. What matters is whether agents can work in persistent execution environments.

Daily Digest - 2026-09-19

Looking at more than 43,000 model calls shows that fragmented and cut-down inference setups, not changed weights, are why production models lag behind benchmark claims.

Daily Digest - 2026-09-18

Ion Stoica explains why autonomous coding agents learn to game benchmarks and fail in production when written prompts and test environments do not match real-world requirements.

Choose your reading rhythm.

Start with a weekly briefing, add daily notes, or hear only when a durable essay or research update is ready.

You've successfully subscribed to Antoine Buteau
You've successfully subscribed to Antoine Buteau
Welcome back! You've successfully signed in.