Antoine Buteau

About Antoine

Daily Digest - 2026-09-20

Shopify founder Tobi Lütke argues that MCP-versus-CLI is the wrong frame. What matters is whether agents can work in persistent execution environments.

Daily Digest - 2026-09-19

Looking at more than 43,000 model calls shows that fragmented and cut-down inference setups, not changed weights, are why production models lag behind benchmark claims.

Daily Digest - 2026-09-18

Ion Stoica explains why autonomous coding agents learn to game benchmarks and fail in production when written prompts and test environments do not match real-world requirements.

Daily Digest - 2026-09-17

Custom Jev-style models using parallel constrained decoding could make existing agent workflows far more token-efficient. Consider an agent workflow where an LLM reviews every support ticket, invoice, or claim before the next step.

Daily Digest - 2026-09-16

Surge AI shows how training Kimi K2. 7 entirely with reinforcement learning across 1,700 coding tasks improves efficiency and transfers across benchmarks without using supervised fine-tuning.

Daily Digest - 2026-09-15

How OpenAI shifted toward an autonomous software factory model driven largely by non-engineers. Non-engineering teams at OpenAI, including legal, recruiting, and finance, now rely on Codex and ChatGPT Work as their primary day-to-day tools.

Daily Digest - 2026-09-14

How to turn custom customer projects into core product features rather than one-off consulting jobs. Forward-deployed engineers are not just technical consultants.

Daily Digest - 2026-09-13

See how giving coding agents domain-specific evaluation skills lets them audit, diagnose, and benchmark AI applications on their own.

Choose your reading rhythm.

Start with a weekly briefing, add daily notes, or hear only when a durable essay or research update is ready.

You've successfully subscribed to Antoine Buteau
You've successfully subscribed to Antoine Buteau
Welcome back! You've successfully signed in.