Posts with tag multi-agent

Claude Fable 5 launched this week — record benchmarks, twice the price of Opus. My production pipeline's config hasn't changed, and that's not neglect. Benchmarks measure raw models. Pipelines run calibrated ones.

A comment under my last post made me rethink where AI ends and static analysis begins. Here is how I draw the line — based on a system I run in production.

Adding a new agent to my system takes 2 days. Not because of the architecture — but because teaching it to think like me is the hard part.

Running the most powerful model on every task is wasteful and slow. Here is how I split work across Haiku, Sonnet, and Opus — and why it matters at scale.

No Python, no Node.js, no custom plugins. Just ~2,900 lines of prompt engineering and a fan-out/fan-in architecture that actually catches real bugs.