Nine Failures, Zero Alerts: I Audited My Own AI Agent Fleet
Most ways of proving you can audit a system are unfalsifiable. A sample report about an invented company shows you can write like an…
Most ways of proving you can audit a system are unfalsifiable. A sample report about an invented company shows you can write like an…
For the last year, if you were building anything serious with agents, you spent most of your time building everything that was not the…
One of my OpenClaw agents spent about ten weeks writing thoughtful weekly summaries of a feed that had gone completely empty. Every week it…
I wrote recently that agentic commerce takes the observer out of the buying journey: the human stops being the one watching the path from…
Picture the choice a consumer-research team faces this year. On one side, a real panel: weeks of fieldwork, incentive costs, recruiting, the slow grind…
A brand's growth team can usually tell you, close to the decimal, how a customer traveled from an ad to a purchase. Which impression,…
The bill arrived, and it changed the conversation about agents. EY illustrates the shift with a single line: a chat that cost about four…
Salesforce put Multi-Agent Orchestration into its Summer '26 release. That matters less because Salesforce did it, and more because Salesforce sits inside a very…
Most teams still evaluate coding agents with one question. Can it write the code? That question made sense when the product was autocomplete. You…
On June 10, 2026, OpenRouter announced Advisor, a server-side tool that lets one model consult another model mid-generation. The obvious reading is cost optimization.…