The Journey
Field reports from building with AI in public. What's working, what isn't, and what it cost me. Open numbers, the failures before the wins.
Subscribe via RSS
"80% of AI projects fail" traces to a footnote one line long
The most repeated number in enterprise AI turns out to be a paraphrase of unnamed opinion surveys, quoted in a magazine profile of a vendor.

Why AI projects fail: 19 headline statistics, and what each one actually counts
Nineteen headline AI failure statistics, traced to source. Seven are not measurements of anything that happened.

I audited 26 risk controls in my trading bot: 7 could never fire
The 15% drawdown halt was wired into two live code paths and could never fire. So I checked all 26 risk controls. Seven were dead. A control you cannot observe firing is not a control, it is a belief.

AI agents escaped containment: 17,600 actions, and the alert that fired too quietly
Three incidents, one headline, three very different severities. What the primary reports actually say, and what to do about it.

Own your AI skill files, or your job becomes one
A skill file is your judgement, written down and executable. Who owns the repo decides whether that's an asset or an extraction.

Soft Systems Methodology explained: Checkland's 1969 method for AI briefs nobody agrees on
Most AI projects fail on the brief, not the model: six people nod at one sentence and mean four different projects. Soft Systems Methodology, built at Lancaster from 1969 for exactly that, in plain English: rich pictures, root definitions, CATWOE, and how to run it on the next thing on your roadmap.

I swapped the model and nothing happened
Swapping the two LLMs behind an autonomous trading bot from Opus 4.7 to 4.8, live, without it missing a beat. The trick was replaying 78 real prompts first.

80% of AI projects fail on the problem, not the model: Soft Systems Methodology
More than 80% of AI projects fail, and RAND's leading root cause is not the technology: nobody agreed what problem was being solved. Soft Systems Methodology, built at Lancaster from 1969 for exactly that failure, gives you a way to surface the disagreement in an afternoon instead of a year.

Opus 5 verbosity: where I wanted three sentences, I got Proust
I judge AI models by how often they make me swear. Opus 5 turned swearing into punctuation, and the reason is more interesting than I first thought.

Why AI checks fail silently: 1,278 restarts behind a healthy status
My auth check said healthy for three weeks on a dead credential. The pattern behind it costs more than the bug did.

Anthropic cut 80% of Claude Code's prompt. I cut my CLAUDE.md by 90% the same day.
An audit of 116 instruction files found one stale copy telling twenty repos they were a different project, two files no session has ever read, and 453 lines loaded twice. Here is what to delete from your own CLAUDE.md, and the method that beats judgement.

Bernardo Kastrup's cosmic mind has no plan: what analytic idealism actually claims
The most rigorous idealist alive gives the non-duality audience far less than they want: a universal consciousness that is instinctive, has no plan, and is not looking back at you. What he actually claims, why he says it cannot be proven, and the one part that survives whoever turns out to be right.

My AI Agents Kept Losing My Work. The Fix Is Older Than Computers.
Seventeen finished pieces of work were invisible to my own tracking system, and the cause wasn't effort or memory. A Berkeley talk gave me the name for what fixed it: an ontology. Here's the version that fits a one-person company.

The Data Said Nobody's Buying AI. Turns Out I Was Selling the Wrong Thing.
Last week I showed the demand numbers for AI consultancy are dire. This week a veteran operator explained the part I missed: the demand isn't dead, it's mute. Nobody buys AI. They buy outcomes, told concretely enough to defend to a board.

Garry Tan Just Described My Setup. Then I Noticed What He Left Out.
The YC president gave a talk describing, almost item for item, the AI system I already run. The validation lasted about ten minutes. Then I noticed the word neither of us was saying.

I Leaked an API Key and It Turned My Own Domain Into a Phishing Weapon
I got phished from my own verified domain, using an API key I'd leaked myself: what actually leaked, how fast it happened, and what I changed.

A friend said it was a business. I decided it was not, and built it anyway.
A friend said it should be a business. I spent forty years around operations and resilience, so I assessed it honestly, decided it was not, and built it anyway, for free. Here is the reasoning, and why the map of your estate should be a file no company owns.

Everyone's selling AI consultancy. The data says nobody's buying.
A $1,000-an-hour AI consultancy playbook is doing the rounds. The demand data, and a friend who sells everything except this, disagree.

A goal prompt beat my build framework on a real job
Same spec, same model, two methods: a seven-part goal prompt against a full agent framework. The lean prompt tied on correctness and won on time, cost, and clutter. Here are the numbers.
