Blog

  • The window is real. So is the graveyard.

    The lab is running.

    The first post set the scene. Real products. Real users. The agentic experiment, done properly.

    This is the part that actually matters.

    The part I’m chasing

    My projects are teaching me things. Not just whether agents work. Something quieter.

    What happens when you run the same work the same way, over and over. Where it holds. Where it cracks. Which pieces are luck, and which pieces are structure.

    I’m trying to separate the two.

    What the lab has already learned

    Some of it is unglamorous. That’s why it’s useful.

    A rejection got merged anyway. Once. It won’t happen again.

    Reviewers reviewed the same code twice. Nothing had changed. Duplicate work isn’t diligence — it’s a design flaw.

    A scanner read a work order as a handoff. Keywords lie. Structure doesn’t.

    The newest verdict hid behind an old comment list. The loop acted on yesterday’s news.

    The reviewer model went cold, and the loop went blind for nearly an hour.

    A fix claimed a commit that didn’t exist yet. Receipts, or it didn’t happen.

    Every one of those became a rule. Not a lesson. A rule.

    The shape is taking form

    An evidence-backed control plane. Everything verified. Nothing self-certified.

    If it only works because it’s mine, it’s a story. If it works because of how it’s built, it’s something else. Something with edges. Rules. A shape. Something another team could pick up without me in the room.

    That’s the thing I’m trying to bottle.

    The market agrees with the paranoia

    Gartner says over 40% of agentic projects get cancelled by 2027. Escalating cost, unclear value, weak controls.

    Deloitte says agent deployment jumps from a quarter of companies to half. PwC found most executives raising AI budgets anyway.

    The window is real. So is the graveyard.

    The real question

    Would anyone pay for the shape?

    Not for the demo. For the outcome.

    That’s the difference between a framework and a diary. One has a market. The other has an audience.

    I want to know which one I’m building.

    What stays in the lab

    The framework itself. The stages. The numbers that matter most.

    They’re not ready to be seen yet. Some of it is mine. Some of it needs to prove itself first.

    You’ll get the shape of it. The lessons, the failures, the surprises. The receipts.

    Phase two

    The experiment has a second phase. This is it.

    I’ll keep posting as it happens. The wins. The embarrassments. The honest parts.

    Either way, it’s going to be interesting.

  • My Products Are My Guinea Pigs

    Everyone’s talking about agents.

    Most of it is noise. Demos that work in a video. Pipelines that die on day two. Hype pretending to be a market.

    I build products. So I did the obvious thing.

    I made my own projects the guinea pigs.

    The lab is real

    No sandbox. No toy data. Real products, real users, real workflows.

    I’m running the agentic experiment on things that actually ship. If it breaks, something I own breaks. That’s the point.

    It’s the only way to know if any of this holds up outside a keynote.

    What I’m testing

    One question, really. Is there a market here, or just a movement?

    The agentic crowd is loud. But most of it lives underground — early adopters talking to each other. I want to know what happens above ground. With normal teams. Normal budgets. Normal expectations.

    Will people pay for this? Not for the demo. For the outcome.

    What I’m not telling you

    The interesting parts stay in the lab for now.

    Some of it is mine. Some of it is worth protecting. Some of it isn’t ready to be seen.

    You’ll get the shape of it. The lessons, the failures, the surprises. The receipts.

    Follow along

    I’ll post what I learn as I learn it. The wins. The embarrassments. The honest parts.

    If it works, you’ll see it work. If it’s all hype, you’ll watch me say so.

    Either way, it’s going to be interesting.