From our SaaS audits
What we find inside SaaS companies
Patterns from the technical due diligences we run for investors and boards.
Newsletter
Human-written AI news, tutorials and insights for software teams and CTOs. Sent when there is something worth sharing, not on a schedule.
From our SaaS audits
Patterns from the technical due diligences we run for investors and boards.
AI has a confidence problem. Human intelligence was shaped by an uncertain world, and uncertainty is exactly the ingredient AI is missing.
Agents follow the patterns in the surrounding code more reliably than they follow your rules, so every inconsistent naming scheme and half-migrated architecture is a lesson you're teaching them. The fix isn't another rule file. It's deleting the bad example.
Two agents implemented and reviewed a full architecture refactor across five iterations, with zero lines of production code written by hand. What neither agent caught was four innocent lines in a coordinator that had quietly started making policy decisions.
Codex got the same god class Claude had failed to remove, plus one thing Claude never had: a target architecture. An implementing agent and an independent reviewer then took three rounds to converge, because agents renovate around bad architecture rather than delete it.
Handing an AI your engineering conventions cleans up the code without fixing the architecture. A 5,000-line god class shrank to 500 lines across tidy domain files, and an independent reviewing agent still returned FAIL: authority never left the central object.
Becoming AI-native in three years does not need a three-year technology roadmap. It needs a usage policy, a map of your data, one bounded experiment, and both the enthusiasts and the sceptics in the room. Build the ability to change, not the overhaul.
Hitting a wall with AI usually means hitting the wall of the chat box, not the wall of AI itself. Give a model memory, tool access and a browser, and the interfaces we built for humans become the integration. The bottleneck stops being the model and starts being your imagination.
An AI-assisted note system captures faster than you can understand, and the gap compounds into internalisation debt: a dense graph of links attached to a thin mental model. The test of a knowledge base is not how much it holds, but how much you could still explain with the file closed.
Synthetic checks like curl and Lighthouse measure a request your users never make. A HAR file captures the real logged-in flow, and handed to Claude alongside the codebase it turns a wall of JSON into a ranked list of fixes that get implemented in the same conversation.
A back-end engineer set out to redesign a React app with an AI agent and came out actually knowing React. The unlock wasn't the prompting, it was understanding the code well enough to steer it. As the cost of entry collapses, knowing the technology matters more, not less.
Analytics only shows the events you remembered to add, never the ones you forgot. Let Claude instrument features as it builds them, then run a skill that audits your tracking against your own docs and live data, flagging what's missing, noisy, or quietly broken.
Integrating a tool with Claude rarely needs an MCP server. A boring script wrapped as a named, version-controlled skill does the job, and the real value shows up when two skills combine: one tool tells you what happened, another tells you why.
Writing code by hand used to be the job. Soon it will be a hobby, like woodworking or growing your own veg. What a company pays for now is judgment: knowing what to build, catching the answer that looks right and isn't, deciding when it's good enough to ship.
Handing everyone an LLM subscription gives them a model. A shared, self-hosted agent gives them a configured environment with company context, connected tools, and clear boundaries. Why a company should own that layer rather than rent it, and how to start small.
Coding agents made implementation cheap, but validation never got cheaper. That shift moves the bottleneck from writing code to deciding what is worth building, and the response is more autonomy and faster iteration, not more detailed upfront plans.