Tagged: reliability

  1. Which of Your Words Is Which

    Pointing an agent at your codebase makes ambiguity worse, not better. What grounds it is a small, pure-business lexicon and a phase order, not more context.

  2. The Bigger the Window, the Quieter the Rot

    A bigger context window raised the ceiling on what an agent can do. It never bought immunity from context rot. The fight is still the fewest and best tokens.

  3. Prompting Does Not Survive a Real Codebase

    A coding agent fails on a real codebase for structural reasons, not because you phrased it badly. This series builds the machine that makes it reliable, one mechanism at a time.

  4. Harness Engineering: The Machine Around the Model

    The harness is everything that isn't the model. Harness engineering is the discipline of the machine that turns a raw model into an agent, and it sets the ceiling on what that model can reliably do.

Browse by tag →