Skip to content

    Guides

    The whole AI cost playbook, published in full

    None of our method is proprietary. These guides cover every lever we use on client work, in the order we use it, including the ones that make an engagement unnecessary. If you can ship it yourself in a fortnight, you should.

    Cost engineeringAI FinOpsToken optimisationModel strategyAgentsAzure
    Start here · Cost engineering

    How to reduce LLM API costs: twelve levers, ranked by risk

    The complete playbook, in the order we actually apply it: the changes that cannot affect your output, then the ones that can — and how to tell the difference before you ship.

    Read the guide

    Why we publish this

    Because the method is not the moat.

    Every technique on this site is documented somewhere in a provider's own guidance. The hard part was never knowing that prompt caching exists — it is finding, in your specific estate, which of a dozen levers is worth the effort, sizing each one honestly, and proving the risky ones before a customer is affected.

    If reading these means you fix it yourself and never speak to us, that is a good outcome. If reading them means you would rather someone did it accurately in two weeks with the measurement left behind, that is what the audit is for.

    Find out what your AI actually costs per completed task.

    Two weeks, a fixed fee, and a ranked savings plan with the quality risk of every move stated up front. If the numbers say an audit is not worth it for you, we will say so on the first call.