#evals

2 posts

  1. 3 min read

    Routing between LLM providers to balance cost and quality

    Why we route requests across more than one model provider, how prompt caching cuts repeat costs, and why routing changes pass offline evals first.

  2. 5 min read

    LLM evals for small teams: a practical starting point

    How to build LLM evals with a small team: start from real failures, use pass/fail checks, calibrate an LLM judge and gate every prompt change.