#evals
2 posts
Routing between LLM providers to balance cost and quality
Why we route requests across more than one model provider, how prompt caching cuts repeat costs, and why routing changes pass offline evals first.
LLM evals for small teams: a practical starting point
How to build LLM evals with a small team: start from real failures, use pass/fail checks, calibrate an LLM judge and gate every prompt change.