Blog

Notes from real runs.

What self-hosted LLM inference actually costs, why the number moves when nothing changed, and how to tell a real saving from noise. Every figure comes from a recorded run.