// blog
Field notes from the local AI frontier
- Self-Hosted vs API: The Real Cost Breakdown Nobody Shows You
Comparing self-hosted LLM costs to commercial APIs honestly: break-even math, hidden costs on both sides, and why hybrid pipelines win for most teams.
- How Much GPU Do You Actually Need to Self-Host an LLM?
A practical VRAM sizing guide for running large language models locally: model size, quantization, context length, and concurrency — with real numbers instead of guesswork.