Local AI & builder infrastructure
Ollama, OpenWebUI, SearXNG, local models, homelabs, GPUs, self-hosted tools, and the tradeoffs between local control and cloud convenience.
- 72 GB of VRAM, No Graphics Card
A mini-PC assigned 72 GB of system RAM to its iGPU and ran a 120B model, but memory bandwidth limited generation to reading speed.
- Why I Self-Host
Privacy, cost, and resilience matter, but the honest reason I self-host is enjoying the work—even when it means drivers and firewall lockouts.
- Why I Generated Slower Speech Instead of Stretching Audio
Fresh speech synthesis preserved sharper consonants than phase-vocoder stretching in 27 of 30 pairs, but extreme slowdown introduced new artifacts.
- Building a Local LLM Box That Doesn't Need Babysitting
A field note on assembling a small inference machine, putting Ubuntu Server on it, and getting a stack that runs stable for daily work without constant firefighting.