- Posted on
- Featured Image
A Linux-first guide to local AI that’s private, fast, and cost-predictable: set up Ollama for instant LLMs, compile llama.cpp for lean performance (CPU/GPU), add whisper.cpp for offline speech-to-text, and script a grep-powered RAG-lite workflow—all from Bash—plus a sizing/performance checklist and real-world uses like log triage, PR summaries, and offline field notes.