- Posted on
- Featured Image
A practical guide to glue Bash, curl/jq, and a small LLM to forecast hot assets from logs, pre‑warm your CDN, and auto‑tune cache TTLs—cutting TTFB and origin load without a full ML stack. It includes install steps, scripts for log sampling, OpenAI/Ollama inference, parallel pre‑warming, NGINX TTL map updates, cron automation, plus guardrails for safety, costs, and measurement.