Welcome to Milkboy Tech Blog

Insights on Software Engineering, DevOps, Containerization, and Agentic AI Systems.

Ollama on Intel Celeron, Part IV: Benchmarking Tailscale, LAN, and Chatbot Optimizations

Benchmarking Ollama Over Tailscale & LAN: Squeezing Sub-Second Latency Out of the Grav AI Chatbot In Part III of this series, we built a 5-tier pipeline for grav-ai-chatbot, running local AI search on a 6-Watt Intel Celeron laptop. But once real traffic started hitting it from different dev...