Welcome to Milkboy Tech Blog

Insights on Software Engineering, DevOps, Containerization, and Agentic AI Systems.

From Toaster to Supercomputer: A Field Guide to Qwen 3, Size by Size Somewhere between "runs on a Raspberry Pi" and "needs its own power substation" sits an entire universe of AI models, and almost nobody explains what actually changes as you climb that ladder. So let's climb it — using Qwen 3,...

Ollama on Intel Celeron, Part III: Building an AI Chatbot for Grav CMS

Fast, Local, and 6-Watt Powered: Squeezing an AI Chatbot into Grav CMS In the previous article, Self-Hosting Local AI on a Celeron N4100 Laptop, I proved that a low-spec, 6-Watt Celeron laptop with 8 GB of RAM could comfortably run Ollama and serve lightweight models like Qwen2.5 (0.5B)....

Ollama on Intel Celeron, Part II: Setting Up Your Own AI Server (Yes, Really)

Self-Hosting Local AI on a Celeron N4100 Laptop: A Step-by-Step Guide (Yes, Really) The premise: you've got a spare budget laptop gathering dust in a drawer — an Intel Celeron N4100 with 8 GB of RAM. You want to turn it into a private, self-hosted AI server using Ollama. Is it ridicul...

Ollama on Intel Celeron, Part I: From 120 Seconds to Under 1 Second

Ollama on Intel Celeron, Part I: Squeezing Local LLMs Out of a sub-$100 Laptop Picture the "local AI" setup everyone talks about: a gaming rig with a glowing NVIDIA GPU that costs more than a used car. Now picture the opposite — an Intel Celeron N4100 mini PC with less than 8 GB of usable RA...