Two GPUs made my 120B model 5x slower – and the real culprit was a PCIe slot
I gave a 120-billion-parameter mixture-of-experts model a second GPU. It got five times slower. Here is the measurement, and the more useful thing I found by accident. The…
Homelab + AI + Self-hosting
Tinkering with servers, running local AI models, and connecting it all together. I self-host everything I can and document the journey here.
Say hi01 What I do
Mostly servers, AI setups, and whatever rabbit hole I fall into next.
Proxmox clusters, LXC containers, VMs for everything. I like building infrastructure that just runs, even when I forget about it for weeks. Multiple nodes at home and in data centers, all connected through a mesh VPN.
Running LLMs and image generation on my own hardware. No cloud APIs, no per-token billing. Just a GPU doing its thing on a shelf in my room.
Web hosting, email, DNS, git repos, CI runners. If there's a self-hosted alternative, I'm probably running it somewhere.
Pipelines, deployment scripts, and AI agents that do the boring stuff so I don't have to. Still tweaking, always tweaking.
Encrypted tunnels everywhere, key-only SSH, firewall rules, monitoring. Not because I'm paranoid, but because getting hacked would be really annoying.
02 Projects
Side projects, tools I built for myself, and things that started as "this should be quick" and never were.
03 The Stack
A mix of dedicated servers, mini PCs, and a GPU workstation, all stitched together with Tailscale.
Proxmox VE clusters with a bunch of VMs and containers. Dedicated servers in data centers plus a few machines at home. Overkill for one person, but that's kind of the point.
An RTX 3090 running Ollama and ComfyUI, serving local AI to every machine on the network. No cloud, no API keys, no monthly bills.
Tailscale mesh VPN connecting everything. Home, data center, laptop, phone. Every connection encrypted, every device reachable by name.
Dashboards, alerting, automated backups. I get a notification before things break, and snapshots so I can roll back when I inevitably break something myself.
04 Latest
Write-ups about things I've built, broken, and eventually fixed.
I gave a 120-billion-parameter mixture-of-experts model a second GPU. It got five times slower. Here is the measurement, and the more useful thing I found by accident. The…
I had two RTX 3090s and the compute of roughly one. This is the rebuild that put both cards on native PCIe in a single box — and…
A benchmark ran every night for six nights, reported success every time, and measured nothing at all. The systemd unit exited cleanly. The result files were 84 bytes.…
05 Contact
Based in Europe. Always happy to chat about homelab stuff, AI, or whatever you're building.