Three and a Half Weeks, 689M Tokens, No Meter
689M tokens. 25 days. ~$40 electricity. The math, the caveats, and what changed when the meter stopped existing.
689M tokens. 25 days. ~$40 electricity. The math, the caveats, and what changed when the meter stopped existing.
Three days into making Qwen3.8 27B my daily local coding driver on dual 3090s — the hybrid linear-attention architecture, an OOM that masqueraded as a backend error, and what the …
I hooked up multiple local models behind a single proxy that routes traffic by complexity. Here is what I found deploying it on a cheap Intel GPU.
Hardening an internet-facing web server sitting in a home DMZ: firewall, SSH, SELinux, fail2ban, and tunnel-only traffic via Cloudflare.
I use GitHub for public collaboration and Forgejo for self-hosted control. They share the same Actions workflow syntax, just in different folders. Here is why running both makes …
How I deployed self-hosted Plausible analytics on a podman server, routed it through Cloudflare tunnels, and integrated it with my Hugo site — with a backup strategy that actually …
How self-hosted AI became the final piece of my homelab puzzle, delivering true parallel processing for multi-user setups and unlocking the real superpower of knowledge management.
How I went from a blank Docker template to 116+ tok/s with speculative decoding, FlashInfer, and a 160k context window on dual 3090s.
A tour of my open-source Proxmox utility toolkit — scripts, docs, and learning paths for homelabbers and enterprise engineers who are tired of repeating themselves.
A real-world comparison of Qwen3.5 27B Q8 and 35B-A3B Q8 running locally on a dual RTX 3090 homelab — which one actually belongs in your daily workflow?
Showing 1-10 of 14 items (Page 1 of 2)