The archive, mapped
Every blog post and case study on the site with the principles and playbooks each piece argues for. Filter by any principle or playbook to see just the evidence behind it.
Matching the filter
6 items- Two-Lane Text GPU Allocation: Quality + Vision/Fast (Plus a Media Lane)Feb 9, 2026 · 11 min
A February 2026 GPU allocation experiment: separating text workloads, testing shared-group priorities, and keeping client routes stable as models moved.
How I evolved a small AMD GPU cluster into a CRD-driven inference platform with an OpenAI-compatible boundary, evidence-based model promotion, and safe rollouts.
- Standing Up a GPU-Ready Private AI Platform (Harvester + K3s + Flux + GitLab)Dec 29, 2025 · 5 min
Field notes from building and operating a small private GPU platform with Harvester, K3s, and a GitLab -> Flux delivery loop.
- Hybrid/On-Prem GPU: The Boring GitOps PathDec 29, 2025 · 5 min
The contracts I use to operate mixed-vendor GPU capacity with Kubernetes and Flux, including cost, storage, scheduling, and rollback limits.
- Welcome to My HomelabNov 27, 2025 · 6 min
An August 2026 tour of the Harvester, K3s, GitLab, Harbor, Flux, and mixed-GPU platform behind the systems I publish here.
How FlexDeck combines Kubernetes, Flux, CI, observability, and model state without hiding freshness, access boundaries, or the limits of a homelab control surface.