509 posts tagged with "Kubernetes"
Container orchestration, Cluster API, and self-hosted control planes

Hetzner's Second Price Hike of 2026: Why Your CCX Fleet Just Got 169% More Expensive
Hetzner's June 15, 2026 hike tripled CCX and CPX while CX and CAX rose 35%. A line-by-line recompute of what a CAPH fleet costs on the wrong family versus the right one — and how to fix your MachineDeployment default.

Hetzner at €3.79 vs OVHcloud at $9.99 vs DigitalOcean at $24: What a 6x Price Gap Still Buys After Three 2026 Price Hikes
February's viral table priced the same 2 vCPU/4 GB box at €3.79 on Hetzner, $9.99 on OVHcloud and $24 on DigitalOcean. After Hetzner's April and June hikes and OVHcloud's 9-11% increase, the gap did not close — bandwidth math made it wider.

Kubernetes 1.36 'Haru': Three GA Graduations Your Small Fleet Actually Needs
Kubernetes 1.36 ships 70 enhancements with 18 graduating to Stable. Only three change day-to-day ops for a modest Hetzner-backed fleet: User Namespaces, MutatingAdmissionPolicy, and Volume Group Snapshots — with YAML, constraints, and the 15 you can safely defer.

Northflank's No-Seat-Fee BYOC Still Has Two Bills: What a Rented Control Plane Costs Against One Hetzner Invoice
Northflank charges no seat fee — just $0.01667/vCPU-hour and $0.00833/GB-hour — but BYOC still means two bills: your cloud provider plus a flat platform fee per vCPU. A line-by-line teardown at three workload sizes against a single Hetzner CX/CPX invoice, with seat-count and utilization sensitivity.

OpenCost 1.121 Finally Answers: What Does Each Token Cost on Your Own GPUs?
OpenCost 1.121 adds Kubernetes-native per-token inference metering via llm-d — allocation vs. usage cost, KV-cache-corrected — turning self-hosted LLM spend from a quarterly guess into a Prometheus metric you can alert on.

Proxmox VE Evaluations Up 340%: What the VMware Exodus Means for Your Next Cluster API Provider
Gartner reports Proxmox VE evaluations up 340% as Broadcom's VMware bills rise 800-1500%. What Proxmox 9.1's OCI-to-LXC feature actually does, how CAPMOX compares to CAPH, and when a second Cluster API provider earns its place in a self-hosted fleet.

vLLM vs Ollama in Production: What PagedAttention's 19x Throughput Gap Really Buys on Owned GPUs
Red Hat's 2026 benchmark put vLLM at 793 tokens per second and Ollama at 41 on the same GPU and model — a 19x gap. The same comparison at one concurrent user shows them within 20%. A concrete breakdown of where PagedAttention earns that gap, where it doesn't, and what each engine costs to run on an owned Hetzner GPU fleet.

Who Backs Up the Management Cluster? Rebuilding the CAPI Brain When It Dies
Your workload clusters keep serving traffic when the management cluster dies — but nothing can scale, heal, or upgrade until you rebuild the brain. A line-by-line walkthrough of three recovery paths with concrete commands, the stale-state trap, and a quarterly drill to prove it works.

Waking Idle Apps at the Proxy Layer: Sablier vs KEDA vs Knative for the 30-App Fleet
Sablier wakes idle containers at the reverse-proxy layer — Traefik, Caddy, Nginx — with no operator or sidecar. How its blocking and waiting-page strategies compare to KEDA and Knative for a fleet of dozens of low-traffic apps, what idle-hour math actually saves, and where HTTP-only wake hits its ceiling.