
DRA or HAMi? What Actually Shares Your GPUs in 2026
Kubernetes DRA now handles the scheduling half of GPU sharing, but runtime enforcement still belongs to HAMi. Here is the converged architecture, the concrete migration path, and which layer each fleet shape should standardize on.

A Chip Vendor Put Its Catalog Inside Your AI Assistant: What Microchip's MCP Server Teaches Every Platform API
Microchip's MCP server lets any AI assistant answer engineering questions from verified parts data. The five things it does that every platform API must copy — and the read-only-vs-read-write caveat that decides whether your agent interface is safe.

Paying $4,700 for a $1,999 GPU: What the RTX 5090 Street Premium Does to Self-Hosted Inference Math
An RTX 5090 at today's $4,700 street price adds about $75 a month to a self-hosted inference node — while H100 rentals fell roughly 70% from their peak. A worked amortization of both price curves, and where buying still wins.

Your GPUs Are 60% Idle and Kubernetes Says Everything Is Fine
A July 2026 CNCF case study found Kubeflow GPUs ~60% idle with every pod green — a scheduler-vs-network conflict fixed with topology constraints (40% to 85% utilization). The three families of GPU waste, a DCGM-based instrumentation ladder, and the rule for earning your next GPU node.

Kubernetes Metrics API Is Stable in v1.37: What a Self-Hosted PaaS Must Expose Before Agents Can Safely Autoscale Apps
Kubernetes 1.37 made the Metrics API stable, but stability doesn't make it safe for an AI agent to autoscale on. Here's the isolation, staleness, and admission contract a self-hosted PaaS needs first.

Kubernetes Metrics API Is Stable in v1.37: The Contract Agents Need Before Autoscaling Apps
Kubernetes v1.37 makes metrics.k8s.io stable, but stable data is not safe automation. Here is the concrete, tenant-scoped contract a self-hosted PaaS should expose before an AI agent changes replicas.

AI-Linked Crypto Scams Extract 4.5× More Per Operation. Build a Defense That Doesn’t Depend on Spotting a Deepfake
AI-linked crypto scams extracted 4.5 times more revenue per operation in 2025. Here is a practical control map for stopping deepfake, identity, and payment fraud without betting everything on detection.

An Agent Can Run Your Home Lab. That Doesn't Make It a Fleet Control Plane.
A concrete boundary between an MCP agent operating a Coolify home lab and a declarative Cluster API fleet, with practical signals for when the architecture needs to change.

Crossplane vs. Cluster API: Draw the Boundary Before AI Draws It for You
Crossplane and Cluster API use Kubernetes reconciliation, but they own different things. Map the boundary between cluster lifecycle, external services, and tenant APIs before exposing infrastructure to AI agents.