531 posts tagged with "AI agents"
AI agents and autonomous systems

Netlify Put Claude Code Inside the Deploy Pipeline: Embedded Runners vs MCP Servers for Agent-Driven Deploys
Netlify runs Claude Code and Codex inside its own infrastructure with full pipeline access; self-hosters expose deploys over MCP instead. A concrete accounting of what each position buys, what each costs, and why a platform on owned machines should ship the MCP surface first.

One MCP Server Beats N Chatbot Plugins: What Railway's ChatGPT and Grok Launches Prove
Railway launched official plugins for Grok and ChatGPT in July 2026, but the ChatGPT plugin runs on Railway's hosted MCP server. A cost-by-cost comparison of per-chatbot plugins versus one protocol-standard server, plus the five-step playbook for self-hosted platforms.

Give Your Agents a Memory That Lives on Your Machines: Qdrant vs Weaviate vs pgvector
AI agents without persistent memory re-derive everything on every run. A number-backed comparison of the three self-hostable vector stores — plus the per-tenant scale ladder that picks pgvector as the default and Qdrant as the escape hatch.

88% of AI Agent Pilots Never Reach Production — the Gate Is Deployment Infrastructure, Not the Model
IDC puts the enterprise AI pilot failure rate at 88% and blames deployment readiness, not models. A control-by-control gap analysis of SSO, audit logging, secret scanning, PR gates, license policy, sandboxing, and runbooks against a Render-compatible git-push PaaS with MCP.

ClawBleed: How One Clicked Link Turned 40,000 Self-Hosted Agent Gateways Into Remote Shells
CVE-2026-25253 let one malicious link steal an OpenClaw gateway token and take over the host. The kill chain, the 40,000 exposed instances, and the gateway-auth checklist for anything an agent can deploy through.

100x Faster Than Containers: What Cloudflare's Dynamic Workers Get Right About Sandboxing Agent Code
Cloudflare's Dynamic Workers run AI-generated snippets in V8 isolates with millisecond startup — but the 100x headline only holds for sub-second tool calls. Here are the real numbers against containers and microVMs, the isolation tradeoff Spectre exposed, and the two-tier sandbox rule for self-hosted platforms.

One Cluster for Training and Inference: How a Bank Cut AI Token Costs 60% on Kubernetes
China Merchants Bank won the CNCF End User Case Study Contest for running AI training and inference on a single Kubernetes stack of nearly 10,000 accelerator cards — lifting utilization from 35% to over 60% and cutting inference cost per million tokens by more than 60%. The component-by-component breakdown, the fairness rules behind it, and what smaller self-hosted fleets can copy.

Coolify's MCP Server Just Made Deploy-From-Chat Table Stakes: What 'Ask Claude to Ship It' Really Covers (and What It Doesn't)
Coolify's native read-only MCP endpoint plus 42-tool community servers let agents deploy to self-hosted infrastructure from chat. A labeled walkthrough of the Postgres-to-FastAPI loop, and the five governance controls a fast follower needs.

Cursor Made Sandboxes Interchangeable: What Eight Backends Behind One Agent Pool Means for Self-Hosted Platforms
Cursor's Self-Hosted Machines launch routes cloud-agent sessions through one worker pool across eight sandbox backends — Lambda, Cloudflare, Coder, Daytona, E2B, Modal, Namespace, and Vercel. The pool and the idle-machine lifecycle are the real product, and here is the four-item checklist for making your own fleet backend number nine.