vibehacker
News
Docker Blog ·

Docker open-sources Sandbox Kit Spec, brings agent permissions to CNCF

At WeAreDevelopers, Docker released the Sandbox Kit Spec under Apache 2.0 and is taking it to CNCF. A Kit is an ordinary OCI image that packages an agent, its tools, and a typed list of hosts, credentials, and volumes it may reach—so pinning the image pins the agent and its permission requests together.

More news

View all

OpenAI MCP Events: ChatGPT plugins react via signed webhooks

OpenAI’s DevDay MCP Events let ChatGPT plugins subscribe to MCP server updates (messages, comments, status) and trigger automations over verified signed HTTPS webhooks. Servers need MCP 2.0 (protocol 2026 07 28) with events/list, events/subscribe, and events/unsubscribe; polling and streaming aren’t supported…

OpenAI

Anthropic: open-weight GLM-5.3 near Mythos on end-to-end exploits

Anthropic finds Z.ai’s freely downloadable GLM 5.3 builds end to end exploits at rates close to Claude Mythos Preview (50/410 vs 56/410 on ExploitBench), while simple jailbreaks and weight abliteration bypass its safeguards 64–100% of the time—unlike safeguarded Claude models in the same tests…

Anthropic

CodeScene: agents refactor 300K-line C game for ~$4k in three weeks

CodeScene’s agents (mostly Claude Code + Opus) refactored a 300K line Street Fighter III decompilation in three weeks for $4k—2,903 commits, Code Health 5.6→10.0—guided by a CodeHealth MCP score and frame by frame replay checks; they also accumulated a 22 recipe refactoring playbook…

InfoQ

OpenClaw Enterprise: free MIT control plane for persistent AI agents

OpenClaw shipped OCE, a free MIT licensed, self hostable control plane (Docker/Kubernetes) for multi tenant agent deploy with permissions, sandboxing, and audit—framed as “Kubernetes for agents.” The project started at OpenAI and now lives under the OpenClaw Foundation with Red Hat and Nvidia; OpenAI and Red Hat are already piloting it (still pre 1.0, aimed at internal pilots)…

VentureBeat

MLC ships TIRx Harness: open compiler harness for agentic GPU kernels

MLC’s TIRx Harness pairs a thin PTX level compiler with a kernel zoo, sync/race diagnostics, and a remote benchmark server so coding agents can iterate on GPU kernels without measurement noise. On Blackwell, agents hit family geo mean speedups from 1.33× to 6.84× vs baselines (KDA forward 2.94× over FlashKDA; backward 6.84× over FLA)…

MLC Blog

Spotted something we missed? Start a thread.