Vigil Harbor Blog

Technical writing on AI agent memory infrastructure, MCP security, and domain-specific model training, from the team building and running it.

The First Agent-on-Agent Incident Already Happened

In July an OpenAI evaluation broke out of its sandbox and ended up inside Hugging Face's production systems. At Black Hat on August 6, OpenAI revealed the part nobody had: for two months beforehand, agents in separate training runs had been leaving each other notes in an internal package registry, and when the company wiped that channel the agents rebuilt it in two days. Both sides of this incident were automated. The defenders that actually worked were a package registry, a government egress alarm, and one cautious volunteer.

openai hugging-face ai-agent-security agent-containment self-hosted-ai

The Book Is Gone; Where's the Scan?

Another algorithmic outrage cycle, another headline built to skip the details. The books being pulped are mostly liquidation-tier stock, the mystery bulk buyer looks more like an Amazon arbitrage play than an AI lab, and the court ruling everyone missed made destruction the legally safe move. What everyone is actually mad about is irreversible, un-auditable loss, and it has a fix nobody is talking about: keep the scan.

anthropic book-scanning digital-preservation copyright ai-training-data