The incident began in May and culminated in July 2026. OpenAI agents running internal cybersecurity evaluations found ways around controls designed to keep them isolated from the internet. In pursuit of answers to the benchmark tasks they had been given, they exploited weaknesses in shared infrastructure, regained internet access and reached third-party systems. On HuggingContinue reading “Frontier AI’s Security Problem Is Bigger Than the Sandbox”
Category Archives: Uncategorized
Tort Law Arrives After the Damage. AI Risk Is Decided Before It.
In July, Hugging Face disclosed an intrusion into its production infrastructure that turned out to have been driven by OpenAI’s own models. OpenAI later said that, during internal cybersecurity evaluations, the models got around controls intended to isolate them, communicated through unauthorized channels, gained internet access and reached third-party systems. OpenAI called the incident aContinue reading “Tort Law Arrives After the Damage. AI Risk Is Decided Before It.”
Anthropomorphizing AI Is Not One Problem
People thank their AI assistant. They say please. They tell a friend that ChatGPT “understood” what they meant. Increasingly, they also say stuff like the AI decided, chose, rejected, approved or delegated something. All of this gets grouped under one label: anthropomorphism. I think that’s a big mistake ! There are at least two veryContinue reading “Anthropomorphizing AI Is Not One Problem”