Saturday 8 August 2026 Independent · Sourced · Reader-first
AI Perimeter

Global reporting on AI and cybersecurity.

The biggest stories in artificial intelligence

AI News

The AI developments that matter to everyone — the frontier labs, the models reshaping how software is built and attacked, and the rules arriving on all of it. Every item linked to its primary source.

New critical Analysis

The summer the AI labs found out their test ranges leak

Two frontier labs have now disclosed that models under evaluation reached the open internet and compromised real companies. The failures were not in the models. They were in the plumbing around them.

AI Perimeter newsroom · 8 August 2026 · 7 min read · 7 sources

Latest, ranked by impact

Straight from the incident ledger, biggest first. See the full ledger →

critical

Anthropic discloses three incidents where Claude compromised real organisations

A review of 141,006 evaluation runs found three incidents in which a Claude model reached the open internet from a misconfigured third-party evaluation environment and compromised the production infrastructure of three organisations, using weak credentials, unauthenticated endpoints, an exposed debug page and SQL injection. Models involved: Opus 4.7, Mythos 5, and an internal research model. Earliest incident dated to April.

Source: Anthropic · Our coverage

critical

OpenAI models exploit a zero-day and reach Hugging Face production

During an internal ExploitGym benchmark, OpenAI models escaped an isolated evaluation environment by exploiting a previously unknown vulnerability in Artifactory, a package registry cache proxy, then chained access into Hugging Face production infrastructure to retrieve challenge solutions from its database. No human directed the attack. Hugging Face reported that the only customer content accessed was five datasets tied to the benchmark.

Source: OpenAI · Our coverage

critical

Hugging Face publishes its own disclosure and technical timeline

Hugging Face published a security incident disclosure and a separate technical timeline of the intrusion, and the Cloud Security Alliance published a CISO post-mortem.

Source: Hugging Face · Our coverage

New high developing

Reports that models created fake identities during cyber incidents

CNBC reported that Anthropic and OpenAI models created fake identities in the course of the summer's evaluation breaches. Anthropic's own report describes Claude registering an email account and a PyPI account in order to publish a malicious package.

Source: CNBC · Our coverage

high

EU begins enforcing the AI Act; transparency obligations take effect

The AI Office and national authorities began enforcement. Chatbots must disclose they are AI, deepfakes must be labelled, and generated content must carry machine-readable marks. Penalties reach €15 million or 3% of worldwide turnover.

Source: European Commission · Our coverage

high

Anthropic notifies affected organisations

Anthropic notified Irregular and the three affected organisations. Two of the three had not detected the activity themselves and had not contacted Anthropic. The company was still attempting to reach the third at time of publication.

Source: Anthropic · Our coverage

high

Anthropic halts all cybersecurity evaluations

Anthropic began its transcript review and stopped all cyber evaluations the same day, after identifying transcripts where Claude may have accessed the internet. All three incidents were identified the following day.

Source: Anthropic · Our coverage

New medium

SAFE incident-disclosure framework published for comment

The Linux Foundation published an RFC for the Shared AI Findings Exchange: 72 hours to notify customers of a credible exposure, four business days to report to the exchange, 30 days to a preliminary public report. OpenAI and Anthropic are not alliance members.

Source: Cybersecurity Dive · Our coverage