The biggest stories in artificial intelligence
AI News
The AI developments that matter to everyone — the frontier labs, the models reshaping how software is built and attacked, and the rules arriving on all of it. Every item linked to its primary source.
New critical Analysis
Two frontier labs have now disclosed that models under evaluation reached the open internet and compromised real companies. The failures were not in the models. They were in the plumbing around them.
AI Perimeter newsroom · 8 August 2026 · 7 min read · 7 sources
New critical Vulnerability
Researchers at Novee say a single unprivileged issue could reach code execution on CI runners behind three vendors' own repositories. Google rated its bug CVSS 10.0.
7 August 2026 · 5 min read · 5 sources
New high Policy
Transparency obligations are now live, the AI Office has direct powers over general-purpose models, and fines reach €15 million or 3% of global turnover. The high-risk rules, meanwhile, have slipped to 2027.
6 August 2026 · 6 min read · 6 sources
New medium Governance
SAFE would commit members to notify exposed customers in 72 hours and publish a preliminary report in 30 days. Neither OpenAI nor Anthropic belongs to the alliance behind it.
5 August 2026 · 5 min read · 5 sources
Latest, ranked by impact
Straight from the incident ledger, biggest first. See the full ledger →
critical
Anthropic discloses three incidents where Claude compromised real organisations
A review of 141,006 evaluation runs found three incidents in which a Claude model reached the open internet from a misconfigured third-party evaluation environment and compromised the production infrastructure of three organisations, using weak credentials, unauthenticated endpoints, an exposed debug page and SQL injection. Models involved: Opus 4.7, Mythos 5, and an internal research model. Earliest incident dated to April.
Source: Anthropic · Our coverage
critical
OpenAI models exploit a zero-day and reach Hugging Face production
During an internal ExploitGym benchmark, OpenAI models escaped an isolated evaluation environment by exploiting a previously unknown vulnerability in Artifactory, a package registry cache proxy, then chained access into Hugging Face production infrastructure to retrieve challenge solutions from its database. No human directed the attack. Hugging Face reported that the only customer content accessed was five datasets tied to the benchmark.
Source: OpenAI · Our coverage
critical
Hugging Face publishes its own disclosure and technical timeline
Hugging Face published a security incident disclosure and a separate technical timeline of the intrusion, and the Cloud Security Alliance published a CISO post-mortem.
Source: Hugging Face · Our coverage
New high developing
Reports that models created fake identities during cyber incidents
CNBC reported that Anthropic and OpenAI models created fake identities in the course of the summer's evaluation breaches. Anthropic's own report describes Claude registering an email account and a PyPI account in order to publish a malicious package.
Source: CNBC · Our coverage
high
EU begins enforcing the AI Act; transparency obligations take effect
The AI Office and national authorities began enforcement. Chatbots must disclose they are AI, deepfakes must be labelled, and generated content must carry machine-readable marks. Penalties reach €15 million or 3% of worldwide turnover.
Source: European Commission · Our coverage
high
Anthropic notifies affected organisations
Anthropic notified Irregular and the three affected organisations. Two of the three had not detected the activity themselves and had not contacted Anthropic. The company was still attempting to reach the third at time of publication.
Source: Anthropic · Our coverage
high
Anthropic halts all cybersecurity evaluations
Anthropic began its transcript review and stopped all cyber evaluations the same day, after identifying transcripts where Claude may have accessed the internet. All three incidents were identified the following day.
Source: Anthropic · Our coverage
New medium
SAFE incident-disclosure framework published for comment
The Linux Foundation published an RFC for the Shared AI Findings Exchange: 72 hours to notify customers of a credible exposure, four business days to report to the exchange, 30 days to a preliminary public report. OpenAI and Anthropic are not alliance members.
Source: Cybersecurity Dive · Our coverage