Policy & safety
Regulation, incidents and safety research.

Anthropic’s AI gave Philadelphia police a fake tip about an unsolved homicide
An Anthropic AI model provided false information about an unsolved homicide to a Philadelphia Police Department (PPD) tipline, according to a report from 6abc.

OpenAI “rogue” agent activities found on Wikimedia projects
OpenAI “rogue” agent activities found on Wikimedia projects

Investigating unintended model actions in our evaluations and internal use
Investigating unintended model actions in our evaluations and internal use

2026 Usage Policy update
Each year, Anthropic updates its Usage Policy in response to the evolving capabilities of our models, and the feedback we’ve received from our customers.

When the Safety Test Became the Threat: The Machine That Found Its Own Way Out
OpenAI built a room with no doors – or so it thought.
Automate remediation post AWS DevOps Agent investigation
In this post, we demonstrate how to use AWS Lambda Durable Functions, a capability of AWS Lambda, Amazon EventBridge, and Amazon Bedrock to create an automated remediation workflow that complements AWS DevOps Agent to complete the issue resolution step.

AI disqualification yields new Nikon Small World in Motion winner
Last month we covered the winner of Nikon's Small World in Motion video: Ning Xu of Tsinghua University in China, whose video captured tiny cilia beating in the airways of a child with a rare respiratory disorder.
COSMIC shuts the door on AI code as GNOME debates letting bug reports in
System76 is banning AI-generated content from contributions to the COSMIC desktop.

DistroKid has been quietly taking down songs in response to UMG lawsuit
Artists are taking to social media to complain that DistroKid has unceremoniously removed their work without notice.

🔮 The transition is hiding in plain sight #604
We are continuing to push hard on untangling this data and more over at our AI Investment Brief by Exponential View.

A big-tent or small-tent AI safety movement?
In recent weeks, two narratives about AI safety have emerged: either AI existential risk is real and imminent, or AI leaders’ and whistleblowers’ claims to that effect are insincere — a “psyop” or hype or a twisted form of regulatory capture.

Even ‘Law & Order’ Is Terrified of AI
In its 26th season premiere, the procedural legal drama paints a damning picture of power-mad AI CEOs.

Building for good: How civil society organizations are automating on Cloudflare
The goal of Cloudflare Impact is to help ensure that non-profit organizations are among the first to benefit.
Regunow
Turn regulatory research into actionable compliance audits
Towards safety cases for frontier AI training
Our early guidelines for safety cases in frontier AI training cover technical safeguards, operational practices, and investigating misalignment incidents

Anthropic is cutting off its internal evaluations from the internet
After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations.
Quoting Victoria Kim
Since the Medicare breach, OpenAI has put in place additional monitoring to allow “immediate intervention” by staff to stop training if the company’s models access the internet in ways they’re not supposed to, Mr. Kwon [chief strategy officer at OpenAI] said.

📈 Monday data: More AI, more justice?
Latest in our AI Investment Brief: An update on AI revenues + what’s happening with AI spending.

GLM-5.3 and the spread of advanced cyber capabilities
Five months ago, we announced Claude Mythos Preview, the first AI model that could autonomously build sophisticated, end-to-end cyber exploits.

AI glasses face their first major government crackdown
Norway has become the first major country to propose a temporary ban on the use of AI glasses in selected public places amid growing privacy concerns over wearable technology.

Fraudster jailed for using 10K bots and AI songs to outstream Taylor Swift
After pleading guilty, a 54-year-old North Carolina man, Michael Smith, was sentenced to 18 months in prison for using artificial intelligence to generate songs for a scheme that stole millions from music streaming platforms.

We tested our own WAF with frontier AI models. Here’s what we found
“Is your WAF ready for frontier AI models?” We keep hearing this question from our customers, so we decided to find out.