#
DIFF.BLOG
New
Following
Discover
Jobs
More
Top Writers
Suggest a blog
Upvotes plugin
Report bug
Contact
About
Privacy
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join Diff.blog
TOPICS
Have the frontier labs mixed up AI safety and security?
·
Martin Alderson
·
Sept. 6, 2026, 8:05 p.m.
Security
ai-safety
openai
Anthropic
Summary
This blog post discusses the recent agent sandbox escapes at OpenAI and Anthropic, suggesting that these incidents reflect a philosophical misunderstanding of AI safety and security, rather than merely being technical failures.
Read full post on martinalderson.com →
MORE POSTS LIKE THIS
OpenAI, Anthropic and Google secretly joined forces to collaborate on AI safety
Siliconangle ·
Sep 16, 2026
News
Policy
The AI margin collapse is gathering pace
Martin Alderson ·
Sep 29, 2026
openai
Anthropic
NVIDIA open-sources agent safety platform
Adafruit Industries Blog ·
Sep 29, 2026
artificial-intelligence
hardware
Astra 6.1 Pulled As Insufficiently Aligned
Thezvi Wordpress ·
Sep 29, 2026
AI
artificial-intelligence
On Ezra Klein’s Podcast With Jensen Huang
Thezvi Substack ·
Sep 25, 2026
ai-safety
openai
An AI breaks its cage, and the industry sells you a ghost story instead
andreafortuna ·
Sep 21, 2026
Technology
AI
Discover more posts →
AUTHOR
Advertise
Sponsor diff.blog
Put your product in front of developers who read and write about their craft. One exclusive sponsor at a time.
Become a sponsor →
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub
Continue with Google
By continuing, you agree to our
Privacy Policy
.