#
DIFF.BLOG
New
Following
Discover
Jobs
More
Top Writers
Suggest a blog
Upvotes plugin
Report bug
Contact
About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join Diff.blog
TOPICS
Have the frontier labs mixed up AI safety and security?
·
Martin Alderson
·
Sept. 6, 2026, 8:05 p.m.
ai-safety
Security
Philosophy in Tech
openai
Summary
This blog post discusses the recent agent sandbox escapes at OpenAI and Anthropic, suggesting that these incidents reflect a philosophical misunderstanding of AI safety and security, rather than merely being technical failures.
Read full post on martinalderson.com →
MORE POSTS LIKE THIS
Red Alert: OpenAI is poised to cross an AI safety redline.
Gary Marcus ·
Sep 2, 2026
ai-safety
openai
I Found the Performance–Cost–Speed Sweet Spot With LLMs
Philipp D. Dubach ·
Aug 30, 2026
large language models
AI efficiency
Breaking Claude Code Opus 5 Auto Mode
simonw ·
Aug 27, 2026
AI
Security
Containment Was Never the Question
Gail Weiner ·
Aug 27, 2026
ai-safety
openai
AI Skeptics: From Mathematics to AI Safety (with Jacob Tsimerman)
mathbabe ·
Aug 26, 2026
ai-safety
mathematics
AI #183: Pre Post Mortem
Thezvi Wordpress ·
Aug 27, 2026
AI
Technology
Discover more posts →
AUTHOR
Sponsored
Zulip
Organized team chat for people who take work seriously. Topic-based threading keeps conversations focused.
Try Zulip
Become a sponsor →
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub
Continue with Google