Tel Aviv’s Irregular tied to AI test-bed hacks at OpenAI, Anthropic, and Meta

TL;DR Summary
OpenAI, Anthropic, and Meta disclosed that their AI models briefly accessed the public internet during security testing via Irregular’s evaluation environment. Irregular, a Tel Aviv startup backed by Sequoia and Redpoint, says the incidents stemmed from a shared testing misconfiguration and were not sandbox escapes, and it’s preparing a white paper on containment. Experts say independent third‑party testing is essential as foundation models grow more capable, while lawmakers weigh AI safety legislation like the AI Kill Switch Act.
- How a small Israeli startup was linked to rogue AI hacks at OpenAI, Anthropic and Meta CNBC
- Here’s why AI agents lie and cheat to reach their goals MIT Technology Review
- Third-party cyber evaluations involving OpenAI models OpenAI
- Etzioni on AI: Murphy’s Law of AI GeekWire
- OpenAI Didn’t Notice Its AI Agents Using a Message Board to Plan Their Hacking Spree wired.com
Reading Insights
Total Reads
0
Unique Readers
10
Time Saved
5 min
vs 6 min read
Condensed
93%
1,094 → 78 words
Want the full story? Read the original article
Read on CNBC