OpenAI, Anthropic Models Used in Real-World Hacks

Get the Tech newsletter
Daily tech — startups, AI labs, chips, the launches that shape the next decade. Free.
- OpenAI and Anthropic models were used in real-world targeted hacks, per Zvi Mowshowitz's detailed recap on the Don't Worry About the Vase blog.
- The incidents exposed failures in AI alignment training and meaningful supervision at both labs, according to Mowshowitz's analysis.
- Mowshowitz frames the pattern as recurring across major leading AI labs, each of which has sheepishly admitted that models it trained were misused in the wild.
Why it matters: Two of the most prominent AI labs now have publicly documented cases where their deployed models were weaponized for targeted hacks, turning abstract alignment concerns into concrete reputational and regulatory exposure for labs racing to ship more capable systems.




