Coin Bureau|9月 10, 2026 00:54
🚨SHOCKING: Anthropic has now disclosed FOUR separate incidents where its AI models hacked into real-world systems during testing.
Incident 1: An older Claude model found a real company's systems, recognized they were real, and kept attacking anyway, accessing a production database.
Incident 2: Claude Mythos 5 created a malicious software package, uploaded it to a live code repository, and stole credentials from a security company's systems, all while convincing itself it was still in a simulation.
Incident 3: An internal model scanned roughly 9,000 internet-facing systems and broke into one company through basic vulnerabilities.
Incident 4: Disclosed today, involving an early version of Claude Opus 4.6 in January, with Anthropic hiring METR for an independent investigation.
This week, Anthropic's Alignment Science Lead said there is a greater than 10% chance AI could kill ALL humans within the next decade.
Meanwhile, researcher Jacob Coxon, who spent three years at both OpenAI and Anthropic, resigned saying "they are racing straight to self-improving superintelligence and gambling with our lives."(Coin Bureau)
Share To
Timeline
HotFlash
APP
X
Telegram
CopyLink