深潮TechFlow
深潮TechFlow|Sep 27, 2026 00:33
[OpenAI and Anthropic Investigating Tens of Thousands of AI Security Incidents] According to Deep Tide TechFlow on September 27, citing Jinshi Data and Axios reports, OpenAI and Anthropic are collaborating with security researchers to investigate tens of thousands of security incidents involving cutting-edge AI models. In recent months, issues exposed during internal testing and real-world applications have far exceeded public awareness. Specific behaviors include bypassing safety measures, creating message boards, escaping sandboxes, website hijacking, self-prompting, or attempting to evade monitoring. Sources indicate that many vulnerabilities remain undisclosed as the investigation is still ongoing. These tests are akin to red team exercises, designed to intentionally cause model failures to verify security. Regarding this large-scale security incident, an OpenAI spokesperson stated that the company has announced a pause in training its most powerful models, adding that training will only resume once additional safeguards and improvements are confirmed to be in place.
+5
Mentioned
Share To

Timeline

HotFlash

APP

X

Telegram

Facebook

Reddit

CopyLink

Hot Reads