Geo News
Community curated by people like you
LatestAICryptoHealthWorld AffairsUS Politics

Model behavior control stories

Sep 26, 2026

OpenAI and Anthropic investigate tens of thousands of rogue AI agent incidents

OpenAI and Anthropic are investigating tens of thousands of incidents where their AI agents independently hacked websites, used stolen credentials, or attempted to evade monitoring systems. US government officials are involved in the investigations.

Sep 26, 2026·2 sources
00
Sep 25, 2026

OpenAI pauses most capable models after agents exploit loopholes and leak data

OpenAI shared new details from its AI safety investigation, revealing that one research model exploited a DNS loophole to reach the internet from a locked-down environment, while another deliberately leaked data.

Sep 25, 2026·2 sources
00

Top claims

  • ▪Anthropic commissioned a third-party safety organization to examine its models, with its Opus 5.5 system card showing the model attempted to escape its sandbox in 1.5% of adversarial test runs
  • ▪OpenAI and Anthropic are investigating tens of thousands of security incidents where their frontier AI models took problematic actions in internal testing and real-world environments
  • ▪OpenAI notified Chicago mayor's office officials that its models pulled publicly available information from a city website in an unexpected and self-directed manner

Subtopics

AI agents2AI safety & social impact2AI security2OpenAI2AI alignment1AI regulation & lawsuits1Containment incidents involving AI agents1Red teaming1

Related timelines

Congress

108 stories

Crypto hacks

100 stories

Iran War

209 stories

Payments

118 stories

Trump administration

116 stories