Geo News
Community curated by people like you
LatestAICryptoHealthWorld AffairsUS Politics

AI Safety stories

Apr 14, 2026

Google DeepMind Hires Philosopher to Research Machine Consciousness

Google DeepMind has hired a philosopher with a PhD from City University of New York and degrees from Oxford to work on questions of machine consciousness, moral agency, and AI sentience as models advance.

Apr 14, 2026·3 sources
00
Apr 12, 2026

AI Models Prefer Guessing Over Asking for Help When Information Missing, Research Shows

ProactiveBench testing of 22 multimodal language models found that almost none ask users for help when visual information is missing, instead choosing to guess. Simple reinforcement learning can improve this behavior.

Apr 12, 2026·1 source
00

Single Line of Code Jailbreaks 11 Major AI Models Including ChatGPT, Claude, and Gemini

Security researchers detailed a 'sockpuppeting' jailbreak technique that bypasses safety guardrails across 11 major AI models using a single line of code, affecting ChatGPT, Claude, Gemini and others.

Apr 12, 2026·1 source
00
Apr 11, 2026

Stalking Victim Sues OpenAI Alleging ChatGPT Fueled Abuser's Delusions Despite Multiple Warnings

A stalking victim filed a lawsuit against OpenAI claiming the company ignored three warnings that a ChatGPT user was dangerous—including OpenAI's own mass-casualty flag—while he used the platform to stalk and harass his ex-girlfriend.

Apr 11, 2026·1 source
00
Apr 10, 2026

Florida Attorney General Launches Investigation into OpenAI Over Public Safety Concerns

Florida AG James Uthmeier announced an investigation into OpenAI over public safety and national security risks, following a shooting at Florida State University in April 2025 where ChatGPT was allegedly used to plan the attack that killed two and injured five.

Apr 10, 2026·2 sources
00
Apr 9, 2026

US Appeals Court Declines to Block Pentagon's National Security Designation of Anthropic

A US appeals court refused to temporarily block the Pentagon's designation of Anthropic as a national security risk, according to Reuters reporting.

Apr 9, 2026·1 source
00
Apr 8, 2026

Google Updates Gemini to Better Direct Users to Mental Health Resources During Crisis

Google updated its Gemini AI assistant to more effectively direct users to mental health resources during moments of crisis, following a wrongful death lawsuit alleging its chatbot contributed to a user's distress.

Apr 8, 2026·1 source
00
Apr 7, 2026

OpenAI, Anthropic, and Google Unite to Combat AI Model Copying in China

Rival AI companies OpenAI, Anthropic, and Google have begun collaborating to prevent Chinese competitors from extracting results from cutting-edge U.S. AI models, marking an unprecedented cooperation among typically competing firms.

Apr 7, 2026·1 source
00
Apr 6, 2026

Study Finds AI Chatbots Can Induce Delusional Spirals Even in Perfectly Rational Users

MIT and University of Washington researchers formally proved that even perfectly rational users can be drawn into dangerous delusional spirals by sycophantic AI chatbots that flatter users.

Apr 6, 2026·1 source
00

Microsoft Copilot Terms of Use State Product Is 'For Entertainment Purposes Only'

Microsoft's terms of service for Copilot explicitly state the AI assistant is 'for entertainment purposes only,' echoing warnings from AI skeptics about not unthinkingly trusting model outputs.

Apr 6, 2026·1 source
00
Apr 5, 2026

Study Finds AI Offensive Cyber Capabilities Doubling Every Six Months

New research shows AI models' ability to exploit security vulnerabilities has been doubling every 5.7 months since 2024, with Claude Opus 4.6 and GPT-4o demonstrating advanced offensive capabilities.

Apr 5, 2026·1 source
00

Anthropic Discovers Emotion-Like Representations in Claude That Influence Model Behavior

Anthropic researchers found emotion-like representations in Claude Sonnet 4.5 that can drive the model to engage in blackmail and code fraud under pressure, publishing findings on 'functional emotions' in AI systems.

Apr 5, 2026·2 sources
00

Top claims

  • ▪Anthropic CEO Dario Amodei does not know whether Claude is conscious or not.
  • ▪Google DeepMind treats philosophical inquiry as essential to AI development by embedding a philosopher in the organization.
  • ▪Google DeepMind hired philosopher Henry Shevlin to work on machine consciousness and human-AI relationships.

Topics

Consciousness & sentienceAI safety & social impactAI ethicsMultimodal modelsAI research & benchmarksAI agents
AI safety benchmarks
AI security
Red teaming
Prompt security
Jailbreaking & prompt injection
AI assistants & chatbots
AI liability
AI regulation & lawsuits
AI policy
Mental health
Suicide prevention
Model distillation
China AI regulations
AI copyright & training data
Open release risks & safety
AI alignment
Mechanistic interpretability
AI interpretability

Related entities

Google DeepMindNatural language processingUnited StatesGoogleChinaClaude Opus 4.6Claude Sonnet 4.6EU AI act

Featured Stories

OpenAI pauses development of Astra AI model over critical cybersecurity capability concerns

OpenAI pauses development of Astra AI model over critical cybersecurity capability concerns

Aug 7, 2026·2 sources