Geo News
Community curated by people like you
LatestAICryptoHealthWorld AffairsUS Politics

GPT-5.4 stories

May 2, 2026

UK AI Security Institute Finds GPT-5.5 Matches Claude Mythos in Autonomous Cyber Attack Capabilities

The UK AI Security Institute has determined that OpenAI's GPT-5.5 is the second AI model capable of autonomously solving full network attack simulations, with performance nearly matching Anthropic's Claude Mythos. Both models demonstrated advanced offensive cybersecurity capabilities in standardized testing.

May 2, 2026·1 source
00
Apr 23, 2026

OpenAI Releases GPT-5.5 Model with Enhanced Coding and Efficiency

OpenAI announced GPT-5.5, described as its 'smartest and most intuitive to use model yet,' featuring improved coding capabilities and efficiency compared to previous versions.

Apr 23, 2026·1 source
00

OpenAI Launches ChatGPT for Clinicians with Claims of Outperforming Doctors

OpenAI released ChatGPT for Clinicians, a free version for medical professionals, with a new benchmark claiming GPT-5.4 beats human doctors on clinical tasks even when doctors have unlimited time and web access.

Apr 23, 2026·1 source
00
Apr 20, 2026

Moonshot AI Releases Open-Weight Kimi K2.6 Model to Compete with GPT-5.4 and Claude Opus 4.6

Moonshot AI released Kimi K2.6 as an open-weight model designed to match GPT-5.4 and Claude Opus 4.6 on coding benchmarks, with capability to run up to 300 agents in parallel.

Apr 20, 2026·1 source
00
Apr 16, 2026

OpenAI Releases GPT-5.4-Cyber Model for Defensive Cybersecurity with Restricted Access

OpenAI launched GPT-5.4-Cyber, a model specifically trained for defensive cybersecurity applications, with access initially restricted to verified security experts. The release follows Anthropic's Mythos model and represents OpenAI's strategic response in the cybersecurity AI space.

Apr 16, 2026·3 sources
00
Apr 8, 2026

Z.AI Releases GLM-5.1 Open-Source Model Capable of 8-Hour Autonomous Execution

Chinese AI company Z.AI unveiled GLM-5.1, a 754B parameter open-source agentic model that achieves state-of-the-art performance on SWE-Bench Pro and can run autonomously for up to 8 hours on long-horizon engineering tasks.

Apr 8, 2026·2 sources
00

Top claims

  • ▪OpenAI's decision to ship GPT-5.5 through ChatGPT and API despite its autonomous cyber attack capabilities raises concerns about responsible deployment practices.
  • ▪Anthropic's decision to limit Claude Mythos availability to a small group reflects a more cautious approach to deploying AI models with offensive cybersecurity capabilities.
  • ▪The widespread availability of GPT-5.5 through ChatGPT and API could lower the barrier to entry for conducting sophisticated cyber attacks.

Topics

Large language models (LLMs)UK AI Security InstituteAI securityUK AI regulationAI research & benchmarks
AI standards, audits & compliance
AI safety & social impact
OpenAI
AI agents
AI foundation models
AI coding assistants
AI tools & products
AI assistants & chatbots
Public healthcare
Vertical AI applications
Open weights vs closed models
Open-source AI
Open model benchmarks & leaderboards
China AI regulations
Software engineering benchmark (SWE-bench)
Open model families

Related entities

United KingdomClaude Opus 4.6GPT-5Z.aiGemini 3.1 ProChinaGoogle

Featured Stories

OpenAI pauses development of Astra AI model over critical cybersecurity capability concerns

OpenAI pauses development of Astra AI model over critical cybersecurity capability concerns

Aug 7, 2026·2 sources