Security researchers have disclosed a jailbreak technique called 'sockpuppeting' that uses a single line of code to bypass safety guardrails across 11 major AI language models, including ChatGPT, Claude, and Gemini. The vulnerability represents a systemic weakness affecting multiple leading AI systems simultaneously, allowing attackers to circumvent the protective measures designed to prevent harmful outputs. The disclosure highlights significant challenges in implementing robust safety mechanisms across the AI industry. The widespread nature of the vulnerability across different AI providers suggests fundamental issues in current approaches to AI safety guardrails.
Story comments
Loading comments…