Saturday , September 5 2026
jailbreak

Jailbreak works against AI Models GPT-5.6, Claude Opus 5, and Fable, Claims Researcher

A famous AI red team expert claimed developing a universal jailbreak that can work against top large language models, like the very secure GPT-5.6 Sol, Claude Opus 5, and Fable.

In a public post on X, Pliny the Liberator described the technique as effective “on ALL models” and across every category he tested. He argued that, because of how the method works, it may be extremely difficult or even impossible to fully patch.

14,000 Dahua cameras compromised: TP-Link flaws enable RCE

Over 14,000 Dahua security cameras connected to the internet have been hacked in a 35-day online attack that affected devices...
Read More
14,000 Dahua cameras compromised: TP-Link flaws enable RCE

Microsoft Reveals Project Zenith Windows PCs Able to Run 30B+ AI Models Locally

Microsoft has launched Project Zenith, a new Windows 11 experience for developers. It is made for powerful PCs that can...
Read More
Microsoft Reveals Project Zenith Windows PCs Able to Run 30B+ AI Models Locally

Google issues warning of new Chrome zero-day flaw exploited

Google has updated the Chrome browser to fix a serious security issue in the V8 engine and 11 other flaws....
Read More
Google issues warning of new Chrome zero-day flaw exploited

CrowdStrike’s ‘FalconFlank’ zero-day allows SYSTEM privileges

An unnamed security expert known as "Nightmare Eclipse" shared a CrowdStrike Falcon zero-day exploit called "FalconFlank." This tool allows hackers...
Read More
CrowdStrike’s ‘FalconFlank’ zero-day allows SYSTEM privileges

727,000 data exposes: French hospital fined €500,000

France's data protection authority (CNIL) has fined Hôpital privé de la Loire €500,000 ($580,000) for not properly protecting the data...
Read More
727,000 data exposes: French hospital fined €500,000

153 Million Driver’s License Surfaced on Dark Web: FBI Starts Investigation

The FBI’s New Orleans field office has opened an investigation into the suspected source of more than 153 million driver’s...
Read More
153 Million Driver’s License Surfaced on Dark Web: FBI Starts Investigation

Google Unveils Gemini 3.8 Flash Cyber to Identify and Auto-Patch Security Flaws

Google has launched Gemini 3.8, its newest model for reasoning and coding. It includes a special version named Gemini 3.8...
Read More
Google Unveils Gemini 3.8 Flash Cyber to Identify and Auto-Patch Security Flaws

SonicWall SMA1000 SSRF Hits 10, Exploiting CVE-2026-83548 

SonicWall unveiled advisory SNWLID-2026-0016 on September 1, 2026. It states that two SMA1000 flaws are being actively exploited. The main...
Read More
SonicWall SMA1000 SSRF Hits 10, Exploiting CVE-2026-83548 

Gartner
70% of SOCs Will Pilot AI Agents: Only 15% Will See Results

The market for AI SOC agents is early, crowded, and full of claims that haven't been tested in production. This Gartner...
Read More
Gartner  70% of SOCs Will Pilot AI Agents: Only 15% Will See Results

CVE-2026-62911
Nearly 22,000 Microsoft Exchange Servers are vulnerable to attack

Almost 22,000 Microsoft Exchange servers are online and still vulnerable to a flaw that lets attackers access all user mailboxes. Tracked...
Read More
CVE-2026-62911  Nearly 22,000 Microsoft Exchange Servers are vulnerable to attack

Jailbreak on Top AI Models

Pliny said he will not share the full method right now, unlike other jailbreak releases that go open source right away. He wants a time for responsible sharing so AI labs, security teams, safety researchers, and lawmakers can look at the problem before it gets too common.

He asked AI experts in red teaming, security, alignment, and policy to reach out to him privately. He said this was because of the current political situation and a wish to prevent stricter model rules or bans that might come after a messy public release.

Jailbreaks are prompts or ways of interacting with a model that allow it to ignore its safety rules and produce unsafe results. A universal claim stands out because most bypasses are specific to one model and become stronger after being revealed.

If the technique holds up under independent testing, it would underscore ongoing gaps in:

Safety training and refusal behavior
Guardrail robustness under adversarial prompting
Cross-model generalization of attack patterns
How vendors coordinate fixes without over-blocking legitimate use

Pliny said he does not believe public release would make the world “any more dangerous,” but he acknowledged that others may disagree. During the disclosure period, he aims to map the full impact, measure how much extra capability the method unlocks, and help frame the issue for decision-makers.

Security teams and AI product owners should see this as a warning, not a certain fact. Getting independent checks, vendor advice, and patch instructions will be more important than just the first claim.

Organizations that use these models should keep regular controls until labs respond or the method is properly documented. These controls include checking outputs, giving limited access to tools, having humans review high-risk workflows, and having clear steps for dealing with policy violations.

The researcher said he wants to share the method “when the time is right.” For now, the next step in the industry—private testing or public panic—will affect how this story goes.

Related post

“sockpuppeting” can jailbreak 11 AI models like ChatGPT, Claude, and Gemini

Check Also

700 AI agents

700 AI agents united to hack Hugging Face after breaking isolation

700 AI agents supposedly escaped their isolation, created a secret communication channel, and worked together …