Sunday , July 26 2026
jailbreak

Jailbreak works against AI Models GPT-5.6, Claude Opus 5, and Fable, Claims Researcher

A famous AI red team expert claimed developing a universal jailbreak that can work against top large language models, like the very secure GPT-5.6 Sol, Claude Opus 5, and Fable.

In a public post on X, Pliny the Liberator described the technique as effective “on ALL models” and across every category he tested. He argued that, because of how the method works, it may be extremely difficult or even impossible to fully patch.

Jailbreak works against AI Models GPT-5.6, Claude Opus 5, and Fable, Claims Researcher

A famous AI red team expert claimed developing a universal jailbreak that can work against top large language models, like...
Read More
Jailbreak works against AI Models GPT-5.6, Claude Opus 5, and Fable, Claims Researcher

Researchers found security flaws in every script generated by ChatGPT, Copilot, and Gemini

A new study from Beacom College shows that all automation scripts produced by top AI models like ChatGPT, Microsoft Copilot,...
Read More
Researchers found security flaws in every script generated by ChatGPT, Copilot, and Gemini

Australian Energy Giant Origin confirms unauthorized access and disclosure of customer data

Origin Energy Limited, a major energy provider in Australia, has said there was a cybersecurity issue with unauthorized access to...
Read More
Australian Energy Giant Origin confirms unauthorized access and disclosure of customer data

Anthropic Unveils Claude Security Plugin for Code Flaw Scanning

Anthropic launched the Claude Security plugin in beta. This tool uses AI to find serious security flaws in Claude Code....
Read More
Anthropic Unveils Claude Security Plugin for Code Flaw Scanning

Apple, ASUS Router, Meta, Windmill & Ubuntu Patch Critical Security Flaws

ASUS has put out important security updates for a serious router flaw. This issue could let remote hackers run any...
Read More
Apple, ASUS Router, Meta, Windmill & Ubuntu Patch Critical Security Flaws

SolarWinds Patches 15 Critical Serv-U Flaws

SolarWinds has shared important security updates for its Serv-U file transfer software. These updates fix 15 problems that could let...
Read More
SolarWinds Patches 15 Critical Serv-U Flaws

Oracle fixes 1,400+ vulnerabilities; critical flaws threaten enterprise servers

Oracle has fixed over 1,400 security holes in its July 2026 Critical Patch Update (CPU). Most of these flaws were...
Read More
Oracle fixes 1,400+ vulnerabilities; critical flaws threaten enterprise servers

Zimbra Patches 4 XSS and Critical SNMP Command Injection Flaws

Zimbra has launched updates to fix serious security flaws, including a command injection bug in the SNMP monitoring part. As...
Read More
Zimbra Patches 4 XSS and Critical SNMP Command Injection Flaws

Qilin ransomware gang exploiting critical Palo Alto VPN Flaw

The Qilin ransomware group is exploiting a flaw in PAN-OS GlobalProtect to break into victims' networks, says the cybersecurity firm...
Read More
Qilin ransomware gang exploiting critical Palo Alto VPN Flaw

“PentestCode” AI Agent Automating Penetration Testing with 18 Tools

A new free tool is adding AI helpers into security work. PentestCode is a version of OpenCode made just for...
Read More
“PentestCode” AI Agent Automating Penetration Testing with 18 Tools

Jailbreak on Top AI Models

Pliny said he will not share the full method right now, unlike other jailbreak releases that go open source right away. He wants a time for responsible sharing so AI labs, security teams, safety researchers, and lawmakers can look at the problem before it gets too common.

He asked AI experts in red teaming, security, alignment, and policy to reach out to him privately. He said this was because of the current political situation and a wish to prevent stricter model rules or bans that might come after a messy public release.

Jailbreaks are prompts or ways of interacting with a model that allow it to ignore its safety rules and produce unsafe results. A universal claim stands out because most bypasses are specific to one model and become stronger after being revealed.

If the technique holds up under independent testing, it would underscore ongoing gaps in:

Safety training and refusal behavior
Guardrail robustness under adversarial prompting
Cross-model generalization of attack patterns
How vendors coordinate fixes without over-blocking legitimate use

Pliny said he does not believe public release would make the world “any more dangerous,” but he acknowledged that others may disagree. During the disclosure period, he aims to map the full impact, measure how much extra capability the method unlocks, and help frame the issue for decision-makers.

Security teams and AI product owners should see this as a warning, not a certain fact. Getting independent checks, vendor advice, and patch instructions will be more important than just the first claim.

Organizations that use these models should keep regular controls until labs respond or the method is properly documented. These controls include checking outputs, giving limited access to tools, having humans review high-risk workflows, and having clear steps for dealing with policy violations.

The researcher said he wants to share the method “when the time is right.” For now, the next step in the industry—private testing or public panic—will affect how this story goes.

Related post

“sockpuppeting” can jailbreak 11 AI models like ChatGPT, Claude, and Gemini

Check Also

botnet

Mycelium Framework: First AI-as-a-Service Botnet

A new cybercrime ad is catching attention in the security world. It talks about a …