cd/entity/Pliny the Liberator· home entities Pliny the Liberator
grep -l @pliny the liberator /news/*.json | wc -l → 8

Pliny the Liberator

mentions 8 type Person feed RSS

// recent coverage 8 mentions

04:48
2026-07-26
promptcube3.com
ai-safety

AI Red Teaming: From Checkbox to Evidence

AI red teaming must shift from a checkbox exercise to providing verifiable evidence of security testing, according to a guide on AI workflow deployment. Buyers now demand actual test vectors, guardrai…

01:47
2026-07-26
promptcube3.com
large-language-models

Universal Jailbreak: Pliny the Liberator's Latest Claim

Pliny the Liberator claims a 'universal jailbreak' capable of bypassing restrictions across multiple large language models, including Claude, GPT, and Gemini. If the technique holds across different a…

01:27
2026-07-25
twitter.com
ai-safety

Pliny the Liberator claims universal jailbreak of models

A security researcher known as Pliny the Liberator claims to have discovered a universal jailbreak technique effective on all AI models, including heavily guardrailed flagships like Opus 5, GPT-5.6 So…

18:01
2026-06-26
decrypt.co
ai-safety

This AI Agent Survived 6,000 Hack Attempts—Here’s How

Developer Fernando Irarrázaval's AI assistant Fiu survived over 6,000 prompt injection attempts from more than 2,000 attackers without leaking its secrets.env file, though the experiment triggered a G…

20:15
2026-06-24
gilesthomas.com
large-language-models

Thoughts on Role Confusion

Researchers Charles Ye, Jasmine Cui, and Dylan Hadfield-Menell found that large language models often ignore explicit role tags like <system> or <user> and instead infer roles from text tone, enabling…

00:00
2026-06-15
docs.damsecure.ai
artificial-intelligence

Washington Bans Fable. Anthropic Wins Anyway

The U.S. government ordered Anthropic to disable its Fable 5 and Mythos 5 AI models over national security concerns after a hacker bypassed safety guardrails, marking the first time a Western governme…

16:34
2026-06-13
pentesty.co
ai-safety

The Day the US Government Shut Down the Most Powerful AI

On June 12, 2026, the U.S. Department of Commerce ordered Anthropic to suspend access to its advanced AI models Claude Fable 5 and Mythos 5 for all foreign nationals, citing national security export c…

// co-occurs with top 8 entities
// topics top 6 topics