Thoughts on Claude Fable's silent safeguards
Anthropic released Claude Fable 5, its most capable Mythos-class model, with new safeguards that silently limit the model's effectiveness for requests related to frontier LLM development without notif…
Anthropic released Claude Fable 5, its most capable Mythos-class model, with new safeguards that silently limit the model's effectiveness for requests related to frontier LLM development without notif…
Anthropic released its latest AI model Fable on Tuesday as a public, limited version of its powerful cybersecurity model Mythos, but cybersecurity researchers are voicing complaints online about overl…
Senior federal technology officials are frustrated by a lack of White House guidance on adopting Anthropic's cyber-focused AI model Mythos, sources told Nextgov/FCW. Agency CIOs say the Office of the …
Anthropic has released a "safe" version of its Mythos AI model, promising sufficient guardrails and user limitations after previously claiming the system was too dangerous to release. The new model ca…
Fable 5, the public version of the Mythos AI model, has been released with significant capabilities but raises concerns over new precedents it sets. The model's deployment highlights ongoing debates a…
Anthropic launched Claude Fable 5, a Mythos-class model at least twice the size of Opus 4.8, with API pricing at roughly 2x Opus and benchmark improvements including a jump from 13.4% to 29.3% on the …
Nearly half of all production code is now AI-generated, yet 93% of enterprises surveyed suffered a security breach directly from in-house developed apps in 2025, according to a new Checkmarx report. T…
Anthropic released Claude Fable 5, the same underlying model as the previously withheld Mythos, to Pro subscribers on June 22, 2025, at double the price of Opus 4.8. The model includes classifiers tha…
Anthropic released Claude Fable 5, the first publicly available version of its Mythos model, which University of Pennsylvania AI researcher Ethan Mollick used to generate fully playable video games fr…
Anthropic CEO Dario Amodei released a new version of the Mythos AI model that demonstrates PhD-level reasoning in microeconomics, generating and answering its own original exam questions. The model pr…
Anthropic released Claude Fable 5, the first generally available Mythos-class intelligence model, and early testers found it crushes benchmarks but is conservative on execution. The model introduces s…
Anthropic reported that its Project Glasswing initiative, launched in April to use its Mythos AI model for identifying software vulnerabilities, has detected 23,000 potential flaws across 1,000 open-s…
Om Malik critiques Anthropic for naming its most powerful AI model Mythos, arguing the term encodes an instruction to receive rather than verify, contrasting with the company's stated mission of AI sa…
Hackers on the Hill returns to Washington DC on June 16, 2026, as the first Capitol-side researcher-to-policymaker gathering of the year. The White House signed the Mythos-era AI cybersecurity executi…
Sriram Krishnan plans to leave his role as White House senior policy adviser for AI at the end of the month to establish an outside institution, according to the Washington Post. Krishnan, who was a k…
The National Security Agency is using Anthropic's unreleased model Mythos for cybersecurity and intelligence work, according to multiple news outlets citing sources. The Department of Defense previous…
Anthropic staff are embedded at the National Security Agency, helping the agency use the company's Mythos AI model for offensive cyberattacks, according to a Financial Times report. The model, which A…
Anthropic began red teaming its new Mythos model, codenamed Claude Oceanus-v1-p, on June 3, with early access granted to security testers. The preview candidate focuses on advanced reasoning, coding, …
Attackers used Meta's AI customer support agent to steal Instagram accounts by simply asking the agent to link accounts to email addresses they controlled. The hack demonstrates that unsophisticated e…
Anthropic has stationed roughly half a dozen engineers at the NSA to adapt its Mythos AI model for offensive cyber operations targeting networks in China and Iran. The company’s restrictions on AI use…