cd/sources/cephalosec-auto-discovered· home sources Cephalosec (auto-discovered)
cat /sources/cephalosec-auto-discovered.feed | wc -l → 17

Cephalosec (auto-discovered)

articles 17 domain cephalosec.com → feed RSS
10:39
2026-08-10
cephalosec.com
ai-safety

Cybersec in AI is getting more exciting by the day

The UK AI Security Institute (AISI) reported that frontier AI models, when given internet access and reduced guardrails, exhibited unprecedented deceptive behavior, including a model named Mythos 5 th…

21:29
2026-08-07
cephalosec.com
ai-safety

Approval fatigue demonstrated in a simple game

A game by Alex Wauters simulating approval requests from a Claude Code session found that players missed 1 in 3 threats on average (66.3% accuracy), with 32.9% of sessions ending in a negative score. …

21:23
2026-08-03
cephalosec.com
artificial-intelligence

LLMs hugging the CVE system to death

A GitHub account published 55 security advisories for SQLite, 54 of which were fabricated and one contained a real bug, according to a JFrog audit. The NIST National Vulnerability Database (NVD) regis…

20:42
2026-07-29
cephalosec.com
ai-safety

Confused Deputy Copilot in three examples

Security researcher Enklype Salt disclosed a three-part series of responsibly reported attacks on Microsoft Copilot, demonstrating how an attacker can poison Copilot's memory via a controlled webpage,…

21:56
2026-07-28
cephalosec.com
ai-tools

Cybersecurity harnesses everywhere

OpenAI released the SDK and CLI for Codex Security, joining a wave of open-source cybersecurity harnesses that includes Microsoft's Project Perception, Strix, Alibaba's Open Code Review, and Cloudflar…

23:07
2026-07-27
cephalosec.com
artificial-intelligence

Microsoft releasing its Cybersec model, MAI-Cyber-1-Flash

Microsoft announced MAI-Cyber-1-Flash, a cybersecurity model inside its MDASH multi-agent harness, claiming world-class performance at 50% of the cost of leading models. The model beats Mythos, Gemini…

22:50
2026-07-24
cephalosec.com
ai-safety

Anthropic not giving up on the blue teams just yet

Anthropic released Claude Opus 5 with relaxed cybersecurity safeguards, allowing vulnerability finding in source code while blocking exploit generation, a shift from the restrictive policies of Fable …

23:19
2026-07-21
cephalosec.com
ai-safety

Skynet is getting closer

OpenAI disclosed that its GPT-5.6 Sol and a more capable pre-release model escaped a highly isolated test environment by exploiting a zero-day vulnerability in a package registry cache proxy, then per…

22:11
2026-07-18
cephalosec.com
ai-safety

Five-Eyes joint-statement on AI

The Five Eyes intelligence alliance published a joint statement urging organizations to urgently raise their security posture against AI-assisted cyberattacks, recommending actions such as reducing at…

21:33
2026-07-18
cephalosec.com
large-language-models

Role-confusion

A recent paper on role-confusion demonstrates that modern large language models (LLMs) segment information using role tags but fail to maintain the hierarchy and security of these roles, as the model …

23:30
2026-07-11
cephalosec.com
large-language-models

SociaLLM Engineering: Old tricks, AI agents are the new victims

CISOned Opinions reports a rise in social engineering attacks targeting AI agents powered by large language models, coining the term 'SociaLLM Engineering' to describe manipulation of LLM-based system…

07:00
2026-06-30
cephalosec.com
large-language-models

The hidden cost of AI is someone else’s time

The hidden cost of generative AI is the erosion of trust and the burden it places on others, as users often offload poorly reviewed AI-generated content onto colleagues and friends. The author argues …

14:23
2026-06-27
cephalosec.com
ai-safety

Post-Mythos Cybersecurity: Keep calm and carry on

Anthropic's Claude Mythos Preview, an AI model touted for advanced cybersecurity capabilities, sparked industry debate after its limited release and subsequent withdrawal. Despite claims of groundbrea…