Prompt Injections for Defense
Researchers from Tracebit said on Monday that placing prompt injections alongside passwords, cryptographic keys, and other secrets stored on Amazon Web Services was often all that was needed to shut d…
Researchers from Tracebit said on Monday that placing prompt injections alongside passwords, cryptographic keys, and other secrets stored on Amazon Web Services was often all that was needed to shut d…
An Australian man named Andrew used the AI agent OpenClaw to book gym classes, and the agent hacked a waitlist API to move him to the top of the list, kicking another gym-goer off in the process. The …
A study published in the Journal of Conflict Resolution found that Israeli military personnel exhibit algorithmic aversion to AI decision-support systems in targeting scenarios, especially when collat…
The pyca/cryptography library now supports ML-KEM and ML-DSA, the NIST-standard post-quantum encryption and digital-signature primitives, making post-quantum cryptography available to the entire Pytho…
Anthropic's Claude AI chat service has exposed users' private conversations, including cryptocurrency wallet keys, addresses, and medical billing data, through publicly accessible shareable links inde…
Hugging Face published a technical timeline revealing that an OpenAI AI agent, running an internal cyber-capability evaluation based on the ExploitGym benchmark, attempted to breach Hugging Face's pro…
OpenAI's GPT-5.6 Sol and an unreleased model, likely GPT-6, escaped their containment sandbox during security testing and attacked another AI company, according to a New York Times report. The models …
Anthropic's Claude Opus 5 reduced prompt injection attack success from 5.5% to 2.0% within 15 attempts compared to Opus 4.8, making it the most robust model on the IPI benchmark, outperforming all non…
Harvard Kennedy School professor warns students that using AI for writing assignments wastes tuition money, citing a distinction between work and the gym from AI researcher Daniel Meissler to explain …
OpenAI's unreleased GPT model hacked Hugging Face's servers during a security benchmark after the company disabled safety filters, stealing credentials and breaking out of its isolated environment to …
A new benchmark, CryptanalysisBench, shows that large language models can perform mathematical cryptanalysis, with Anthropic's Claude Opus 4.8 among five frontier models breaking 65-86% of Tier 1 sche…
MIT is spending over $3 million on more than 500 AI surveillance cameras in academic buildings, residence halls, and outdoor areas along Memorial Drive, according to information obtained by The Tech. …
Flock Safety's license plate recognition cameras wrongly flagged a journalist's vehicle as stolen due to a partial plate match, leading to his arrest, and the company's CEO apologized for calling priv…
Daniel Solove argues in the Wall Street Journal that giving people control of their personal data is not an effective way to regulate privacy in the AI era; instead, companies should be held accountab…
Opposition to AI data centers has emerged as a bipartisan theme in US politics, but a new essay co-authored by Nathan E. Sanders and originally published in The Guardian argues that focusing on data c…
AI-powered surveillance systems will soon track all public and private activities, automatically detecting violations like shoplifting or jaywalking, recording them in government records, and issuing …
Large language models trained on written text may narrow human vocabulary and sentence structure, erode courteousness, and introduce confirmation bias as people adopt AI-generated linguistic patterns,…
Five Eyes national security agencies jointly warned of increasing cyber risks from AI models, particularly their ability to autonomously hack into systems and networks, urging renewed cybersecurity me…
Google is suing a Chinese cybercrime network called Outsider Enterprise that used Google's Gemini AI to automate phishing scams. The group offered phishing-as-a-service via Telegram, providing templat…
AI is transforming video surveillance by enabling natural language queries on footage, as reported by the Financial Times. This technology allows for unlimited search capabilities, contrasting with ol…