21:56
2026-07-28
promptcube3.com
ai-safety
Mask2Shield: Hardening LLMs Against Neuron-Pruning Attacks
A new defense method called Mask2Shield (M2S) reduces successful neuron-pruning attacks on large language models from 279 to as few as 1 out of 313 prompts by distributing safety logic across the netwβ¦