04:00
2026-07-28
arxiv.org
artificial-intelligence
Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining
A study from arXiv (2607.23175v1) presents the first comparative evaluation of training-free methods for aligning language models to user-specific toxicity sensitivities, finding that all methods reduβ¦