04:00
2026-08-14
arxiv.org
artificial-intelligence
The "Knowledge-Behavior Gap" in Cultural Taboo Safety of Large Language Models
Researchers introduced CulShield, the first public benchmark for evaluating cultural taboo safety in large language models, covering 77 countries and territories with over 2,020 taboos. Testing on advโฆ