The Implications of Linguistic Illegibility for LLM Security
A paper submitted to arXiv on 2 Sep 2026 introduces the term "linguistic illegibility" to describe scenarios in which an LLM's externalized or mechanistically-probed language artifacts fail to represent how the model act…