In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access?
A new study on arXiv (2609.00904v1) finds that large language models (LLMs) do not demonstrate reliable control over privileged internal representations when the control target is not inferable from t…