04:00
2026-09-28
machinebrief.com
ai-safety
User Model Extraction via Belief Self-Distillation
A new arXiv paper (2609.31603v1) introduces Belief Self-Distillation (BSD), a read-write framework that learns a compact user representation from a frozen LLM's own natural conversations and can both …