04:00
2026-08-11
arxiv.org
artificial-intelligence
Scaling Inherently Interpretable Language Models
A new arXiv paper challenges the notion that interpretability comes at the cost of capability, showing that making interpretability a training constraint allows it to scale with model capability acrosβ¦