00:59
2026-07-18
lesswrong.com
artificial-intelligence
The Most Forbidden Technique is not always forbidden
Goodfire announced a private beta of Silico, its LLM training platform, and reproduced RLFR, a method using probes as reward signals for reinforcement learning. The announcement sparked debate on Twitβ¦