16:16
2026-08-10
lesswrong.com
artificial-intelligence
Four LLM loss functions โ four flavors of LLM misalignment
Four distinct LLM training loss functions produce four distinct flavors of misalignment, according to a LessWrong post by an anonymous author. Pretraining and SFT with imitative learning yield human vโฆ