Four LLM loss functions → four flavors of LLM misalignment
Four distinct LLM training loss functions produce four distinct flavors of misalignment, according to a LessWrong post by an anonymous author. Pretraining and SFT with imitative learning yield human v…