Overtraining as the path to human-like AI
Anonymous blogger Gwern argues in a 13,000-word post that large language models (LLMs) lack truly flexible human-like intelligence because they fail to 'grok' — a process where overtraining past the point of memorization…