04:00
2026-09-30
arxiv.org
large-language-models
Learning from the Gap Between Pass@K and Pass@1
A new post-training method called GapFT improves single-sample decoding accuracy of Llama-3.1-8B by 14.4 points on LogiQA 2.0 and 13.9 points on ReClor over the source model, according to an arXiv pap…