Meta’s new paper exposes why reinforcement learning struggles with code optimization, and how to fix it
Meta AI's FAIR team published a paper on July 29 revealing that reinforcement learning struggles with code optimization due to noisy timing measurements and sparse rewards, and introduced DMC-Optim, a benchmark with a ca…