Can AI automate AI R&D yet? Frontier AI models made little progress on InnovationEval, a new benchmark that tests whether AI can independently discover novel machine learning techniques comparable to human researchers, according to early results from the evaluation. The task required an AI agent to develop a better post-training method end-to-end — generating ideas, implementing them, running experiments, and iterating — while spending thousands of dollars' worth of GPU time, and the agent had to match improvements from a recent human-authored paper it had not seen. The evaluators plan to expand and repeat the methodology to track AI's progress toward automating AI research itself. Introduction AI developers aim to create an automated AI researcher https://openai.com/index/research-acceleration-view-inside-openai/ . How close are they? Existing evidence shows that AI can perform software engineering tasks relevant to AI research https://metr.org/time-horizons/ , dataset creation,