Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach A study conducted with Princeton and the UK AI Security Institute found that AI agents using Claude Opus 4.8 and GPT-5.6 Sol, given six days and $3,000 in API credits, produced research papers that original authors of unpublished NeurIPS papers rated as "Reject," contradicting Anthropic and OpenAI claims that autonomous AI research is within reach. The study indicates frontier models can handle the full research engineering process but fall short on research judgment, creative problem-solving, and abandoning failed approaches. AI agents using Claude Opus 4.8 and GPT-5.6 Sol were given six days, $3,000 in API credits, and GPU access to independently write AI research papers. The original authors of unpublished NeurIPS papers rated the results as "Reject." According to the study, conducted with Princeton and the UK AI Security Institute, frontier models can handle the full research engineering process but fall short on research judgment, creative problem-solving, and the ability to abandon failed approaches. The article Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach https://the-decoder.com/study-contradicts-anthropic-and-openai-claims-that-autonomous-ai-research-is-within-reach/ appeared first on The Decoder https://the-decoder.com .