Chinese AI agent outperforms Anthropic’s Claude Code in autonomous research Zhejiang University's Qiushi Engine topped the ResearchClawBench leaderboard, outperforming Anthropic's Claude Code in autonomous research tasks. The benchmark, created by the Shanghai Artificial Intelligence Laboratory, tests AI agents' ability to independently conduct research and compare results against human-written reference papers. Despite its top ranking, Qiushi Engine cannot reliably make new discoveries yet. Chinese AI agent outperforms Anthropic’s Claude Code in autonomous research Zhejiang University’s Qiushi Engine topped the ResearchClawBench leaderboard, but it can’t reliably make new discoveries yet Anthropic’s https://www.scmp.com/topics/anthropic?module=inline&pgtype=article Claude Code and other top agents. Claude Code https://www.scmp.com/news/china/article/3359901/anthropic-hits-back-after-china-warns-claude-code-backdoor-risks?module=inline&pgtype=article in third. ResearchClawBench tests the ability of AI agents to independently carry out research and compares their results against reference papers written by humans to see if they can reach the same conclusions or even outdo the original authors. The benchmark, created by a team led by the Shanghai Artificial Intelligence Laboratory, was designed to assess whether agents can really conduct the kind of tasks their creators say they can handle. Qiushi Engine, which was officially launched by a Zhejiang University-led team last week, is a large language model-based agent designed to perform scientific research in real physical environments. Its developers said that unlike some other existing systems that could be limited to performing specific tasks, Qiushi Engine was capable of “end-to-end autonomous scientific discovery”.