20:28
2026-08-17
cryptobriefing.com
artificial-intelligence
New benchmark reveals AI agents follow complex instructions less than 30% of the time
A new benchmark called AGENTIF, developed by researchers at Tsinghua University and Zhipu AI, reveals that current advanced AI models achieve less than 30% perfect instruction following on complex, reβ¦