17:00
2026-10-09
infoq.com
artificial-intelligence
Android Bench 2 Adds Support for Long-Horizon Tasks, Agentic Evaluation, and Continuous Scoring
Google released Android Bench 2.0, adding long-horizon tasks (LHTs), agentic evaluation, and continuous scoring to its benchmark for AI models and agents on Android development tasks. The update replaβ¦