Tencent WorkBuddy Bench – Agentic Coding Leaderboard
Tencent's WorkBuddy Bench leaderboard shows no single model dominates agentic coding tasks, with Claude Opus 4.8 leading five of eight scored columns, GLM-5.2 leading two, and GPT-5.5 leading one. The…