{"slug": "benchmark-across-claude-opencode-hermes-and-other-coding-agents-on-10-swe-bench", "title": "Benchmark across Claude, OpenCode, Hermes, and other coding agents on 10 SWE-bench tasks", "summary": "Carlo Capocasa published a 10-task benchmark comparing coding agents Claude, OpenCode, Pi, Zcode, Hermes, and 3code on SWE-bench verified tasks, with 3code solving 9 of 10 tasks using 5 million tokens and Pi solving 6 of 10 with fewer tokens. Capocasa cautioned that harness performance varies with token efficiency and task completion rates, advising users to validate results against their own heuristics.", "body_md": "# Benchmark across Claude, OpenCode, Hermes, and other coding agents on 10 SWE-bench tasks\n\nCarlo Capocasa published a 10-task GLM benchmark comparing Claude, OpenCode, Pi, Zcode, Hermes, and 3code on representative SWE-bench verified tasks. 3code solved 9 of 10 tasks using 5 million tokens, while Pi solved 6 of 10 using fewer tokens than other runners-up. Capocasa noted harness performance varies with token efficiency and task completion rates, cautioning that users should validate results against their own heuristics.\n\n## Topics\n\n## Sources\n\n-   Press   [Read article](https://capocasa.dev/10-task-glm-5-3-harness-bench-claude-opencode-pi-zcode-hermes-and-3code)  \n\n## Go deeper\n\nThis intelligence is sourced automatically from public sources across the web and synthesised by the Prefactor AI pipeline. Stories are reviewed before publication.", "url": "https://wpnews.pro/news/benchmark-across-claude-opencode-hermes-and-other-coding-agents-on-10-swe-bench", "canonical_source": "https://www.getreadyforagents.com/news/coding-agent-harness-benchmark-10-task/", "published_at": "2026-09-07 20:05:22+00:00", "updated_at": "2026-09-07 20:30:24.326107+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-tools", "ai-agents"], "entities": ["Carlo Capocasa", "Claude", "OpenCode", "Pi", "Zcode", "Hermes", "3code"], "alternates": {"html": "https://wpnews.pro/news/benchmark-across-claude-opencode-hermes-and-other-coding-agents-on-10-swe-bench", "markdown": "https://wpnews.pro/news/benchmark-across-claude-opencode-hermes-and-other-coding-agents-on-10-swe-bench.md", "text": "https://wpnews.pro/news/benchmark-across-claude-opencode-hermes-and-other-coding-agents-on-10-swe-bench.txt", "jsonld": "https://wpnews.pro/news/benchmark-across-claude-opencode-hermes-and-other-coding-agents-on-10-swe-bench.jsonld"}}