How much of SWE-bench Pro can 64 DeepSeek agents solve in a day? A Doubleword inference deployment running 64 independent coding agents powered by DeepSeek-V4-Pro completed all 46,784 attempts across SWE-bench Pro's 731 public problems in 20 hours and 23 minutes on a single 8×B300 node, while the SGLang baseline averaged 10.7 problems per agent and throughput-oriented SGLang averaged 24.8 in 24 hours. The Doubleword deployment delivered roughly 30× the request and token rates of throughput-oriented SGLang, which itself more than doubled the SGLang baseline's throughput. A single agent solved 51.2% of the benchmark on average, and accepting a solution from any of the 64 agents raised that to 70.7%, a 19.5-point improvement. How much of SWE-bench Pro can 64 DeepSeek agents solve in a day? Coding agents can now spend hours on a single task, making hundreds of model calls as they inspect code, edit files and test changes.