05:13
2026-09-01
forum.level1techs.com
artificial-intelligence
Does anyone know how the vericoding benchmark ran GLM-4.5 and DeepSeek V3.1? Asking because one of them looks like the non-thinking variant
A developer working on the Beneficial AI Foundation's vericoding benchmark questions whether the open-weight models GLM-4.5 and DeepSeek V3.1 were run with their thinking modes enabled, noting that thβ¦