04:00
2026-09-17
machinebrief.com
artificial-intelligence
Voice of Reason: Reinforcement Learning for Spoken Math
Applying reinforcement learning with verifiable rewards to the GLM-4-Voice speech model raised free-form accuracy on the GSM8K math benchmark to 74.8%, a new state of the art for speech-native models,…