Wafer pushes GLM-5.2 Fast onto AMD as Nvidia inference costs bite
Inference startup Wafer claims it achieved 80% of Nvidia B200 performance on AMD MI355X GPUs running GLM-5.2, with 2.6x cheaper hardware, by optimizing the software stack. The benchmark, published by Wafer, shows 2626 to…