GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia Z.ai released GLM-5.3-Flash, an open-source model with 320 billion parameters that scores just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index, at a seventh of the cost. All inference traffic for the model ran on Chinese AI chips instead of Nvidia hardware. GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia By Maximilian SchreinerSource: The Decoder https://the-decoder.com Z.ai releases GLM-5.3-Flash, an open-source model with 320 billion parameters that lands just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index, at a seventh of the cost. What's notable is that all of the inference /glossary/inference traffic ran on Chinese AI chips instead of Nvidia /glossary/nvidia hardware. The article GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia https://the-decoder.com/the-chinese-ai-model-glm-5-3-flash-runs-without-nvidia-and-costs-a-fraction-of-what-the-competition-does/ appeared first on The Decoder https://the-decoder.com . Get AI news in your inbox Daily digest of what matters in AI.