GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia
Z.ai releases GLM-5.3-Flash, a 320B open-source model that scores within three points of GLM-5.3 at one-seventh the cost, running entirely on Chinese AI chips rather than Nvidia hardware.
Z.ai releases GLM-5.3-Flash, an open-source model with 320 billion parameters that lands just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index, at a seventh of the cost. What's notable is that all of the inference traffic ran on Chinese AI chips instead of Nvidia hardware.
The article GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia appeared first on The Decoder.