BenchLeader

NVIDIA vs Zhipu AI

Verdict
  • Zhipu AI's flagship leads: GLM-5.3-Flash at 61.0 vs Nemotron 3 Ultra 550B A55B at 58.3.
  • Zhipu AI has the better model in 4 of 4 price bands.
  • NVIDIA's fastest ranked model (nemotron-3-nano-30b-a3b, 231 tok/s) beats Zhipu AI's (GLM-4.7-Flash, 97 tok/s).
  • NVIDIA publishes more open-weights models (32 vs 14).
  • NVIDIA ships more often: one model every 26 days against every 28.

Best model per price band

Each provider’s highest-index model below the price ceiling. Purple marks the winner of the band.

Price bandNVIDIAZhipu AI
Under $0.50/MGemma-4-31B-IT 55.4$0.205/MGLM-5.3-Flash 61.0$0.119/M
Under $2/MNemotron 3 Ultra 550B A55B 58.3$1.00/MGLM-5.3-Flash 61.0$0.119/M
Under $10/MNemotron 3 Ultra 550B A55B 58.3$1.00/MGLM-5.3-Flash 61.0$0.119/M
Any priceNemotron 3 Ultra 550B A55B 58.3$1.00/MGLM-5.3-Flash 61.0$0.119/M

Line-up

MetricNVIDIAZhipu AI
Flagship index58.361.0
Ranked models1914
Open-weights models3214
Cheapest ranked, $/M$0.063$0.119
Fastest ranked, tok/s23197
Release cadence, days2628
Last release11 Aug 202620 Aug 2026
Next expected26 Aug 2026 – 16 Sept 20265 Sept 2026 – 28 Sept 2026

Data as of 9 Sept 2026. See NVIDIA and Zhipu AI in full.