NVIDIA vs
Zhipu AI
Verdict
- Zhipu AI's flagship leads: GLM-5.3-Flash at 61.0 vs Nemotron 3 Ultra 550B A55B at 58.3.
- Zhipu AI has the better model in 4 of 4 price bands.
- NVIDIA's fastest ranked model (nemotron-3-nano-30b-a3b, 231 tok/s) beats Zhipu AI's (GLM-4.7-Flash, 97 tok/s).
- NVIDIA publishes more open-weights models (32 vs 14).
- NVIDIA ships more often: one model every 26 days against every 28.
Best model per price band
Each provider’s highest-index model below the price ceiling. Purple marks the winner of the band.
| Price band | NVIDIA | Zhipu AI |
|---|---|---|
| Under $0.50/M | Gemma-4-31B-IT 55.4$0.205/M | GLM-5.3-Flash 61.0$0.119/M |
| Under $2/M | Nemotron 3 Ultra 550B A55B 58.3$1.00/M | GLM-5.3-Flash 61.0$0.119/M |
| Under $10/M | Nemotron 3 Ultra 550B A55B 58.3$1.00/M | GLM-5.3-Flash 61.0$0.119/M |
| Any price | Nemotron 3 Ultra 550B A55B 58.3$1.00/M | GLM-5.3-Flash 61.0$0.119/M |
Line-up
| Metric | NVIDIA | Zhipu AI |
|---|---|---|
| Flagship index | 58.3 | 61.0 |
| Ranked models | 19 | 14 |
| Open-weights models | 32 | 14 |
| Cheapest ranked, $/M | $0.063 | $0.119 |
| Fastest ranked, tok/s | 231 | 97 |
| Release cadence, days | 26 | 28 |
| Last release | 11 Aug 2026 | 20 Aug 2026 |
| Next expected | 26 Aug 2026 – 16 Sept 2026 | 5 Sept 2026 – 28 Sept 2026 |