BREAKING
NVIDIA claims GB300 up to 25x efficiency
Judge AI infra as a Pareto curve
Throughput vs responsiveness
Throughputper MW
Max total tokens
Whole facility
Responsivenessper user
Faster replies
Single user
0
B300 GPUs
0
Grace CPUs
0TB/s
NVLink
0PFLOPS
FP4 Tensor
Per-watt gains vs Hopper by model
DeepSeek V4 Pro25
GLM5.120
Kimi K2.610
Real gains stay workload-dependent
AI NEWS BLITZ
NVIDIA says its new GB300 NVL72 delivers up to 25 times better efficiency than Hopper.