4-hour window · 96 points · 100.0% coverage
- glm-5.3 is the strongest model at 127.99 token/s average throughput (peak 169.43 token/s), while nemotron-3-ultra is the weakest at 4.87 token/s average, never exceeding 10.46 token/s.
- The most operationally significant volatility is gemma4:31b, which swung between 27.97 and 173.46 token/s (CV 42.4%); deepseek-v4-pro also dropped to 9.58 token/s at 06:00 UTC before recovering to 108.17 token/s, and glm-5.3-flash spiked to 166.43 token/s at 07:00.
- No missing-data limitation exists: all eight models have 12 of 12 samples, 96 valid points, and 100.0% coverage across the four-hour window.