4-hour window · 96 points · 100.0% coverage
- glm-5.3-flash is the strongest model at 93.51 token/s average throughput, ahead of minimax-m3 (87.80 token/s) and gemma4:31b (87.16 token/s); nemotron-3-ultra is the weakest at 1.96 token/s average, never exceeding 3.72 token/s.
- deepseek-v4-pro shows the most operationally significant volatility, with a 57.9% coefficient of variation, a peak of 92.28 token/s at 14:20 UTC, and a collapse to 8.89 token/s at 18:00 UTC; glm-5.3-flash also declined 39.0% across the window.
- No missing-data limitation applies: all eight models recorded 12 of 12 samples, 96 valid points total, and 100.0% coverage over the four-hour window.