4-hour window · 96 points · 100.0% coverage
- deepseek-v4-flash posted the strongest average throughput at 120.09 token/s, ahead of gemma4:31b at 100.31 token/s, while nemotron-3-ultra was weakest at 24.03 token/s, never exceeding 52.53 token/s in any interval.
- glm-5.3-flash showed the most operationally significant volatility, with a coefficient of variation of 52.7% and swings from 43.26 token/s at 18:00 to 183.59 token/s at 17:00 and 182.29 token/s at 18:40; glm-5.3 also declined 30.0% overall, ending at 57.67 token/s.
- No missing-data limitation applies: all eight models recorded 12 of 12 expected samples, totaling 96 valid points at 100.0% coverage for the window.