4-hour window · 96 points · 100.0% coverage
- glm-5.3 is the strongest model with an average throughput of 146.23 token/s, peaking at 186.3 token/s at 22:00 UTC, while nemotron-3-ultra is the weakest at 35.57 token/s average, never exceeding 94.35 token/s.
- The most operationally significant volatility comes from nemotron-3-ultra, whose coefficient of variation is 86.1%, swinging between 6.13 and 94.35 token/s and trending up 155.9% over the window; gemma4:31b also fell from a 173.46 token/s peak to 55.4 token/s at 22:00, a 23.4% decline.
- No missing-data limitation applies: all eight models have 12 of 12 samples, valid_point_count is 96, and coverage is 100.0%, so the four-hour window is fully populated.