4-hour window · 96 points · 100.0% coverage
- gemma4:31b is the strongest model at 122.49 token/s average throughput, ahead of glm-5.3 at 102.07 token/s; nemotron-3-ultra is the weakest at 27.83 token/s average, never exceeding 55.94 token/s.
- deepseek-v4-flash shows the steepest decline at -40.1% trend, and deepseek-v4-pro dropped to 11.76 token/s at 07:00; glm-5.3 and deepseek-v4-flash share the highest volatility at 51.9% coefficient of variation.
- No missing-data limitation applies: all eight models have 12 of 12 samples, valid_point_count is 96, and coverage is 100.0%.