4-hour window · 96 points · 100.0% coverage
- deepseek-v4.1-flash is the strongest model at 157.58 token/s average throughput, peaking at 234.58 token/s; nemotron-3-ultra is the weakest at 21.26 token/s average, never exceeding 37.98 token/s.
- glm-5.2 shows the most volatility, with a 71.6% coefficient of variation and swings from 18.11 to 195.72 token/s, including a 195.72 token/s spike at 04:00 followed by a drop to 18.11 token/s at 04:20; deepseek-v4.1-flash also trended up 66.9% over the window.
- No missing-data limitation exists: all eight models report 12 of 12 samples, valid_point_count is 96, and coverage is 100.0% across the four-hour window.