4-hour window · 84 points · 87.5% coverage
- Highest average throughput was gemma4:31b at 108.26 token/s across 12 samples, peaking at 159.92 token/s; lowest was nemotron-3-ultra at 19.05 token/s, never exceeding 27.42 token/s.
- glm-5.3-flash showed the sharpest trend, rising 147.3% from roughly 50 token/s early in the window to 151.0–160.09 token/s by 16:40–17:00 UTC, while glm-5.2 was the most volatile at 61.3% coefficient of variation, swinging between 11.12 and 162.78 token/s.
- deepseek-v4-flash returned 0 of 12 expected samples, so its throughput is unknown; overall coverage is 87.5% (84 of 96 expected points), limiting fleet-wide conclusions for the period.