4-hour window · 84 points · 87.5% coverage
- gemma4:31b is the strongest model with an average throughput of 123.88 token/s (peaking at 173.39 token/s), while nemotron-3-ultra is the weakest at 42.92 token/s average, ranging from 5.24 to 89.59 token/s.
- The most significant volatility comes from glm-5.2 and nemotron-3-ultra: glm-5.2 fell from 108.13 token/s at 08:20 to 5.89 token/s at 12:00 (a -23.3% trend), and nemotron-3-ultra shows a 55.8% coefficient of variation with a -37.2% trend, ending at 19.72 token/s.
- deepseek-v4-flash reported zero samples out of 12 expected, contributing to overall coverage of 87.5% (84 valid points), so its throughput cannot be assessed for this window.