← All models

glm-5.2

756.2B parameters 1,048,576 token context thinkingcompletiontools
74.9
output tokens / second
160 tokens · 2.14s server · 2.28s request

Performance statistics

Calculated live from valid twenty-minute RRDtool samples.

Range Latest Average Lowest Median P50 P05 Highest Std dev Variability CV Trend Data coverage
Last 24 hours 73.7 64.5 4.8 50.6 10.9 218.1 52.2 81.0% -18.0% 72 samples 100.0%
Last 7 days 73.7 62.5 2.4 50.8 16.3 239.9 43.8 70.1% -2.1% 504 samples 100.0%
Last 30 days 73.7 62.3 2.4 60.8 20.3 239.9 32.0 51.3% -8.7% 2160 samples 100.0%

P05: 5% of twenty-minute throughput samples are at or below this value; it represents lower-tail performance. CV: standard deviation ÷ average; lower is more stable. Trend: second-half average versus first-half average. Data coverage: valid samples versus all expected twenty-minute slots in the selected range.

Last 24 hours

glm-5.2 throughput for Last 24 hours