llm-benchmarks
githubdrose.io

GLM 5 Turbo Benchmarks


Measured on z-ai

5 runs
Avg Tokens / Second
9.06
Avg Time to First Token (ms)
12500.00
Runs Analysed
5
Last Updated
Sep 1, 2026, 04:08 AM

Throughput distribution

Throughput over time

GLM 5 Turboz-ai1tok/s

Generated throughput, including reasoning tokens — so a thinking model's line does not collapse; the table ranks on visible tokens · one cell per model, showing its best-covered provider · shared vertical scale, 0 to 32 tok/s (99th percentile) · dashed rule is the model's own mean · 1 further provider not drawn

Every provider serving this model

1 measured
ProviderModelAvg Toks/SecMinMaxAvg TTF (ms)
z-aiGLM 5 Turbo9.061.5514.7012500.00

Frequently Asked Questions

How fast is GLM 5 Turbo?
The latest rolling average throughput is 9.06 tokens per second with an average time to first token of 12500.00 ms across 5 recent runs.
How often are these benchmarks updated?
Benchmarks refresh automatically whenever the monitoring cron runs. The most recent run completed on Sep 1, 2026, 04:08 AM.

Related models