llm-benchmarks
githubdrose.io

cerebras Provider Benchmarks


Measured over the last 30 days

Models Tracked
2
Avg Tokens / Second
112.35
Avg Time to First Token (ms)
620.00
Last Updated
Sep 25, 2026

All cerebras models

2 live · one row per model · fastest first
cerebras site
ModelAvg Toks/SecMinMaxAvg TTF (ms)
Gemma 4 31B127.0071.40158.00470.00
gpt-oss-120b97.7016.70138.00770.00

Frequently Asked Questions

Which cerebras model is fastest?
Based on recent tests, Gemma 4 31B shows the highest average throughput among tracked cerebras models.
How many recent measurements feed this dashboard?
This provider summary aggregates 484 individual prompts measured across 473 monitoring runs over the past month.