llm-benchmarks
githubdrose.io

groq Provider Benchmarks


Measured over the last 30 days

Models Tracked
6
Avg Tokens / Second
195.83
Avg Time to First Token (ms)
0.00
Last Updated
Aug 9, 2026

All groq models

6 tracked · fastest first
groq site
ModelAvg Toks/SecMinMaxAvg TTF (ms)
GPT-oss-safeguard-20b293.0028.10684.000.00
llama-3.1-8b219.0074.40363.000.00
Qwen3.6-27B218.0011.20326.000.00
qwen-3-32b174.0061.90241.000.00
llama-3.3-70b150.0019.20252.000.00
llama-4-scout121.0019.80198.000.00

Frequently Asked Questions

Which groq model is fastest?
Based on recent tests, GPT-oss-safeguard-20b shows the highest average throughput among tracked groq models.
How many recent measurements feed this dashboard?
This provider summary aggregates 2830 individual prompts measured across 2435 monitoring runs over the past month.