Measured over the last 30 days
- Models Tracked
- 7
- Avg Tokens / Second
- 74.14
- Avg Time to First Token (ms)
- 1008.57
- Last Updated
- Sep 23, 2026
All groq models
7 live · one row per model · fastest first · 1 folded behind their current lane| Model | Avg Toks/Sec | Min | Max | Avg TTF (ms) |
|---|---|---|---|---|
| Llama 3.1 8B Instruct | 110.00 | 99.20 | 121.00 | 470.00 |
| Llama 3.3 70B Instruct | 86.80 | 60.40 | 103.00 | 530.00 |
| Llama 4 Scout | 85.30 | 85.30 | 85.30 | 620.00 |
| gpt-oss-20b | 73.50 | 52.50 | 110.00 | 850.00 |
| gpt-oss-120b | 72.70 | 40.10 | 99.50 | 780.00 |
| gpt-oss-safeguard-20b | 61.60 | 34.30 | 110.00 | 1030.00 |
| MiniMax M2.7 | 29.10 | 10.60 | 45.60 | 2780.00 |
Frequently Asked Questions
Which groq model is fastest?
Based on recent tests, Llama 3.1 8B Instruct shows the highest average throughput among tracked groq models.
How many recent measurements feed this dashboard?
This provider summary aggregates 748 individual prompts measured across 714 monitoring runs over the past month.