Measured over the last 30 days
- Models Tracked
- 2
- Avg Tokens / Second
- 17.58
- Avg Time to First Token (ms)
- 12570.32
- Last Updated
- Sep 1, 2026
All open-inference models
2 tracked · fastest first| Model | Avg Toks/Sec | Min | Max | Avg TTF (ms) |
|---|---|---|---|---|
| Gemma 4 31B | 24.60 | 5.50 | 38.20 | 2580.00 |
| DeepSeek V4 Flash 0731 | 1.59 | 1.59 | 1.59 | 980.00 |
Frequently Asked Questions
Which open-inference model is fastest?
Based on recent tests, Gemma 4 31B shows the highest average throughput among tracked open-inference models.
How many recent measurements feed this dashboard?
This provider summary aggregates 425 individual prompts measured across 371 monitoring runs over the past month.