Measured over the last 30 days
- Models Tracked
- 2
- Avg Tokens / Second
- 30.45
- Avg Time to First Token (ms)
- 1555.00
- Last Updated
- Sep 13, 2026
All inferencenet models
2 live · one row per model · fastest first · 2 folded behind their current lane| Model | Avg Toks/Sec | Min | Max | Avg TTF (ms) |
|---|---|---|---|---|
| Schematron V2 Small | 38.20 | 16.70 | 54.50 | 1270.00 |
| Schematron V2 Turbo | 22.70 | 9.00 | 32.60 | 1840.00 |
Frequently Asked Questions
Which inferencenet model is fastest?
Based on recent tests, Schematron V2 Small shows the highest average throughput among tracked inferencenet models.
How many recent measurements feed this dashboard?
This provider summary aggregates 8 individual prompts measured across 6 monitoring runs over the past month.