llm-benchmarks
githubdrose.io

inferencenet Provider Benchmarks


Measured over the last 30 days

Models Tracked
2
Avg Tokens / Second
30.45
Avg Time to First Token (ms)
1555.00
Last Updated
Sep 13, 2026

All inferencenet models

2 live · one row per model · fastest first · 2 folded behind their current lane
ModelAvg Toks/SecMinMaxAvg TTF (ms)
Schematron V2 Small38.2016.7054.501270.00
Schematron V2 Turbo22.709.0032.601840.00

Frequently Asked Questions

Which inferencenet model is fastest?
Based on recent tests, Schematron V2 Small shows the highest average throughput among tracked inferencenet models.
How many recent measurements feed this dashboard?
This provider summary aggregates 8 individual prompts measured across 6 monitoring runs over the past month.