llm-benchmarks
githubdrose.io

wafer Provider Benchmarks


Measured over the last 30 days

Models Tracked
2
Avg Tokens / Second
15.57
Avg Time to First Token (ms)
10942.35
Last Updated
Sep 1, 2026

All wafer models

2 tracked · fastest first
ModelAvg Toks/SecMinMaxAvg TTF (ms)
DeepSeek V4 Flash 073110.208.9711.306390.00
GLM 5.28.736.9311.707750.00

Frequently Asked Questions

Which wafer model is fastest?
Based on recent tests, DeepSeek V4 Flash 0731 shows the highest average throughput among tracked wafer models.
How many recent measurements feed this dashboard?
This provider summary aggregates 561 individual prompts measured across 491 monitoring runs over the past month.