llm-benchmarks
githubdrose.io

Gemma 4 26B A4B Benchmarks


Measured on makora

6 runs
Avg Tokens / Second
43.10
Decode Tokens / Second
—
Avg Time to First Token (ms)
10600.00
Runs Analysed
6
Last Updated
Oct 10, 2026, 07:23 PM

Throughput distribution

Throughput over time

Gemma 4 26B A4Bmakora49tok/s

Generated throughput, including reasoning tokens — so a thinking model's line does not collapse; the table ranks on visible tokens · one cell per model, showing its best-covered provider · models with at least 4 samples in the window first, fastest of those at the top · shared vertical scale, 0 to 92 tok/s (99th percentile) · dashed rule is the model's own mean

Every lane measuring this model

1 measured · live first · a lane is a provider, a catalogue id and a transport
ProviderLaneAvg Toks/SecMinMaxAvg TTF (ms)n
makoravia OpenRouter43.101.0487.7010600.006

Frequently Asked Questions

How fast is Gemma 4 26B A4B?
The latest rolling average throughput is 43.10 tokens per second with an average time to first token of 10600.00 ms across 6 recent runs.
How often are these benchmarks updated?
Benchmarks refresh automatically whenever the monitoring cron runs. The most recent run completed on Oct 10, 2026, 07:23 PM.

Related models