<!-- llms-explorer concept facts · https://llms-explorer.com/tree/prefill-parity-gating-of-prefill-tokens-per-seco/ · pack 2026-10-05 · ~287 tokens -->

# Prefill-parity gating of prefill tokens-per-second in cross-runtime reports

> Prefill-parity gating of prefill tokens-per-second is already covered: llm-benchpacks computes a per-case parity status from prompt tokens and cached prompt tokens and prints prefill_tps only for comparable cases, as recorded in llm-benchpacks-workload-level-agent-benchmarks.md and time-to-first-...

Parent: [Mac local LLMs: Benchmarking and comparisons](https://llms-explorer.com/tree/mac-local-llms-benchmarking-and-comparisons/) · 1 facets · 1 facts · page: https://llms-explorer.com/tree/prefill-parity-gating-of-prefill-tokens-per-seco/

## Facts

- Prefill-parity gating of prefill tokens-per-second is already covered: llm-benchpacks computes a per-case parity status from prompt tokens and cached prompt tokens and prints prefill_tps only for comparable cases, as recorded in llm-benchpacks-workload-level-agent-benchmarks.md and time-to-first-token-measurement-artifacts.md; the gate's inputs per server are listed in prompt-token-counting-parity-across-runtimes.md. — [source](https://raw.githubusercontent.com/ephes/llm-benchpacks/main/docs/decisions.md)
