<!-- llms-explorer concept facts · https://llms-explorer.com/tree/head-to-head-top-40-versus-full-vocabulary-kld-o/ · pack 2026-10-05 · ~269 tokens -->

# Head-to-head top-40 versus full-vocabulary KLD on one model and text

> Already covered and still open: localbench-top-40-union-kl-estimator-versus-kld.md records that no run computing both estimators on one model and text exists; the llama.cpp perplexity README describes only the full-vocabulary `--kl-divergence-base` / `--kl-divergence` path and no top-k option, so...

Parent: [Mac local LLMs: Quantization evaluation](https://llms-explorer.com/tree/mac-local-llms-quantization-evaluation/) · 1 facets · 1 facts · page: https://llms-explorer.com/tree/head-to-head-top-40-versus-full-vocabulary-kld-o/

## Facts

- Already covered and still open: localbench-top-40-union-kl-estimator-versus-kld.md records that no run computing both estimators on one model and text exists; the llama.cpp perplexity README describes only the full-vocabulary `--kl-divergence-base` / `--kl-divergence` path and no top-k option, so no head-to-head can be produced from mainline tooling alone. — [source](https://raw.githubusercontent.com/ggml-org/llama.cpp/master/tools/perplexity/README.md)
