llama.cpp reasoning-preserve flag and template support detection

Parent: Mac local LLMs: llama.cpp internals · Published reference · snapshot 2026-10-05

↓ Facts as markdownall context files

Ollama has no preserve flag. Its Qwen3.5 renderer decides per turn: an assistant turn gets a `<think>` block only when thinking is on and the turn comes after the last real user query, or when the renderer variant forces it.

These notes link each claim to its source. A source may be a research report hosted on this site rather than the primary document. A published reference means the content is available; it does not certify independent review or accuracy.Read the editorial policy and follow the sources before relying on a claim.

Facts

Corrections and disagreements

Children

← the whole tree · 3D view· how to read this page