Mac local LLMs: Agent clients, context and compaction

Parent: Running LLM models locally on a Mac · Published reference · snapshot 2026-10-05

↓ Facts as markdownall context files

Needs only ANTHROPIC_BASE_URL + ANTHROPIC_AUTH_TOKEN and a /v1/messages server: Ollama >=0.14 (`ollama launch claude --model <m>`), LM Studio >=0.4.1, oMLX, Rapid-MLX, llama-server; no proxy needed.

These notes link each claim to its source. A source may be a research report hosted on this site rather than the primary document. A published reference means the content is available; it does not certify independent review or accuracy.Read the editorial policy and follow the sources before relying on a claim.

Connect Claude Code to a local server

Subagent model routing (Claude Code)

Windows and compaction (Claude Code)

Server overflow behaviour

Slimming the first request

OpenCode

Codex

Open questions

Corrections and disagreements

Concepts in this cluster

Children

← the whole tree · 3D view· how to read this page