Mac local LLMs: oMLX, Rapid-MLX and related internals

Parent: Running LLM models locally on a Mac · Published reference · snapshot 2026-10-05

↓ Facts as markdownall context files

oMLX (Apache-2.0, macOS 15+, Python 3.11-3.13, M1-M5): multi-model EnginePool, menu-bar app, /admin dashboard, persistent SSD cache, embeddings and rerank on one endpoint. Install: signed DMG, `brew install jundot/omlx/omlx` (CLI only; `brew tap jundot/omlx` first).

These notes link each claim to its source. A source may be a research report hosted on this site rather than the primary document. A published reference means the content is available; it does not certify independent review or accuracy.Read the editorial policy and follow the sources before relying on a claim.

Choose and install

oMLX memory guard and pressure levels

oMLX store-cache gate and stalled admission

oMLX SSD and hot cache

Hybrid GDN, MTP and boundary snapshots

oMLX engines and Metal stream clears

oMLX pressure reclaim and executors

oMLX health, teardown and wedges

Rapid-MLX, MTPLX, Osaurus

Corrections to earlier claims

Open

Corrections and disagreements

Concepts in this cluster

Children

← the whole tree · 3D view· how to read this page