oMLX and Rapid-MLX runtimes

Parent: Mac local LLMs: oMLX, Rapid-MLX and related internals · Published reference · snapshot 2026-10-05

↓ Facts as markdownall context files

oMLX engine layers: FastAPI HTTP layer, BatchedEngine / VLMBatchedEngine over mlx-lm BatchGenerator, an EnginePool that holds several models with LRU eviction, pinning and per-model TTL, and a per-model settings store in `~/.omlx/settings.json` (profiles, alias, chat-template kwargs, model-type o...

These notes link each claim to its source. A source may be a research report hosted on this site rather than the primary document. A published reference means the content is available; it does not certify independent review or accuracy.Read the editorial policy and follow the sources before relying on a claim.

Facts

Children

← the whole tree · 3D view· how to read this page