Models & Infrastructure

Omlx AI

Large Language Models (LLMs)

A native macOS inference server built on MLX. Paged SSD KV caching drops agent TTFT from 30-90s to under 5s. OpenAI & Anthropic compatible API for Apple Silicon.

Visit website ↗
Omlx AI website screenshot
Omlx AI website preview
Explore more Models & Infrastructure tools →