Models & Infrastructure
Omlx AI
Large Language Models (LLMs)A native macOS inference server built on MLX. Paged SSD KV caching drops agent TTFT from 30-90s to under 5s. OpenAI & Anthropic compatible API for Apple Silicon.
Visit website ↗