oMLX logo

oMLX

Mac LLM server that cuts agent wait times from 90s to 5s

▲ 92 Product Hunt votes Open Source Launched Aug 30, 2026
Visit oMLX →

Mac LLM server that cuts agent wait times from 90s to 5s

oMLX turns your Mac into a full LLM inference server, run from the menu bar. It serves text, vision, OCR, embedding and reranker models with continuous batching, plus a RAM+SSD tiered KV cache that survives restarts, so Claude Code and Cursor respond in about 5s instead of 90s. OpenAI and Anthropic compatible APIs drop straight in. Native Swift, not Electron. Apache 2.0, open source.

oMLX launched on Product Hunt on 2026-08-30 and has 92 upvotes and 18 comments so far.

This tool is open source.

Categories: Design AI
View on Product Hunt · Listing last verified: Aug 31, 2026

More Design AI tools

rhun Audryo slash-editor Moxie