oMLX

oMLX

Mac LLM server that cuts agent wait times from 90s to 5s

oMLX turns your Mac into a full LLM inference server, run from the menu bar. It serves text, vision, OCR, embedding and reranker models with continuous batching, plus a RAM+SSD tiered KV cache that survives restarts, so Claude Code and Cursor respond in about 5s instead of 90s. OpenAI and Anthropic compatible APIs drop straight in. Native Swift, not Electron. Apache 2.0, open source.

Open SourceDeveloper ToolsArtificial IntelligenceGitHub
👍 90 💬 17 评论 🕐 2026年8月30日
广告位 (响应式横幅)

💬 用户评论

该工具在 Product Hunt 上有 17 条评论。

在 Product Hunt 上查看评论 →

📋 基本信息

上线日期
2026年8月30日
点赞数
90
评论数
17
数据来源
网络公开数据
广告位 (响应式横幅)

🔗 相关推荐