Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python.
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Zig backend, Swift frontend macOS app with chat, music, voice, video generation.
Tags
Similar apps
More in this categoryRapid-MLX is an open-source (Apache 2.0) OpenAI- and Anthropic-compatible LLM inference server for Apple Silicon, built on MLX, focused on reliable tool…
4kPythonmacOSDesktopHigh-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon.
1.6kPythonApache-2.0macOSDesktopRun any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command.
281TypeScriptSelf-hostedWebBash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
78kPythonMITLocal UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
78kPythonApache-2.0Self-hostedYour AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research.
38kPythonAGPL-3.0Self-hosted