Find the best models and how to run them locally.
Tags
Similar apps
More in this categoryHigh-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon.
1.6kPythonApache-2.0macOSDesktopFine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
1.4kPythonApache-2.0macOSDesktopRun a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest experts…
453SwiftMITmacOSDesktopllama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.
281GoMITSelf-hostedA local LLM server built for concurrent work. Drop-in replacement for Ollama and/or OpenAI and Ollama APIs on one port.
189Rust