Local LLM Testing & Benchmarking for Apple Silicon
Tags
Similar apps
More in this categoryRun a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest experts…
453SwiftMITmacOSDesktopNative MLX runtime for Laya typed decision models — 7–14 ms short decisions on M3 Max. No text generation, PyTorch, or cloud API.
6.9kPythonApache-2.0High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon.
1.6kPythonApache-2.0macOSDesktopFine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
1.4kPythonApache-2.0macOSDesktopYour iPhone helps your Mac run a 27B model: faster prompt reading and more context over a USB-C cable
920PythonMITmacOSDesktopiOS