OpenAppsSubmit
  • Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest experts…

    453SwiftMITmacOSDesktop
    1mo ago
  • Find the best models and how to run them locally.

    247TypeScriptMIT
    7mo ago
  • Native MLX runtime for Laya typed decision models — 7–14 ms short decisions on M3 Max. No text generation, PyTorch, or cloud API.

    6.9kPythonApache-2.0
    21d ago
  • High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon.

    1.6kPythonApache-2.0macOSDesktop
    10mo ago
  • Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.

    1.4kPythonApache-2.0macOSDesktop
    9mo ago
  • Your iPhone helps your Mac run a 27B model: faster prompt reading and more context over a USB-C cable

    920PythonMITmacOSDesktopiOS
    9d ago