OpenAppsSubmit

Backburner

by StayLameBro

AI & LLMmacOSDesktopiOS

Your iPhone helps your Mac run a 27B model: faster prompt reading and more context over a USB-C cable

Backburner on GitHub

Tags

  • A local inference engine for Apple silicon, built around the model.

    1.3kC++Apache-2.0macOSDesktop
    21d ago
  • Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest experts…

    453SwiftMITmacOSDesktop
    1mo ago
  • The fastest way to run Qwen 3.8 Flash Next, Qwen 3.8 27B and Ternary Bonsai 2 27B on a Mac: 125 tok/s in OpenCode on an M5 Max, and a 27B model on 16 GB Macs.

    2.6kPythonApache-2.0macOSDesktop
    5mo ago
  • Up to 4× faster LLM decoding on Apple Silicon, lossless. Native MLX port of DeepSeek's DSpark & z-lab's DFlash speculative decoding — Gemma-4, Qwen3.8,…

    702PythonMITmacOSDesktop
    3mo ago
  • HumlaAI & LLMalt. to Granola

    Open-source AI meeting notes for Mac. Records mic + system audio with no bot, transcribes on-device or via OpenAI / Deepgram / Groq, identifies speakers…

    303RustMITmacOSDesktop
    6mo ago
  • Local LLM Testing & Benchmarking for Apple Silicon

    210SwiftGPL-3.0macOSDesktop
    8mo ago