OpenAppsSubmit

Turbo Fieldfare

by drumih

AI & LLMmacOSDesktop

Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook

Turbo Fieldfare on GitHub

Tags

  • Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest experts…

    453SwiftMITmacOSDesktop
    1mo ago
  • The fastest way to run Qwen 3.8 Flash Next, Qwen 3.8 27B and Ternary Bonsai 2 27B on a Mac: 125 tok/s in OpenCode on an M5 Max, and a 27B model on 16 GB Macs.

    2.6kPythonApache-2.0macOSDesktop
    5mo ago
  • Visual AI agent workflow automation platform with local LLM integration - build intelligent workflows using drag-and-drop interface, no cloud dependencies…

    178TypeScriptSelf-hosted
    1.2y ago
  • Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust.

    32kRustMITWindowsmacOSDesktop
    1.8y ago
  • ODS V3 Pre-Release: Public testing and refinement ahead of the official V3 launch. Turn your PC, Mac, or Linux box into a private AI server.

    7.2kPythonApache-2.0Self-hosted
    8mo ago
  • Native MLX runtime for Laya typed decision models — 7–14 ms short decisions on M3 Max. No text generation, PyTorch, or cloud API.

    6.9kPythonApache-2.0
    21d ago