OpenAppsSubmit

Evidence-backed AMD Strix Halo local-AI setup and benchmarks: Qwen3.8, Ollama, llama.cpp, Vulkan/ROCm, large GGUFs, and cross-OEM results.

Strix Halo Guide on GitHub

Tags

  • ODS V3 Pre-Release: Public testing and refinement ahead of the official V3 launch. Turn your PC, Mac, or Linux box into a private AI server.

    7.2kPythonApache-2.0Self-hosted
    8mo ago
  • GgrunAI & LLMalt. to Ollama

    llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.

    281GoMITSelf-hosted
    7mo ago
  • FoxAI & LLMalt. to Ollama

    A local LLM server built for concurrent work. Drop-in replacement for Ollama and/or OpenAI and Ollama APIs on one port.

    189Rust
    7mo ago
  • Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust.

    32kRustMITWindowsmacOSDesktop
    1.8y ago
  • Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.

    8.5kPythonApache-2.0CLI
    8mo ago
  • Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook

    6.9kSwiftApache-2.0macOSDesktop
    3mo ago