OpenAppsSubmit

FluidAudio

by FluidInference

Audio & MusicmacOSDesktopiOS

Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization.

Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.

FluidAudio on GitHub

Tags

  • Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP…

    21kPythonMIT
    3.9y ago
  • Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet…

    15kC++Apache-2.0WindowsmacOSLinux
    4.1y ago
  • Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

    15kJupyter NotebookApache-2.0AndroidiOS
    7.1y ago
  • A PyTorch-based Speech Toolkit

    12kPythonApache-2.0
    6.5y ago
  • Spotify, native and fast. One lightweight Rust app for your whole library, local playback, and Spotify Connect on Linux, macOS, and Windows.

    7kRustMITWindowsmacOSLinux
    1mo ago
  • Hold a key, speak, release — AI-polished text appears at your cursor in any app. Open-source voice input for macOS & Windows. (按住快捷键说话,松开即得润色后的文字)

    3.8kRustAGPL-3.0WindowsmacOSLinux
    6mo ago