OpenAppsSubmit

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP…

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

FunASR on GitHub

Open-source alternative to

Tags

  • FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

    6.4kPythonMIT
    3.4y ago
  • A PyTorch-based Speech Toolkit

    12kPythonApache-2.0
    6.5y ago
  • Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

    9.5kCMIT
    2.3y ago
  • Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization.

    3kSwiftApache-2.0macOSDesktopiOS
    1.3y ago
  • Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

    1.6kCApache-2.0
    10mo ago
  • Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.

    391RustMIT
    9mo ago