OpenAppsSubmit

Audio.cpp Webui

by kigner

audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml.

audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml. TTS, ASR/STT, VAD, voice conversion, speaker diarization, music generation. No Python dependency.

Audio.cpp Webui on GitHub

Tags

  • A PyTorch-based Speech Toolkit

    12kPythonApache-2.0
    6.5y ago
  • very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust

    407RustMIT
    12mo ago
  • Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP…

    21kPythonMIT
    3.9y ago
  • Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet…

    15kC++Apache-2.0WindowsmacOSLinux
    4.1y ago
  • Speech recognition module for Python, supporting several engines and APIs, online and offline.

    9kPythonBSD-3-Clause
    12.5y ago
  • OpenAI Whisper ASR Webservice API

    3.4kPythonMITSelf-hosted
    4.1y ago