audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml.
audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml. TTS, ASR/STT, VAD, voice conversion, speaker diarization, music generation. No Python dependency.
Tags
Similar apps
More in this categoryvery fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
407RustMITOpen-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP…
21kPythonMITSpeech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet…
15kC++Apache-2.0WindowsmacOSLinuxSpeech recognition module for Python, supporting several engines and APIs, online and offline.
9kPythonBSD-3-Clause