OpenAppsSubmit

Phonon

by fermionresearch

Audio & MusicCLISelf-hosted

Phonon: open speech recognition models (Phonon-2, Phonon-1) — CLI, CPU and CUDA images

Phonon on GitHub

Tags

  • OpenAI Whisper ASR Webservice API

    3.4kPythonMITSelf-hosted
    4.1y ago
  • Fast, cross-platform CLI and GUI for batch transcription, translation, speaker annotation and subtitle generation using OpenAI’s Whisper on CPU, Nvidia GPU and…

    162PythonMITCLI
    2.4y ago
  • VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation…

    56kPythonAGPL-3.0Desktop
    6mo ago
  • A PyTorch-based Speech Toolkit

    12kPythonApache-2.0
    6.5y ago
  • Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.

    1.7kC++MPL-2.0LinuxDesktop
    5y ago
  • Optimized Whisper models for streaming and on-device use

    899PythonMIT
    12mo ago