Real Time Voice Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
Digital audio workstations, recording and mixing tools.
Clone a voice in 5 seconds to generate arbitrary speech in real-time
The swiss army knife of lossless video/audio editing
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
Background Music, a macOS audio utility: automatically pause your music, set individual apps' volumes and record system audio.
Buzz transcribes and translates audio offline on your personal computer. Powered by OpenAI's Whisper.
A PyTorch-based Speech Toolkit
so-vits-svc fork with realtime support, improved interface and more features.
vits2 backbone with multilingual-bert
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
Block and skip sponsors, while also muting and skipping ads on YouTube.
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
Noise supression using deep filtering