audio-processing
60 个项目 · ⭐ 136.9kA Streamilt web app for music source separation & karaoke
Transcribe and translate audio to text using Whisper and DeepL.
Code and Pretrained Models for ICLR 2023 Paper "Contrastive Audio-Visual Masked Autoencoder".
A library for real-time voice processing in web browsers
Python API & command-line tool to easily transcribe speech-based video files into clean text
A text-to-speech and speech-to-text server compatible with the OpenAI API, supporting Whisper, FunASR, Bark, and CosyVoice backends.
A collection of Audio and Speech pre-trained models.
Your one-stop solution for voice dataset creation
Fine-tuning toolkit for Chatterbox TTS & Chatterbox TURBO models. Supports 23 languages with smart vocabulary extension. Features offline preprocessing, automatic VAD trimming, and voice cloning capabilities. Train custom TTS models with your own dataset in LJSpeech and file-based format.
共 60 条 · 第 2 / 3 页