asr
215 个项目 · ⭐ 227.7kWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
🤖 wukong-robot 是一个简单、灵活、优雅的中文语音对话机器人/智能音箱项目,支持ChatGPT多轮对话能力,还可能是首个支持脑机交互的开源智能音箱项目。
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation, and more!
Multilingual Automatic Speech Recognition with word-level timestamps and confidence
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
共 215 条 · 第 1 / 11 页