asr
215 个项目 · ⭐ 227.7kgpt_server是一个用于生产级部署LLMs、Embedding、Reranker、ASR、TTS、文生图、图片编辑和文生视频的开源框架。
A Keras CTC implementation of Baidu's DeepSpeech for model experimentation
A list of publically available audio data that anyone can download for ASR or other speech activities
Blazing fast whisper turbo for ASR (speech-to-text) tasks
Obsidian plugin to create high-quality transcriptions from markdown linked audio files
A text-to-speech and speech-to-text server compatible with the OpenAI API, supporting Whisper, FunASR, Bark, and CosyVoice backends.
Input text from speech in any Linux window, the lean, fast and accurate way, using whisper.cpp OFFLINE. Speak with local LLMs via llama.cpp.
VietASR - Vietnamese Automatic Speech Recognition
SOVA ASR (Automatic Speech Recognition)
An opensource speech-to-text software written in tensorflow
Simplified diarization pipeline using some pretrained models - audio file to diarized segments in a few lines of code
[EMNLP Main '25] LiteASR: Efficient Automatic Speech Recognition with Low-Rank Approximation
共 215 条 · 第 5 / 11 页