speech-to-text
986 个项目 · ⭐ 742.1k🎧 Automatic Speech Recognition: DeepSpeech & Seq2Seq (TensorFlow)
Mila — native macOS local transcription app (whisper.cpp) with optional speaker diarization. Apache-2.0.
MooER: Moore-threads Open Omni model for speech-to-speech intERaction. MooER-omni includes a series of end-to-end speech interaction models along with training and inference code, covering but not limited to end-to-end speech interaction, end-to-end speech translation and speech recognition.
A testing server for a speech to text service based on coqui.ai
A text-to-speech and speech-to-text server compatible with the OpenAI API, supporting Whisper, FunASR, Bark, and CosyVoice backends.
This plugin integrates Azure Speech Cognitive Services in Unreal Engine.
Official one-stop shop for AI Agents and developers building with Telnyx.
NodeJS Bindings for Whisper - the CPU version of OpenAI's Whisper, as initially crafted in C++ by ggerganov.
Private, local transcription and system-wide dictation for Apple Silicon.
Self-hostable multimodal chat with local LLMs (Ollama/OpenAI): PDF RAG, image chat, and Whisper voice, Streamlit + Docker.
Experimental open-source macOS menubar app for speech-to-text workflows
一款基于 PySide6 和 ElevenLabs API 的桌面应用,能将音视频或JSON转录稿智能地转换为高质量SRT字幕。特别为中、日、韩、英等语言优化了排版规则。
TypeWhisper for Windows - Local speech-to-text with translation
Deep Learning based Automatic Speech Recognition with attention for the Nvidia Jetson.
共 986 条 · 第 14 / 50 页