speech-to-text
986 个项目 · ⭐ 742.2kA WebRTC-native, audio-first conversational-AI framework for Go.
Local-first, open-source YouTube dubbing: clones the original voice into another language, time-synced. Chatterbox voice cloning + faster-whisper + NLLB, with multi-voice, burned subtitles and optional Wav2Lip lip-sync. Runs free on your own machine via a simple CLI.
Telegram bot with voice message recognition and generation. Speech to Text and Text to Speech
A self-hosted meeting transcription app that doesn't need to join your meetings as a bot.
Wyoming protocol server for Microsoft Azure speech-to-text
An efficient implementation of RNN-T Prefix Beam Search in C++/CUDA.
Syn.Speech is a flexible speaker independent continuous speech recognition engine for Mono and .NET framework
Tool for automatic transcription and speaker diarization based on whisper and pyannote.
Fix the 2-5 second push-to-talk activation delay on macOS. Keeps microphone hardware awake for instant voice transcription with AirPods, Bluetooth headsets, and built-in mic. Works with SuperWhisper, WhisperFlow, Wispr Flow, and any push-to-talk app on Apple Silicon (M1/M2/M3/M4).
Android native AI inference library, bringing text, image, video, STT, TTS inference
共 986 条 · 第 26 / 50 页