speech-to-text
986 个项目 · ⭐ 742.1kOpen-source AI meeting copilot - real-time transcription, echo cancellation, and AI assistance. Captures system audio + mic, cancels echo via WebRTC AEC3, transcribes with Deepgram, and gives you Claude/OpenAI help during meetings. Runs locally on macOS and Windows.
Transcribe and translate audio to text using Whisper and DeepL.
Stream-Omni is a GPT-4o-like language-vision-speech chatbot that simultaneously supports interaction across various modality combinations.
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
A live speech recognition using Facebooks wav2vec 2.0 model.
Speech-to-text in Obsidian using Whisper
Speech-to-text input for Claude Code with live streaming dictation
A lightweight Python package for Automatic Speech Recognition using ONNX models
共 986 条 · 第 10 / 50 页