stt
132 个项目 · ⭐ 103.4k🗣 An overlay that gets your user’s voice permission and input as text in a customizable UI
A text-to-speech and speech-to-text server compatible with the OpenAI API, supporting Whisper, FunASR, Bark, and CosyVoice backends.
VietASR - Vietnamese Automatic Speech Recognition
SOVA ASR (Automatic Speech Recognition)
Voice notes for iPhone and macOS - 100% Rust, Dioxus, local-first (SQLite + LanceDB + RIG)
Automatic Speech Recognition in Unity using Vosk library
🗣️ Real‑time, low‑latency voice, vision, and conversational‑memory AI assistant built on LiveKit and local LLMs
Speech-to-text and keyboard input captions for OBS.
Voice-driven writing, input, and cross-app work for your desktop.
ChunkFormer: Masked Chunking Conformer For Long-Form Speech Transcription
共 132 条 · 第 3 / 7 页