speech-to-text
986 个项目 · ⭐ 742.2kCLI for audio, video, and text transcription with ASR providers and LLM-powered summarization via local or cloud backends.
The Web AI Toolkit is a powerful, privacy-first JavaScript library that brings advanced AI capabilities directly to your web applications. Run OCR, speech-to-text, text summarization, image classification, and more — all locally in the browser with no data sent to external servers.
End-to-end subtitle translation workstation with cloud and local OpenVINO model support.
A SpeechToText application that uses OpenAI's whisper via faster-whisper to transcribe audio and send that information to VRChats textbox system and/or KillFrenzyAvatarText over OSC. Also supports various other methods like OBS via Browsersource and a SteamVR overlay!
Transformer based Bangla Speech Recognition | Encoder Decoder Architecture
Experiments to test different speech recognition systems for SEPIA Framework
a free, customizable, osc capable speech-to-text interface for relaying text to different types of applications
Windows 本地 AI 语音输入:Qwen3-ASR CPU/GPU 双运行形态、本地屏幕 OCR 上下文、热词纠错与可选 AI 润色。
A voice recognition-based tool for translating languages in real-time.
☕🇧🇷 Scripts para o Kaldi em Português Brasileiro
共 986 条 · 第 27 / 50 页