speech-to-text
986 个项目 · ⭐ 742.1kA nearly-live implementation of OpenAI's Whisper, using sounddevice. Requires existing Whisper install.
Browser-only Canva-style presentation studio, powered by local Web AI.
An API to transcribe audio with OpenAI's Whisper Large v3!
Проект для распознавания речи на русском языке на основе pykaldi.
Python Kaldi speech recognition with grammars that can be set active/inactive dynamically at decode-time
Self-hosted, OpenAI-compatible AI gateway for private RAG, natural-language data access, and tool-calling agents.
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API. https://www.npmjs.com/package/msedge-tts
Simple self-hosted web application, which can be used to convert audio to subtitles by OpenAI's Whisper model
共 986 条 · 第 11 / 50 页