whisper
386 个项目 · ⭐ 456.0kDocker image for a self-hosted WhisperLive real-time speech-to-text server, powered by faster-whisper. Provides WebSocket streaming for live audio transcription and an OpenAI-compatible REST API. Supports all Whisper models, VAD, NVIDIA GPU (CUDA) acceleration, offline mode, and multi-arch (amd64, arm64).
Voice memos recorded from the microphone, transcribed offline to text and converted to Joplin notes
PAFTS : Library That Preprocessing Audio For TTS.
Ear is a desktop app that will help you transcribe what is playing on your computer!
A stand-alone application with GUI for OpenAI's Whisper
This is an OpenAI Whisper automatic speech recognition microservice
Enterprise-grade browser extension bringing multilingual voice interaction to AI chatbots (Pi, Claude, ChatGPT). Features real-time speech detection with Silero VAD, accurate transcription via OpenAI Whisper, and ElevenLabs TTS. Built with TypeScript, XState, and modern web standards. Progressive enhancement across Chrome, Firefox, and Safari
共 386 条 · 第 16 / 20 页