speech-to-text
986 个项目 · ⭐ 742.0kSpeech recognition module for Python, supporting several engines and APIs, online and offline.
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Silero Models: pre-trained text-to-speech models made embarrassingly simple
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
开源 AI 视频本地化工具:自动完成 YouTube/Bilibili 视频下载、字幕识别与翻译、语音克隆配音、音轨混合和字幕压制。
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.
The python library for real-time communication
On-device subtitle generation that connects directly to DaVinci Resolve, Premiere, and After Effects.
Lightweight and powerful real-time audio/speech translation tool based on Windows LiveCaptions.
共 986 条 · 第 2 / 50 页