speech
141 个项目 · ⭐ 372.2kUltra fast and portable Parakeet implementation for on-device inference in C++ using Axiom with MPS+Unified Memory
Working online speech recognition based on RNN Transducer. ( Trained model release available in release )
T-one is a high-performance streaming ASR pipeline for Russian, specialized for the telephony domain.
A Keras CTC implementation of Baidu's DeepSpeech for model experimentation
A list of publically available audio data that anyone can download for ASR or other speech activities
Python API & command-line tool to easily transcribe speech-based video files into clean text
This plugin integrates Azure Speech Cognitive Services in Unreal Engine.
OSSSpeechKit offers a native iOS Speech wrapper for AVFoundation and Apple's Speech.
SOVA ASR (Automatic Speech Recognition)
[EMNLP Main '25] LiteASR: Efficient Automatic Speech Recognition with Low-Rank Approximation
Timething is a library for aligning text transcripts with their audio recordings.
Russian text normalization pipeline for speech-to-text and other applications based on tagging s2s networks
共 141 条 · 第 4 / 8 页