text-to-speech
291 个项目 · ⭐ 865.8kMultilingual TTS model with voice cloning and duration control, based on T5Gemma encoder-decoder LLM
LLMVoX: Autoregressive Streaming Text-to-Speech Model for Any LLM
Sayna is a unified Voice Layer for AI Agents with a seemless integration to an existing agentic frameworks
Implementation of Spear-TTS - multi-speaker text-to-speech attention network, in Pytorch
Golang framework to build an AI that can understand and speak back to you, and everything else you want
PlayHT Python SDK - AI Text-to-Speech Streaming & Voice Cloning API
A text-to-speech and speech-to-text server compatible with the OpenAI API, supporting Whisper, FunASR, Bark, and CosyVoice backends.
This plugin integrates Azure Speech Cognitive Services in Unreal Engine.
MGM-Omni: Scaling Omni LLMs to Personalized Long-Horizon Speech
共 291 条 · 第 6 / 15 页