speech-synthesis
107 个项目 · ⭐ 322.3kICASSP 2022: "Text2Video: text-driven talking-head video synthesis with phonetic dictionary".
Stream-Omni is a GPT-4o-like language-vision-speech chatbot that simultaneously supports interaction across various modality combinations.
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API. https://www.npmjs.com/package/msedge-tts
Multilingual TTS model with voice cloning and duration control, based on T5Gemma encoder-decoder LLM
Access the latest AI models like ChatGPT, LLaMA, Deepseek, Diffusion, Hugging face, and beyond through a unified prompt layer and performance evaluation
React hooks for Speech Recognition and Speech Synthesis
This plugin integrates Azure Speech Cognitive Services in Unreal Engine.
共 107 条 · 第 3 / 6 页