audio
145 个项目 · ⭐ 772.1kA video composition framework build on top of AVFoundation. It's simple to use and easy to extend.
Model compression toolkit engineered for enhanced usability, comprehensiveness, and efficiency.
🗜️Compress Image, Video, and Audio same like Whatsapp 🚀✨
Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine
The Hugging Face Course on Transformers for Audio
HuggingSound: A toolkit for speech-related tasks based on Hugging Face's tools
Transcribe and translate audio to text using Whisper and DeepL.
Joint speech-language model - respond directly to audio!
LLMVoX: Autoregressive Streaming Text-to-Speech Model for Any LLM
Code and Pretrained Models for ICLR 2023 Paper "Contrastive Audio-Visual Masked Autoencoder".
Radient turns many data types (not just text) into vectors for similarity search, RAG, regression analysis, and more.
llama.cpp (GGUF LLMs) and llava.cpp (GGUF VLMs) for ROS 2
Golang framework to build an AI that can understand and speak back to you, and everything else you want
共 145 条 · 第 5 / 8 页