audio-processing
60 个项目 · ⭐ 136.9kPre-training, fine-tuning, and inference code with the MAEST models for music analysis applications.
Repository for the paper "Combining audio control and style transfer using latent diffusion", accepted at ISMIR 2024
📣 Find sentiments, tags, entities, and actions in your voice recordings instantly
Speaker change detection using SincNet and an LSTM/Transformer
stm32-speech-recognition-and-traduction is a project developed for the Advances in Operating Systems exam at the University of Milan (academic year 2020-2021). It implements a speech recognition and speech-to-text translation system using a pre-trained machine learning model running on the stm32f407vg microcontroller.
This repository contains a short introduction on the topic of audio and speech processing -- from basics to applications.
A Comprehensive Speech Processing Algorithms Library for research and production use
An optimized FastAPI server for OpenAI's Whisper whisper-large-v3-turbo model using MLX optimization
共 60 条 · 第 3 / 3 页