FinRAGBench-V: A Benchmark for Multimodal RAG with Visual Citation in the Financial Domain (EMNLP 2025)
Agent-based Multimodal Urban Moblity Simulator resulting from the ERC MAGnUM project
OpenSIPS + RTPEngine Recording + Speech Recognition in HEP
Voice assistant SDK for iOS devices written in Swift
A simple way to add speech to text functionality to your website :microphone:
A python toolbox for analyzing and plotting free recall data
Speech to text and text to speech Vue library
API playground for Deepgram built with Streamlit
Speech corpora for the speech recognition evaluation system
SDKs and docs for Skit's speech to text service
A mobile web application that helps you convert spoken words to sharable/editable text 🎊
共 40709 条 · 第 1962 / 2036 页