ContextQA - The open-source tool for data-driven conversations
Fine-tune the newly released Llama-3.2 lightweight models.
Access different AI models in a one place
Minimal, async-first Python framework for production LLM apps- 2 hard deps, no magic, no SaaS.
This repository provides core code for managing large volumes of video footage, enabling content understanding, automatic tagging, and vector database storage. It integrates multimodal models and LLMs for accurate descriptions and semantic search. A web interface allows visualization.
Ensemble Integration: a customizable pipeline for generating multi-modal, heterogeneous ensembles
FinRAGBench-V: A Benchmark for Multimodal RAG with Visual Citation in the Financial Domain (EMNLP 2025)
A python toolbox for analyzing and plotting free recall data
API playground for Deepgram built with Streamlit
SDKs and docs for Skit's speech to text service
This voice assistant is buit in VS Code. It has an ability to understand human speech, process it and provide relevant, requested output to the user. When the user speaks out any appropriate trigger words, the virtual assistant gets activated to serve the user's command.
Local-first, capability-aware, traceable audio and video transcription Skill for Codex
共 7255 条 · 第 334 / 363 页