Project Context Clues aim to vectorize CyberOps recon and threat intel data for AI driven systems.
Build Fullstack GenAI apps and AI Agents
Agent-based Multimodal Urban Moblity Simulator resulting from the ERC MAGnUM project
YesBut - Multimodal Satire Comprehension Dataset
[IEEE T-BIOM] FaceXBench: Evaluating Multimodal LLMs on Face Understanding
小红书图文笔记 Agent Skill,支持看图理解、主动提问、智能选题、实时热梗搜索、标题正文、标签评论钩子和风险检查。
Code for "MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching", TIP2025
Implementation of "With a Little Help from my Temporal Context: Multimodal Egocentric Action Recognition, BMVC, 2021" in PyTorch
ABC: Achieving Better Control of Multimodal Embeddings using VLMs [TMLR2025]
React component and hook for speech recognition, with voice commands and pluggable STT engines
MCP server for transcript processing — formatting, contextual repair & smart summarization with deep-thinking LLMs
HTK Toolkit with Linux 64 bit and Docker support
共 26086 条 · 第 1251 / 1305 页