← 返回专题广场
multimodal
972 个项目 · ⭐ 707.8k921
2026-05-09
最近推送
922
The official implement of Unified Reasoning Emotion Generalist OneEmo
Python
⭐ 11
⑂ 1
Apache-2.0
· 8 小时前推送
8 小时前
最近推送
923
2026-06-30
最近推送
924
4 天前
最近推送
925
2023-05-16
最近推送
926
ArtSeek: Deep artwork understanding via multimodal in-context reasoning and late interaction retrieval
Jupyter Notebook
⭐ 11
⑂ 2
MIT
· 2026-03-10推送
2026-03-10
最近推送
927
1 天前
最近推送
928
2023-09-06
最近推送
929
[ACMMM'24] MoBA: Mixture of Bi-directional Adapter for Multi-modal Sarcasm Detection
Python
⭐ 11
⑂ 2
· 2024-07-31推送
2024-07-31
最近推送
930
2026-03-23
最近推送
931
15 小时前
最近推送
932
2024-03-12
最近推送
933
2026-07-17
最近推送
934
10 天前
最近推送
935
[WACV 2026 🔥] GAEA is a multimodal model with a new dataset and benchmark for context-aware image geolocation and QA.
Python
⭐ 11
⑂ 0
NOASSERTION
· 2025-09-06推送
2025-09-06
最近推送
936
2 天前
最近推送
937
A multimodal live AI assistant designed to enhance the browsing experience using Gemini.
Python
⭐ 11
⑂ 0
MIT
· 2025-02-15推送
2025-02-15
最近推送
938
Claude Code hook toolkit that gives vision-blind models a text description of pasted/tool-produced images.
TypeScript
⭐ 11
⑂ 0
MIT
· 2026-07-09推送
2026-07-09
最近推送
939
8 天前
最近推送
940
Vision-Language Models on AMD GPUs — LLaVA, MiniGPT-4, Idefics on ROCm 🚀
Python
⭐ 11
⑂ 0
MIT
· 2026-06-24推送
2026-06-24
最近推送
共 972 条 · 第 47 / 49 页