25521
2026-03-19
最近推送
25522
Pytorch Implementation of CLIP-Lite | Accepted at AISTATS 2023
Python
⭐ 14
⑂ 2
MIT
· 2023-03-18推送
2023-03-18
最近推送
25523
PegasusX: The Future of Multimodal Embeddings 🦄 🦄
Python
⭐ 14
⑂ 6
Apache-2.0
· 2024-10-17推送
2024-10-17
最近推送
25524
2025-10-15
最近推送
25525
4 天前
最近推送
25526
2 天前
最近推送
25527
It is the implementation of paper "Multi-Modal Sarcasm Detection in Twitter with Hierarchical Fusion Model"
Jupyter Notebook
⭐ 14
⑂ 5
MIT
· 2023-04-18推送
2023-04-18
最近推送
25528
2024-06-12
最近推送
25530
Code for the paper [SIGIR'26]“Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation”
Python
⭐ 14
⑂ 0
MIT
· 2025-12-06推送
2025-12-06
最近推送
25531
12 天前
最近推送
25532
Summit Vitals: Multi-Camera and Multi-Signal Biosensing at High Altitudes
Python
⭐ 14
⑂ 1
· 2024-10-22推送
2024-10-22
最近推送
25533
2 天前
最近推送
25534
An Interactive Game-based Vision Planning benchmark
Python
⭐ 14
⑂ 0
· 2025-02-24推送
2025-02-24
最近推送
25535
2025-12-30
最近推送
25536
[ICRA26] OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
Python
⭐ 14
⑂ 1
Apache-2.0
· 2026-02-24推送
2026-02-24
最近推送
25537
2026-06-03
最近推送
25538
2025-04-19
最近推送
25539
2023-10-07
最近推送
25540
A 5-way embedding model for text, audio, image, video, and 3D point clouds.
Python
⭐ 14
⑂ 4
NOASSERTION
· 2025-11-13推送
2025-11-13
最近推送
共 26060 条 · 第 1277 / 1303 页