39981
2026-03-19
最近推送
39982
Pytorch Implementation of CLIP-Lite | Accepted at AISTATS 2023
Python
⭐ 14
⑂ 2
MIT
· 2023-03-18推送
2023-03-18
最近推送
39983
PegasusX: The Future of Multimodal Embeddings 🦄 🦄
Python
⭐ 14
⑂ 6
Apache-2.0
· 2024-10-17推送
2024-10-17
最近推送
39984
2025-10-15
最近推送
39985
5 小时前
最近推送
39986
3 天前
最近推送
39987
It is the implementation of paper "Multi-Modal Sarcasm Detection in Twitter with Hierarchical Fusion Model"
Jupyter Notebook
⭐ 14
⑂ 5
MIT
· 2023-04-18推送
2023-04-18
最近推送
39988
2024-06-12
最近推送
39990
Code for the paper [SIGIR'26]“Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation”
Python
⭐ 14
⑂ 0
MIT
· 2025-12-06推送
2025-12-06
最近推送
39991
14 天前
最近推送
39992
Summit Vitals: Multi-Camera and Multi-Signal Biosensing at High Altitudes
Python
⭐ 14
⑂ 1
· 2024-10-22推送
2024-10-22
最近推送
39993
3 天前
最近推送
39994
An Interactive Game-based Vision Planning benchmark
Python
⭐ 14
⑂ 0
· 2025-02-24推送
2025-02-24
最近推送
39995
2025-12-30
最近推送
39996
[ICRA26] OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
Python
⭐ 14
⑂ 1
Apache-2.0
· 2026-02-24推送
2026-02-24
最近推送
39997
2026-06-03
最近推送
39998
2025-04-19
最近推送
39999
2023-10-07
最近推送
40000
A 5-way embedding model for text, audio, image, video, and 3D point clouds.
Python
⭐ 14
⑂ 4
NOASSERTION
· 2025-11-13推送
2025-11-13
最近推送
共 40707 条 · 第 2000 / 2036 页