25741
--
综合分
25742
--
综合分
25743
--
综合分
25744
--
综合分
25745
--
综合分
25746
--
综合分
25747
--
综合分
25748
--
综合分
25749
Self-hosted multimodal AI workspace — chat, vision QA, text-to-image, image-to-image in one conversation
Python
⭐ 12
⑂ 0
NOASSERTION
· 11 小时前推送
--
综合分
25750
Multimodal and multilingual topic model with pretrained embeddings
HTML
⭐ 12
⑂ 1
MIT
· 2023-04-11推送
--
综合分
25751
[NeurIPS2023] LoRA: A Logical Reasoning Augmented Dataset for Visual Question Answering
Jupyter Notebook
⭐ 12
⑂ 1
· 2024-01-06推送
--
综合分
25752
Unified-modal Salient Object Detection via Adaptive Prompt Learning
Python
⭐ 12
⑂ 1
· 2026-07-17推送
--
综合分
25754
[Reproduce] Code for the EMNLP2018 paper "A Visual Attention Grounding Neural Model for Multimodal Machine Translation".
Python
⭐ 12
⑂ 3
Apache-2.0
· 2020-01-19推送
--
综合分
25755
Exploring Visual Interpretability for Contrastive Language-Image Pretraining
⭐ 12
⑂ 0
· 2022-09-24推送
--
综合分
25757
--
综合分
25758
Batch LLM Inference with Ray Data LLM: From Simple to Advanced
Jupyter Notebook
⭐ 12
⑂ 4
MIT
· 2026-02-13推送
--
综合分
25759
--
综合分
25760
An out-of-tree vLLM plugin for Mobilint NPU runtime integration.
Python
⭐ 12
⑂ 1
Apache-2.0
· 21 天前推送
--
综合分
共 26058 条 · 第 1288 / 1303 页