← 返回专题广场
vision
81 个项目 · ⭐ 227.7k41
5 小时前
最近推送
42
2025-09-20
最近推送
43
2026-06-07
最近推送
44
19 小时前
最近推送
45
11 小时前
最近推送
46
Code for NOLA, an implementation of "nola: Compressing LoRA using Linear Combination of Random Basis"
Python
⭐ 59
⑂ 4
MIT
· 2024-08-26推送
2024-08-26
最近推送
47
Transform screenshots into searchable Obsidian notes using AI vision and text analysis
TypeScript
⭐ 57
⑂ 1
AGPL-3.0
· 2025-04-02推送
2025-04-02
最近推送
48
[CVPR 2024] The official implementation of paper "synthesize, diagnose, and optimize: towards fine-grained vision-language understanding"
Jupyter Notebook
⭐ 52
⑂ 0
· 2025-06-16推送
2025-06-16
最近推送
49
Vision Transformers Needs Registers. And Gated MLPs. And +20M params. Tiny modality gap ensues!
Python
⭐ 47
⑂ 1
MIT
· 2025-06-03推送
2025-06-03
最近推送
50
给 DeepSeek 补上「眼睛和耳朵」的多模态视觉插件:看图 / OCR / 物体检测 / 视频理解 / 语音转写 / 截图直读,一键安装(DSH 插件)。
PowerShell
⭐ 41
⑂ 3
MIT
· 6 天前推送
6 天前
最近推送
52
Stable Diffusion with Text-to-Image and Image-to-Text
Jupyter Notebook
⭐ 38
⑂ 6
· 2023-06-24推送
2023-06-24
最近推送
53
2026-07-16
最近推送
54
2025-12-14
最近推送
55
16 天前
最近推送
56
9 天前
最近推送
57
2025-09-19
最近推送
59
Search knowledge by what documents mean and how they look — not one or the other.
Python
⭐ 25
⑂ 2
Apache-2.0
· 28 天前推送
28 天前
最近推送
60
Awesome-HCI (Ubiquitous, LLM, MLLM, Agent, RAG, Embodied-AI, RLHF)
Python
⭐ 25
⑂ 2
· 2026-03-08推送
2026-03-08
最近推送
共 81 条 · 第 3 / 5 页