← 返回专题广场
multimodal-llm
9 个项目 · ⭐ 4.5k1
2026-02-25
最近推送
2
22 天前
最近推送
3
Official implementation of paper "MiniGPT-5: Interleaved Vision-and-Language Generation via Generative Vokens"
Python
⭐ 868
⑂ 52
Apache-2.0
· 2025-05-08推送
2025-05-08
最近推送
4
[ICCV25 Oral] Token Activation Map to Visually Explain Multimodal LLMs
Python
⭐ 189
⑂ 10
· 2025-12-14推送
2025-12-14
最近推送
5
13 天前
最近推送
6
vLLM Qwen 3.6-27B (AWQ-INT4) + DFlash speculative decoding on AMD Strix Halo (gfx1151 iGPU, 128 GB UMA, ROCm 7.13). 24.8 t/s single-stream, vision, tool calling, 256K context, OpenAI-compatible, Docker. Matches DGX Spark FP8+DFlash+MTP at a third of the cost. No CUDA.
Python
⭐ 51
⑂ 4
Unlicense
· 2026-05-10推送
2026-05-10
最近推送
7
2026-04-27
最近推送
8
2026-02-12
最近推送