← 返回专题广场
phi-3-vision
3 个项目 · ⭐ 3.0k1
streamline the fine-tuning process for multimodal models: PaliGemma 2, Florence-2, and Qwen2.5-VL
Python
⭐ 2.7k
⑂ 222
Apache-2.0
· 5 天前推送
5 天前
最近推送
2
Phi-3, -3.5, and -4 for Mac: Locally-run Vision and Language Models for Apple Silicon
Jupyter Notebook
⭐ 278
⑂ 22
MIT
· 22 天前推送
22 天前
最近推送
3
Chat with Phi 3.5/3 Vision LLMs. Phi-3.5-vision is a lightweight, state-of-the-art open multimodal model built upon datasets which include - synthetic data and filtered publicly available websites - with a focus on very high-quality, reasoning dense data both on text and vision.
Jupyter Notebook
⭐ 33
⑂ 11
MIT
· 2025-01-02推送
2025-01-02
最近推送