llava
54 个项目 · ⭐ 67.4k[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
SUPIR aims at developing Practical Algorithms for Photo-Realistic Image Restoration In the Wild. Our new online demo is also released at suppixel.ai.
MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.
Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
A C#/.NET library to run LLM (🦙LLaMA/LLaVA) on your local device efficiently.
Eagle: Frontier Vision-Language Models with Data-Centric Strategies
[EMNLP-2024] Build multimodal language agents for fast prototype and production
ChatGPT爆火,开启了通往AGI的关键一步,本项目旨在汇总那些ChatGPT的开源平替们,包括文本大模型、多模态大模型等,为大家提供一些便利
Tag manager and captioner for image datasets
A Framework of Small-scale Large Multimodal Models
Super lightweight Ollama + Qwen Code alternative to run Llama 3.3, DeepSeek-R1, Phi-4, Gemma 3, Mistral Small 3.1 and other large language models.
[CVPR'25 highlight] RLAIF-V: Open-Source AI Feedback Leads to Super GPT-4V Trustworthiness
An open-source implementation for training LLaVA-NeXT.
共 54 条 · 第 1 / 3 页