LLaVA icon
Open Source LLMs

LLaVA

🇨🇳 中文 / 🇬🇧 English — 双语介绍
🇨🇳 中文

LLaVA 是开源的多模态大模型(视觉+语言):把视觉编码器和大语言模型结合,能看图聊天、理解截图和图表,被社区广泛用作本地多模态 AI 的基座。项目由威斯康星大学团队开源。 🔗 GitHub: https://github.com/haotian-liu/LLaVA | ⭐ 25k+

🇬🇧 English

LLaVA is an open-source multimodal model (vision + language): it combines a vision encoder with a large language model to chat about images, screenshots and charts, widely used as the base for local multimodal AI. Open-sourced by a University of Wisconsin team. 🔗 GitHub: https://github.com/haotian-liu/LLaVA | ⭐ 25k+

📋 Product Info

Category: Open Source LLMs
Listed on: 2026-08-22
Official site: https://llava-vl.github.io
Open source: GitHub repository ↗
Tags: 多模态,视觉语言模型,开源,AI研究,大模型

🎯 Why This Tool

LLaVA belongs to the Open Source LLMs category on hedirbase. Tagged with: 多模态,视觉语言模型,开源,AI研究,大模型.

LLaVA is an open-source multimodal model (vision + language): it combines a vision encoder with a large language model to chat about images, screenshots and charts, widely used as the base for local m

✓ Manually reviewed 📁 Open source 🌐 Curated for independents
🌎 访问 LLaVA → 📦 GitHub