haotian-liu/LLaVA
[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond. observed · 2026-08-28
Health v2 · maintenance only
20/100
- Activity 0
- Release rhythm 8
- Longevity 88
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 1234
- days_rel: n/a
- days_push: 751
- n_releases_24m: 0
Adoption not part of the score
25000 stars · 2778 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
LLaVA (Large Language and Vision Assistant) is an open-source multimodal large language model framework implementing visual instruction tuning, combining vision encoders with LLMs like LLaMA for image understanding and conversation. It provides training code, pretrained checkpoints, and a demo toward GPT-4V-level vision-language capabilities.
Use cases
- build a chatbot that understands images
- fine-tune a vision-language model on custom data
- run multimodal LLM inference locally
- ask questions about images with an open-source model
- train a multimodal model with visual instruction tuning
- evaluate large multimodal models on benchmarks
When to choose
- you need an open-source, self-hosted alternative to GPT-4V for image+text tasks
- you want to fine-tune or extend a proven vision-language model
- you need pretrained multimodal checkpoints with permissive Apache-2.0 licensing
When to avoid
- you only need text-only LLM capabilities
- you lack GPU resources for training or inference
- you need a fully managed multimodal API rather than self-hosted models
Facets
library · maturity active
llm-inference llm-training machine-learning deep-learning chatbot transformers large-language-models computer-vision artificial-intelligence deep-learning python multimodal vision-language-model visual-instruction-tuning gpt-4v foundation-models image-understanding natural-language-processing gpu linux docker
1 source
- readme: https://github.com/haotian-liu/LLaVA · fetched 2026-08-28 · 8b8b4b4ea008
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| haotian-liu/LLaVA | main | 20 |
For agents
markdown · JSON · MCP: product_card(name="haotian-liu/LLaVA")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem