LLM Reference

LLaVA Models by Haotian Liu

Haotian LiuCC-BY-NC-4.0Open weights
4 models2023Up to 4k ctx

Last refreshed 2026-04-19. Next refresh: weekly.

Details

ResearcherHaotian Liu
Commercial useCommercial use: non-commercial
Models4
Released2023
Max context4k

Capabilities

Vision1 of 4 models
Multimodal1 of 4 models

About

LLaVA, or Large Language and Vision Assistant, is an advanced family of open-source large multimodal models (LMMs) developed by a collaborative team from the University of Wisconsin-Madison, Microsoft Research, and Columbia University 126. These models uniquely integrate a vision encoder, such as CLIP ViT-L/14, with large language models like Vicuna, Mistral, and Nous-Hermes to enable robust visual and language understanding 126. A key innovation of LLaVA models is their end-to-end training process, enriched with GPT-4 generated multimodal instruction-following data to optimize performance 12. The evolution of LLaVA models includes LLaVA-1.5, which added an MLP vision-language connector and academic task-oriented data, and LLaVA-NeXT (1.6), which improved image resolution and broadened LLM support 6. Prioritizing data efficiency, these models are highly accessible for research purposes 12.

Decision facts

Best fit
vision and multimodal workcodingchatbot and role-playing use cases
Capability starting point
LLaVA 13B with 4k context and multimodal inputs
Lowest tracked input
Not tracked
Closest related family
LLaVA 1.5

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

4 in view

Use when the workload needs 13B parameters.

2023-0413B parameters

Use when the workload needs 13B parameters.

2023-0413B parameters

Use when the workload needs 7B parameters.

2023-047B parameters
LLaVA 13BCurrent

Use when the workload needs 4k context, 13B parameters, and multimodal inputs.

2023-044k context13B parametersmultimodal inputs

Release Timeline

1 release group
2023-04
4 current
LLaVA 13B
4k context13B parametersmultimodal inputs
Current
LLaVA Llama 2 13B
13B parameters
Current
LLaVA Llama 2 7B
7B parameters
Current
LLaVA Vicuna 13B
13B parameters
Current

Specifications(4 models)

LLaVA model specifications comparison
ModelReleasedContextParametersVisionMultimodal
LLaVA Vicuna 13B2023-0413BNoNo
LLaVA Llama 2 13B2023-0413BNoNo
LLaVA Llama 2 7B2023-047BNoNo
LLaVA 13B2023-044k13BYesYes

Available From(1 provider)