LLM Reference

LLaVA Llama 2 7B

Released
2023-04-17
Last refreshed
2026-04-19
Status
Researched 244d ago
Open weightsCommercial use: non-commercial

LLaVA Llama 2 7B is released 2023-04-17 in the LLaVA family with open-weight; evaluate it while provider pricing coverage matures.

Use it for

  • Teams evaluating general LLM work

Do not use it for

  • Cost-sensitive launches that need sourced token pricing
  • Vision or document-understanding workloads
  • Strict JSON or tool-calling flows
Specifications
Family
LLaVA
Released
2023-04-17
Parameters
7B
Architecture
Decoder Only
Knowledge cutoff
2023-01
Specialization
general
Openness
Open weights
License
CC-BY-NC-4.0Commercial use: non-commercial
Weights
Unknown
Code
Unknown
Training
Fine-tuned
Created by

Academic researcher focused on vision models

N/A
Founded N/A
Website
Pricing

No tracked provider token pricing is available yet.

About

LLaVA, short for Large Language and Vision Assistant, is a multimodal AI model that integrates the Llama 2 7B language model with a vision encoder, often using CLIP, through a projection matrix or multilayer perceptron (MLP). This combination empowers LLaVA to handle both textual and visual data, enabling tasks like visual question answering, image captioning, optical character recognition (OCR), and multimodal dialogue. Its training involves a two-stage process: feature alignment pre-training followed by fine-tuning on multimodal instruction-following data.

Top use-case fit

No primary decision-task fit is mapped for this model yet.

Provider price ladder

No tracked provider token pricing is available for this model yet.

Capabilities

No model capability flags are currently sourced.

Benchmark peer barsfor Coding

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.