LLaVA 13B

Released
2023-04-17
Last refreshed
2026-09-30
Status
Researched today
Open weightsCommercial use: non-commercialMultimodalVision
Specifications
Family
LLaVA
Released
2023-04-17
Context
4k
Parameters
13B
Architecture
Decoder Only
Specialization
general
Openness
Open weights
License
CC-BY-NC-4.0Commercial use: non-commercial
Weights
Unknown
Code
Unknown
Training
Fine-tuned
Created by

Academic researcher focused on vision models

N/A
Founded N/A
Website
Pricing
Output / 1M
-
Input / 1M
-

Cheapest of 1 route · Replicate API

About

Original LLaVA (Large Language-and-Vision Assistant) 13B model. Multimodal vision+language model combining a vision encoder with a language model for visual understanding tasks.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
Replicate API--
ServerlessPartial

Capabilities

VisionMultimodal

Benchmark peer barsfor Vision

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.