LLaVA 1.6 Vicuna 13B
LLaVA 1.6 Vicuna 13B is worth evaluating for general LLM work when its provider route and context window match the workload.
Use it for
- Teams evaluating general LLM work
- Workloads that can use a 4k context window
- Buyers comparing 1 tracked provider route
Do not use it for
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
- Family
- LLaVA 1.6
- Released
- 2024-01-31
- Context
- 4k
- Parameters
- 13B
- Architecture
- Decoder Only
- Knowledge cutoff
- 2023-01
- Specialization
- general
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Unknown
- Code
- Unknown
- Training
- Fine-tuned
Cheapest of 1 route · Replicate API
About
LLaVA 1.6 Vicuna 13B is a sophisticated multimodal language model designed to handle multimodal chatbot tasks, integrating text and image processing seamlessly. It features a pre-trained LLM, Vicuna-13B, and a likely CLIP ViT-L/14 vision encoder, linked using a trainable projection matrix, allowing it to comprehend both textual and visual content efficiently. The model offers capabilities such as image captioning, visual question answering, and enhanced reasoning and OCR, with the added advantage of processing high-resolution images up to 672x672 pixels.
Top use-case fit
No primary decision-task fit is mapped for this model yet.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Replicate API | $0.100 | $0.500 | Serverless |
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Cheapest of 1 route · Replicate API