Palmyra Vision Models by Writer
Last refreshed 2026-07-10. Next refresh: weekly.
Details
Capabilities
Links
WebsiteAbout
Palmyra-Vision is Writer's sophisticated multimodal large language model (LLM) that specializes in interpreting and generating text from images. Equipped to handle a variety of tasks—such as extracting handwritten text, classifying objects and colors, and describing visual data like charts and infographics—it performs exceptionally in real-world applications. Notably, it achieved an 84.4% accuracy score on the VQAv2 benchmark, outperforming other leading multimodal models like GPT-4V. This makes it ideal for enterprise tasks including compliance checks, generating product descriptions, and creating accessible ALT text. Accessible via Writer's image analyzer app, Palmyra-Vision can also be integrated into custom AI solutions through Writer's AI Studio, offering flexibility for tailored business needs 13.
Decision facts
- Best fit
- visionvision and multimodal work
- Capability starting point
- Palmyra Vision with 8k context and multimodal inputs
- Lowest tracked input
- Palmyra Vision · $7.5/1M · Writer AI Studio
- Closest related family
- Camel
Archived Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
Keep for legacy integrations; evaluate Palmyra X5 before new work.
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| Palmyra Vision | Keep for legacy integrations; evaluate Palmyra X5 before new work. | 2024-02 | vision8k contextmultimodal inputs | Replaced |
Release Timeline
1 release groupReplaced By
Keep for legacy integrations; evaluate Palmyra X5 before new work.
Specifications(1 models)
| Model | Released | Context | Vision | Multimodal |
|---|---|---|---|---|
| Palmyra Vision | 2024-02 | 8k | Yes | Yes |
Available From(1 provider)
Pricing
| Model | Provider | Input / 1M | Output / 1M | Type |
|---|---|---|---|---|
| Palmyra Vision | Writer AI Studio | $7.5 | $22.5 | Serverless |





