Doubao Embedding Vision Models by ByteDance
ByteDanceProprietary
ByteDance releases · 19 in the last 12 months · this family litChangelog →
1 model2025
Last refreshed 2026-07-10. Next refresh: weekly.
Details
ResearcherByteDance
LicenseProprietary
Commercial useCommercial use: conditional
Models1
Released2025
Capabilities
VisionAll models
MultimodalAll models
Links
WebsiteAbout
ByteDance's multimodal embedding family for representing text and images in a shared vector space for retrieval, semantic search and RAG applications.
Decision facts
- Best fit
- embeddingvision and multimodal work
- Capability starting point
- Doubao Embedding Vision with multimodal inputs
- Lowest tracked input
- Not tracked
- Closest related family
- BAGEL
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
1 in view
Doubao Embedding VisionCurrent
Use when the workload needs embedding and multimodal inputs.
2025-12embeddingmultimodal inputs
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| Doubao Embedding Vision | Use when the workload needs embedding and multimodal inputs. | 2025-12 | embeddingmultimodal inputs | Current |
Release Timeline
1 release group2025-12
1 current
Doubao Embedding Vision
Currentembeddingmultimodal inputs
Specifications(1 models)
| Model | Released | Vision | Multimodal |
|---|---|---|---|
| Doubao Embedding Vision | 2025-12 | Yes | Yes |


