Fuyu Models by Adept AI

Adept AIApache 2.0Open source
2 models2024

Last refreshed 2026-05-19. Next refresh: weekly.

Details

ResearcherAdept AI
LicenseApache 2.0OSI-approved
Commercial useCommercial use: permitted
Models2
Released2024

About

Developed by Adept AI, the Fuyu large language model (LLM) family is distinct for its innovative approach in handling multimodal tasks, specifically catering to digital agents. It boasts a simplified architecture, employing a decoder-only transformer without needing a specific image encoder. Instead, image patches are directly fed as linear projections into the transformer's initial layer, which allows for flexible processing of various image resolutions. This configuration not only streamlines training but also accelerates inference, providing rapid response times under 100 milliseconds for handling large images. Although the Fuyu models are primarily optimized for digital agent applications, they demonstrate robust performance on standard image understanding benchmarks. One of the models in this family, the Fuyu-8B, is publicly available under a CC-BY-NC license, offering opportunities for research and development with the caveat that it may require fine-tuning to meet specific application needs.

Decision facts

Best fit
codingagent workflows
Capability starting point
Fuyu-Heavy
Lowest tracked input
Not tracked
Closest related family
Persimmon

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

1 in view1 retired
Fuyu-HeavyCurrent

Use when provider availability and model metadata match the workload.

2024-01

Release Timeline

1 release group
2024-01
1 current · 1 retired
Fuyu-8B
8B parameters
Archived
Current

Specifications(2 models)

Fuyu model specifications comparison
ModelReleasedParameters
Fuyu-Heavy2024-01—

Available From(1 provider)

Models(2)