LLM Reference

Persimmon Models by Adept AI

Adept AIApache 2.0Open source
1 model2023Up to 16k ctx

Last refreshed 2026-05-19. Next refresh: weekly.

Details

ResearcherAdept AI
LicenseApache 2.0OSI-approved
Commercial useCommercial use: permitted
Models1
Released2023
Max context16k

About

The Persimmon family of large language models (LLMs) by Adept AI features decoder-only transformer models that excel despite their relatively small parameter counts. Persimmon-8B, the most renowned model in this family, offers an impressive context size of 16K tokens, enabling it to manage longer inputs and preserve more context during text generation 12. Adept AI emphasizes practical evaluation of the models, focusing on direct text generation over implicit probabilities 1. They are distributed under an Apache license to encourage community contributions and further development 12. Although the base model's performance is similar to Llama 2 with less training data, its instruction-tuned variant, Persimmon-8B-FT, outperforms in various benchmarks 1. The architecture incorporates enhancements like squared ReLU activation and query/key layernorm for increased efficiency, and Adept provides fast inference code, blending C++ speed with Python's flexibility 1.

Decision facts

Best fit
coding
Capability starting point
Persimmon 8B with 16k context
Lowest tracked input
Not tracked
Closest related family
Fuyu

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

1 in view

Use when the workload needs 16k context and 8B parameters.

2023-0916k context8B parameters

Release Timeline

1 release group
2023-09
1 current
Persimmon 8B
16k context8B parameters
Current

Specifications(1 models)

Persimmon model specifications comparison
ModelReleasedContextParameters
Persimmon 8B2023-0916k8B