Persimmon 8B
Persimmon 8B is released 2023-09-16 in the Persimmon family with open-source; evaluate it while provider pricing coverage matures.
Use it for
- Teams evaluating general LLM work
- Workloads that can use a 16k context window
Do not use it for
- Cost-sensitive launches that need sourced token pricing
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
- Family
- Persimmon
- Released
- 2023-09-16
- Context
- 16k
- Parameters
- 8B
- Architecture
- Decoder Only
- Specialization
- general
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Unknown
- Code
- Unknown
- Training
- Fine-tuned
No tracked provider token pricing is available yet.
About
Persimmon-8B is a sophisticated open-source large language model developed by Adept AI, featuring approximately 8 billion parameters. It is a decoder-only transformer enhanced with squared ReLU activation functions and rotary positional encodings, offering a substantial context window of 16,000 tokens, more than quadrupling the capacity of models like LLaMA 2 and GPT-3. Trained on a dataset consisting of 737 billion tokens blended with text and code, it employs an advanced version of FlashAttention for efficient handling of long sequences. Despite utilizing less data than LLaMA 2, it achieves comparable performance on various benchmarks.
Top use-case fit
No primary decision-task fit is mapped for this model yet.
Provider price ladder
No tracked provider token pricing is available for this model yet.
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.