Platypus2 7B
Platypus2 7B is released 2023-12-15 in the Platypus2 family with open-weight; evaluate it while provider pricing coverage matures.
Use it for
- Teams evaluating general LLM work
Do not use it for
- Cost-sensitive launches that need sourced token pricing
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
- Family
- Platypus2
- Released
- 2023-12-15
- Parameters
- 7B
- Architecture
- Decoder Only
- Specialization
- general
- Openness
- Open weights
- License
- Llama 2 CommunityCommercial use: conditional
- Weights
- Unknown
- Code
- Unknown
- Training
- Fine-tuned
No tracked provider token pricing is available yet.
About
The Platypus2-7B is a large language model based on the Llama 2-7B transformer architecture, specifically instruction-fine-tuned to excel in English language tasks. Developed by Cole Hunter and Ariel Lee, it was trained using a STEM and logic-focused dataset, allowing it to adeptly handle tasks that demand logical reasoning and problem-solving. Training was conducted on a single A100 80GB GPU utilizing LoRA for fine-tuning efficiency. Notable considerations include using fp16=False and bf16=True during fine-tuning for optimal performance.
Top use-case fit
No primary decision-task fit is mapped for this model yet.
Provider price ladder
No tracked provider token pricing is available for this model yet.
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
No tracked provider token pricing is available yet.