Starling LM 7B Beta
Starling LM 7B Beta is a released coding and classification model with open-weight; evaluate it while provider pricing coverage matures.
Use it for
- Teams evaluating coding and classification
- Workloads that can use a 8k context window
Do not use it for
- Cost-sensitive launches that need sourced token pricing
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
- Family
- Starling
- Released
- 2024-02-05
- Context
- 8k
- Parameters
- 7B
- Architecture
- Decoder Only
- Specialization
- general
- Openness
- Open weights
- License
- Llama 2 CommunityCommercial use: conditional
- Weights
- Unknown
- Code
- Unknown
- Training
- Fine-tuned
No tracked provider token pricing is available yet.
About
Starling LM 7B Beta is an open-source large language model crafted by Nexusflow, leveraging a 7-billion parameter transformer architecture tailored for conversational AI. Fine-tuned with Reinforcement Learning from AI Feedback (RLAIF), it aims to enhance helpfulness and minimize harm. Built on the foundation of the Openchat-3.5-0106 and Mistral-7B-v0.1 models, it utilizes the berkeley-nest/Nectar ranking dataset, Nexusflow/Starling-RM-34B reward model, and Proximal Policy Optimization (PPO) strategy. Achieving an improved MT Bench score of 8.12, its capabilities span engaging conversations, informative responses, and tasks like content and code generation.
Top use-case fit: coding, agents, and build tasks
Coding
1 relevant benchmark in the decision map.
Classification
2 relevant benchmarks in the decision map.
Provider price ladder
No tracked provider token pricing is available for this model yet.
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
Benchmark scores(4)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Google-Proof Q&A | 49.7 | diamondObserved 2026-03-06 | — | Source |
| HellaSwag | 89.2 | 10-shotObserved 2026-03-06 | — | Source |
| HumanEval | 74.2 | pass@1Observed 2026-03-06 | — | Source |
| Massive Multitask Language Understanding | 77.8 | 5-shotObserved 2026-03-06 | — | Source |
Migration checks
No linked migration route is available for this model yet.