LLM Reference

Starling Models by Nexusflow

NexusflowLlama 2 CommunityOpen weights
1 model2024Up to 8k ctx

Last refreshed 2026-05-19. Next refresh: weekly.

Details

ResearcherNexusflow
Commercial useCommercial use: conditional
Models1
Released2024
Max context8k

About

The Starling family of Large Language Models (LLMs) stems from the innovative Berkeley-Nest AI research group. Among its models, Starling-LM-7B-alpha stands out as a 7-billion parameter language model specifically fine-tuned using Reinforcement Learning from AI Feedback (RLAIF). This fine-tuning process harnessed the extensive Nectar dataset, comprising GPT-4-ranked chat prompts and responses. Starling-LM-7B-alpha focuses on enhancing its helpfulness and maintaining a non-harmful approach, evolving from the Openchat 3.5 model. The project also introduced the Starling-RM-7B-alpha reward model, pivotal for RLAIF processes. To foster advancements in RLHF mechanisms and AI safety, the dataset, reward model, and language model are openly accessible. Additionally, a more refined iteration, Starling-LM-7B-beta, has been made available for further research and development.

Decision facts

Best fit
coding
Capability starting point
Starling LM 7B Beta with 8k context
Lowest tracked input
Not tracked
Closest related family
NexusRaven

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

1 in view

Use when the workload needs 8k context and 7B parameters.

2024-028k context7B parameters

Release Timeline

1 release group
2024-02
1 current
Starling LM 7B Beta
8k context7B parameters
Current

Specifications(1 models)

Starling model specifications comparison
ModelReleasedContextParameters
Starling LM 7B Beta2024-028k7B