LLM Reference
Concepts & capability filters

alignment

See matching models with benchmark scores and pricing.

Definition

Alignment is the tuning that turns a raw predictive model into a deployable one: how it follows instructions, when and how it refuses, its tone, and how reliably it stays on task. Techniques include RLHF, DPO, and constitutional methods. For a build decision, alignment shows up as refusal behavior and instruction adherence you can test on your own prompts.

Models Mentioning alignment(12)