LLM Reference
Concepts & capability filters

alignment

Alignment is the tuning that turns a raw predictive model into a deployable one: how it follows instructions, when and how it refuses, its tone, and how reliably it stays on task.

Category
Not classified
Difficulty
Not classified
Aliases
None tracked
Last reviewed
2026-07-02

Key facts

  • Techniques include RLHF, DPO, and constitutional methods.
  • For a build decision, alignment shows up as refusal behavior and instruction adherence you can test on your own prompts.

Models Mentioning alignment(12)