Concepts & capability filters
Quantized Low-Rank Adaptation
QLoRA extends LoRA by combining 4-bit quantization (via NF4 and double quantization) with paged optimizers to fine-tune billion-parameter LLMs on consumer GPUs.
- Category
- Not classified
- Difficulty
- Not classified
- Aliases
- None tracked
- Last reviewed
- 2026-07-02
Key facts
- It maintains performance parity with 16-bit full tuning while dramatically reducing memory requirements.