A technical note comparing parameter-efficient fine-tuning techniques for large language models, including Low-Rank Adaptation (LoRA), its quantized variant QLoRA, and Weight-Decomposed Low-Rank Adaptation (DoRA).
Paper
The full text of this publication is not hosted on 44B due to licensing.
Read it at OpenAlex