QLORA: Efficient Finetuning of Quantized LLMs

4-bit quantized LLM fine-tuning with LoRA adapters: near full-precision performance, drastic memory reduction, single-GPU training.
Efficient Adaptation
Author

Imad Dabbura

Published

July 15, 2023

QLORA: Efficient Finetuning of Quantized LLMs

Back to top