← Flash PapersQLoRA: Efficient Finetuning of Quantized LLMs

Share this Flash Paper

QLoRA: Efficient Finetuning of Quantized LLMs

https://flashpapers.ai/p/ntq7uuag9g
Make your own
Tim Dettmers, Artidoro Pagnoni +2 more
Details ▸
VRAM requirement reduction:from >780GB to <48GBGuanaco 65B relative performance:99.3%· Vicuna BenchmarkDouble Quantization savings:0.37 bits/parameter
Concepts:LoRA