Preprint2023
QLoRA
Unknown
QLoRA enables fine-tuning of large language models on a single GPU by combining 4-bit NormalFloat quantization, double quantization, and paged optimizers, matching 16-bit performance.
0May 1, 2023Quantization
arXiv
A comprehensive index of artificial intelligence and machine-learning research with AI-generated summaries, citation metrics, and direct links to papers and code.