2305.14314
arxiv.org · 7,464 words · saved by 2 readers
N/A
QL O RA: Efficient Finetuning of Quantized LLMs Tim Dettmers∗ Artidoro Pagnoni∗ Ari Holtzman Luke Zettlemoyer arXiv:2305.14314v1 [cs.LG] 23 May 2023 University of Washington {dettmers,artidoro,ahai,lsz}@cs.washington.edu…
saved by
related reading
- [2510.11696] QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMsarxiv.org
- 2005.14165arxiv.org
- LoRA Without Regret - Thinking Machines Labthinkingmachines.ai
- In-depth guide to fine-tuning LLMs with LoRA and QLoRA | Mercity Researchmercity.ai
- Finetune LLMs on your own consumer hardware using tools from PyTorch and Hugging Face ecosystem – PyTorchpytorch.org
- A Guide to Quantization in LLMs | Symbl.aisymbl.ai
- Quantization from the ground upngrok.com
- Efficient LLM inferencefinbarrtimbers.substack.com
- Efficient LLM Finetuning with Unsloth | Modal Docsmodal.com
- What Makes Low-Bit Quantization-Aware Training Work for Reasoning LLMs? A Systematic Studyarxiv.org
- gpt-4.pdfcdn.openai.com
- Fine-tuning LLMs Guide | Unsloth Documentationdocs.unsloth.ai