LoRA Learns Less and Forgets Less research paper by Databricks, 2024
Databricks · May 15, 2024 · Training and scaling · 405 citations · 91 upvotes
What it shows
LoRA learns less than full fine-tuning on code and math, but forgets less of what the model already knew.
Summarised by hand from the abstract.
Other training and scaling papers
TopicAbout this paper
- Authors
- Dan Biderman, Jose Gonzalez Ortiz, Jacob Portes and 9 more
- arXiv
- 2405.09673 · PDF
- Venue
- Trans. Mach. Learn. Res.
- Citations
- 405, 38 influential · Semantic Scholar
- Upvotes
- 91 · Hugging Face
- Code
- github.com/danbider/lora-tradeoffs
- Lab
- Databricks · on Companies · on Acquisitions · on Rounds · on Paydays · on Releases · on TechConf
Changes
| What changed | |
|---|---|
| Sep 24, 2026 | New paperAdded to the listSep 24, 2026 |
Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.