huggingface / huggingface/alignment-handbook
Is QLoRA better than finetuning?
Open
- Dominant language
- Python
- Stars
- 5.7k
- Forks
- 490
- Avg merge
- 2m
- Merged PRs (30d)
- 1
Description
The results reported in https://github.com/huggingface/alignment-handbook/pull/88 suggest that QLoRA is better for both SFT and DPO. Is this accurate, and have people seen this happen in any other settings?
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.