AI-Hypercomputer / AI-Hypercomputer/maxtext

Support LoRA training

未關閉
#609 5 則留言 3 個 reaction 已指派 1 人 已指派給 @xibinliu 在 GitHub 檢視
feature request
主要語言
Python
星號
2.4k
分支
607
平均合併
2 天 19 小時
30 天內合併 PR
158

描述

Is there a plan to support PEFT methods like LoRA training in maxtext to support larger model fine-tuning / continue pretraining so that bigger models like LLaMA-3-70B can be trainined even with small amount of TPU/GPUs?

貢獻指南

開啟貢獻指南

評估

這個 Issue 還沒有評估資料。

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。