AI-Hypercomputer / AI-Hypercomputer/maxtext
Support LoRA training
未關閉
feature request
- 主要語言
- Python
- 星號
- 2.4k
- 分支
- 607
- 平均合併
- 2 天 19 小時
- 30 天內合併 PR
- 158
描述
Is there a plan to support PEFT methods like LoRA training in maxtext to support larger model fine-tuning / continue pretraining so that bigger models like LLaMA-3-70B can be trainined even with small amount of TPU/GPUs?
貢獻指南
評估
這個 Issue 還沒有評估資料。