AI-Hypercomputer / AI-Hypercomputer/maxtext
Support LoRA training
未关闭
feature request
- 主要语言
- Python
- 星标
- 2.4k
- 派生
- 607
- 平均合并
- 2 天 19 小时
- 30 天内合并 PR
- 158
描述
Is there a plan to support PEFT methods like LoRA training in maxtext to support larger model fine-tuning / continue pretraining so that bigger models like LLaMA-3-70B can be trainined even with small amount of TPU/GPUs?
贡献指南
评估
这个 Issue 还没有评估数据。