AI-Hypercomputer / AI-Hypercomputer/maxtext

Support LoRA training

未关闭
#609 5 条评论 3 个 reaction 已指派 1 人 已指派给 @xibinliu 在 GitHub 查看
feature request
主要语言
Python
星标
2.4k
派生
607
平均合并
2 天 19 小时
30 天内合并 PR
158

描述

Is there a plan to support PEFT methods like LoRA training in maxtext to support larger model fine-tuning / continue pretraining so that bigger models like LLaMA-3-70B can be trainined even with small amount of TPU/GPUs?

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。