kohya-ss / kohya-ss/sd-scripts
If GGPO is NOT used for flux lora training, all "avg_key_norm" are NaN starting from step0 in Tensorboard
Open
- Dominant language
- Python
- Stars
- 7.2k
- Forks
- 1.2k
- Avg merge
- 11m
- Merged PRs (30d)
- 2
Description
Tried different optimizers, lower/higher lr, rank/alpha combinations, datasets, base models. If i do not enable GGPO for Flux lora training the average key norms are always NaN
Contributor guide
No contributing guide indexed for this repository
Research direction
Start by reproducing Flux LoRA training with GGPO disabled and inspect the TensorBoard avg_key_norm values from step 0. Compare the training behavior with GGPO enabled while narrowing down where the NaN values first appear; done means avg_key_norm remains finite without GGPO.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- machine-learning
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100