kohya-ss / kohya-ss/sd-scripts
rl-stablediffusion training
Open
- Dominant language
- Python
- Stars
- 7.2k
- Forks
- 1.2k
- Avg merge
- 11m
- Merged PRs (30d)
- 2
Description
any chance you could implement this?
https://github.com/vinhkhuc/ddpo/tree/support_gpu
it's for RLHF type of stuff, [check the paper](https://rl-diffusion.github.io/)
could be really interesting for lora and finetuning
Contributor guide
No contributing guide indexed for this repository
Research direction
No repository files, tests, or entry points are named. Start by reviewing the linked ddpo support_gpu branch and the linked paper, then determine the scope for RLHF-style training support. Done would require an agreed implementation covering the proposed LoRA and fine-tuning use cases.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100