kohya-ss / kohya-ss/sd-scripts

rl-stablediffusion training

Open
#575 2 comments 1 reaction 0 assignees View on GitHub
Dominant language
Python
Stars
7.2k
Forks
1.2k
Avg merge
11m
Merged PRs (30d)
2

Description

any chance you could implement this?
https://github.com/vinhkhuc/ddpo/tree/support_gpu
it's for RLHF type of stuff, [check the paper](https://rl-diffusion.github.io/)
could be really interesting for lora and finetuning

Contributor guide

No contributing guide indexed for this repository

Research direction

No repository files, tests, or entry points are named. Start by reviewing the linked ddpo support_gpu branch and the linked paper, then determine the scope for RLHF-style training support. Done would require an agreed implementation covering the proposed LoRA and fine-tuning use cases.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.