huggingface / huggingface/alignment-handbook

Reward Modeling Support

Open
#109 0 comments 3 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
5.7k
Forks
490
Avg merge
2m
Merged PRs (30d)
1

Description

Hi the team, great work! I wonder whether there will be demo / example about training reward models in multi-GPU env or deepspeed seetings?

Thanks!

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.