facebookresearch / facebookresearch/MLGym

Question: RL Algorithms to train LLMs

Open
#13 8 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
622
Forks
59
PR merge metrics
No merged PRs in 30d

Description

Quick question: The repo README mentions that the goal is to implement RL algorithms to train LLMs in a research environment. Upon going through the repository I couldn't find any specific code for RL training LLMs. Requesting the organizers to help with the following questions:
1. Could you confirm that at this moment, no specific code for RL training is released
2. Could you let us know if this is in the future roadmap and if yes, when can we expect it to be released?

Thank you very much for the excellent work. Loved it!

Contributor guide

Open the contributing guide

Research direction

Start with the repository README, which describes the goal of implementing RL algorithms for training LLMs, and compare that statement with the current repository contents. Done means clearly answering whether RL-for-LLM training code is currently released and, if applicable, documenting the future roadmap or expected release timing.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
documentation, machine-learning
Issue type
Documentation
Difficulty
1/5
Estimated time
Under an hour
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.