facebookresearch / facebookresearch/MLGym
Question: RL Algorithms to train LLMs
- Dominant language
- Python
- Stars
- 622
- Forks
- 59
- PR merge metrics
- No merged PRs in 30d
Description
Quick question: The repo README mentions that the goal is to implement RL algorithms to train LLMs in a research environment. Upon going through the repository I couldn't find any specific code for RL training LLMs. Requesting the organizers to help with the following questions:
1. Could you confirm that at this moment, no specific code for RL training is released
2. Could you let us know if this is in the future roadmap and if yes, when can we expect it to be released?
Thank you very much for the excellent work. Loved it!
Contributor guide
Research direction
Start with the repository README, which describes the goal of implementing RL algorithms for training LLMs, and compare that statement with the current repository contents. Done means clearly answering whether RL-for-LLM training code is currently released and, if applicable, documenting the future roadmap or expected release timing.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- documentation, machine-learning
- Issue type
- Documentation
- Difficulty
- 1/5
- Estimated time
- Under an hour
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100