NVIDIA-NeMo / NVIDIA-NeMo/RL

Async DAPO support

Open
#1,365 0 comments 0 reactions 1 assignee Claimed by @ashors1 View on GitHub
enhancement r0.6.0
Dominant language
Python
Stars
2k
Forks
561
Avg merge
4d 5h
Merged PRs (30d)
145

Description

With #602, DAPO reward shaping and dynamic sampling support is only added to the sync path. We should add async support as well.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.