NVIDIA-NeMo / NVIDIA-NeMo/RL

On-policy distillation tutorial with Latest models

Open
#1,679 2 comments 1 reaction 1 assignee Claimed by @zpqiu View on GitHub
enhancement t-onpolicydistillation x-sesai
Dominant language
Python
Stars
2k
Forks
561
Avg merge
4d 5h
Merged PRs (30d)
145

Description

Maybe let's do gpt-oss? It is just supported in the mcore path

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.