NVIDIA-NeMo / NVIDIA-NeMo/RL

[Feature Request] Qwen3VL GRPO, SFT training

Open
#1,256 5 comments 1 reaction 1 assignee Claimed by @eagle705 View on GitHub
enhancement external multimodal new model r0.6.0 x-ss
Dominant language
Python
Stars
2k
Forks
561
Avg merge
4d 5h
Merged PRs (30d)
145

Description

**Additional context**

Our customer would like to apply RL methods (GRPO, GSPO, and SPO) to VLM with MoE (such as Qwen3-VL).

Would it be possible to extend the current VLM support to Qwen3-VL?

(cc. @terrykong, @snowmanwwg )

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.