AlibabaResearch / AlibabaResearch/RoTS

Some questions regarding your research

Open
#3 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
No language data
Stars
6
Forks
0
PR merge metrics
No merged PRs in 30d

Description

Although you control the final fine-tuning set size and use a comparable rollout budget, the current ablations only demonstrate the effectiveness of the overall FDE/EIR data-synthesis pipeline over root-level parallel sampling. You seem to not isolate the contribution of the proposed heuristic node-selection mechanisms. In particular, comparisons against random or uniform intermediate-node expansion, random failed-prefix recovery, and greedy priority-based selection without the UCB term are missing. Therefore, the observed gains may stem from prefix reuse, increased state/error diversity, or a higher-quality candidate data pool, rather than specifically from the proposed fragility and recovery-node selection criteria.

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.