huggingface / huggingface/datablations
Figure issue about your paper (Figure 4 and Figure 15)
- Dominant language
- Jupyter Notebook
- Stars
- 345
- Forks
- 19
- PR merge metrics
- No merged PRs in 30d
Description
Hi,
I am reading your paper and I have noticed that figure 4 and figure 15 are exactly the same. Are they meant to be the same? I believe that figure 4 shows the result of models trained on a different training set rather than OSCAR corpus (because you mentioned in Appendix I that 'To ensure our findings are not dataset-dependent, we train models with the same configurations from Figure 4 on the OSCAR corpus'). I wonder if you placed a wrong figure here.
Appreciate your work! I believe this work will definitely give researchers and engineers more insights when training and developing LLMs.
Matthew
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.