allenai / allenai/fluid-benchmarking

Issue with reproducing the paper results

Ouverte
#2 0 commentaires 0 réactions 0 personnes assignées Voir sur GitHub
Langage dominant
Python
Étoiles
29
Forks
4
Métriques de merge des PR
Aucune PR mergée en 30 j

Description

Hello,

Could you please provide the evaluation metrics codes to produce Tables 1 and 2 ? After all my attempts I could not reproduce those numbers (I did get the main trend but the numbers are far off good as reported in the paper). Even when using the experiments.jsonl provided in the repository (and the one obtained from running scripts/run_experiments.py) I still can not get the numbers stated in the paper. Also, the both files contain 2,712 ckpts, not 2802 as mentioned on the paper.

Any help would be greatly appreciated. Thanks in advance

CC @valentinhofmann @davidheineman

Guide de contribution

Aucun guide de contribution indexé pour ce dépôt

Évaluation

Cette issue n'a pas encore été évaluée.

Recevez les nouvelles issues par e-mail

Un résumé court des issues GitHub adaptées aux débutants.