AlibabaResearch / AlibabaResearch/AdvancedLiterateMachinery

VGT evaluation result not matching for DoclayNet.

Aperta
#188 0 commenti 0 reazioni 0 assegnatari Vedi su GitHub
Lingua principale
C++
Stelle
1.8k
Fork
195
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

Hi,

Thank you for creating and sharing the Vision Grid Transformer repository.

I am currently trying to evaluate the model on the DocLayNet test dataset in order to replicate the published results (mAP of 83.7). I am using the weights available here: [doclaynet_VGT_model.pth](https://github.com/AlibabaResearch/AdvancedLiterateMachinery/releases/download/v1.3.0-VGT-release/doclaynet_VGT_model.pth).

I executed the evaluation using the following command:

bash
Copy code
python path/to/train_VGT.py --config-file VGT/object_detection/Configs/cascade/doclaynet_VGT_cascade_PTM.yaml --eval-only --num-gpus 1 MODEL.WEIGHTS VGT/downloads/weights/doclaynet_VGT_model.pth OUTPUT_DIR VGT/AdvancedLiterateMachinery/DocumentUnderstanding/VGT/downloads
However, the results I obtained differ from the published ones. Please see the attached matrix for reference:

![image](https://github.com/user-attachments/assets/9b6a9443-62b2-441f-83f8-d3fb6be02a51)

Could you please advise if I might be overlooking something in the evaluation process?

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.