Mismatch between pre-training and fine-tuning phase
未關閉
- 主要語言
- Python
- 星號
- 157
- 分支
- 39
- PR 合併指標
- 30 天內沒有已合併 PR
描述
As far as I'm aware, PRIMERA replaces sentences with \ tokens during pre-training. However, these \ tokens do not appear on fine-tuning phase, or when doing inference. Still, it was shown to achieve impressive results on Zero-Shot and Few-Shot Evaluation. I was wondering if PRIMERA had any strategies to reduce the mismatch between fine-tuning and pre-training phase ? (I did not see related information mentioned in the paper)
貢獻指南
這個儲存庫沒有索引到貢獻指南
評估
這個 Issue 還沒有評估資料。