allenai / allenai/dont-stop-pretraining

TypeError: stat: path should be string, bytes, os.PathLike or integer, not NoneType

オープン
#37 コメント 2 件 リアクション 0 件 担当者 0 名 GitHub で見る
主要言語
Python
スター
544
フォーク
72
PR マージ指標
30日以内にマージされた PR はありません

説明

Hi,

I am trying to train the biomed_roberta_base model on the chemprot dataset using the provided scripts.train python command and encounter the below issues.

![Screenshot from 2022-04-20 00-16-11](https://user-images.githubusercontent.com/17106288/164149421-62dc8af6-d4a3-46f6-bde1-74d3576d1e90.png)

The above dataset and models have been downloaded as stated in the README of the master branch. Also since the mentioned environment wasn't working for me I am using the below conda environment

![Screenshot from 2022-04-20 00-20-29](https://user-images.githubusercontent.com/17106288/164149793-2ef78e6e-0db3-4521-a752-afc15145203d.png)

Please let me know how to solve the above issue. It seems like the tokenizer asks for a vocab file but I am not sure how to provide one.

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

評価

この issue はまだ評価されていません。

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。