AnswerDotAI / AnswerDotAI/ModernBERT
Vocab.txt for ONNX?
Open
- Dominant language
- Python
- Stars
- 1.7k
- Forks
- 145
- PR merge metrics
- No merged PRs in 30d
Description
I want to try this out with ONNX and I'm not using Python, using .Net 9. How can I get the vocab.txt file because it is larger than the current one I'm using for BERT tokenization?
Contributor guide
No contributing guide indexed for this repository
Research direction
Start by checking the repository's tokenizer and model asset documentation for how vocab.txt is produced or distributed, then compare the ONNX and .NET 9 usage described in the issue. Done means documenting or providing the larger vocabulary file for the reported BERT tokenization workflow.
Written by the indexing model from the issue text.
Assessment
- Domain
- machine-learning
- Issue type
- Documentation
- Difficulty
- 3/5
- Estimated time
- 1-2 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100