AnswerDotAI / AnswerDotAI/ModernBERT

Vocab.txt for ONNX?

Open
#160 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
1.7k
Forks
145
PR merge metrics
No merged PRs in 30d

Description

I want to try this out with ONNX and I'm not using Python, using .Net 9. How can I get the vocab.txt file because it is larger than the current one I'm using for BERT tokenization?

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by checking the repository's tokenizer and model asset documentation for how vocab.txt is produced or distributed, then compare the ONNX and .NET 9 usage described in the issue. Done means documenting or providing the larger vocabulary file for the reported BERT tokenization workflow.

Written by the indexing model from the issue text.

Assessment

Domain
machine-learning
Issue type
Documentation
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.