how to generate sentence pair similarity data from the raw data?
Open
- Dominant language
- Python
- Stars
- 1.2k
- Forks
- 306
- PR merge metrics
- No merged PRs in 30d
Description
Hi, how can I use these data to train my sentence semantic similarity model?
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue names no files, tests, or entry points for generating sentence-pair similarity data. First clarify the raw-data format and the expected training examples, then document the supported preparation workflow and how to verify that the resulting data is suitable for training.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- data, machine-learning
- Issue type
- Documentation
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100