huggingface / huggingface/cookbook
Idea for contribution : Recipe on Training FineTuning Embeddings / Re-Rankers on Custom Domain Specific Dataset
Open
- Dominant language
- Jupyter Notebook
- Stars
- 2.7k
- Forks
- 417
- Avg merge
- 17h
- Merged PRs (30d)
- 3
Description
Recipe to guide how to prepare data , benchmark to follow , methods to fine tune the bi encoders , cross encoders to adapt the domain specific dataset
Contributor guide
No contributing guide indexed for this repository
Research direction
Start by reviewing the existing cookbook recipes and the requested topics in the issue: data preparation, benchmarking, and fine-tuning bi-encoders and cross-encoders. Done means adding a recipe that covers those steps for a custom domain-specific dataset.
Written by the indexing model from the issue text.
Assessment
- Domain
- documentation, machine-learning
- Issue type
- Documentation
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Mostly clear
- Newbie friendliness
- 35/100