huggingface / huggingface/cookbook

Idea for contribution : Recipe on Training FineTuning Embeddings / Re-Rankers on Custom Domain Specific Dataset

Open
#25 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
2.7k
Forks
417
Avg merge
17h
Merged PRs (30d)
3

Description

Recipe to guide how to prepare data , benchmark to follow , methods to fine tune the bi encoders , cross encoders to adapt the domain specific dataset

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reviewing the existing cookbook recipes and the requested topics in the issue: data preparation, benchmarking, and fine-tuning bi-encoders and cross-encoders. Done means adding a recipe that covers those steps for a custom domain-specific dataset.

Written by the indexing model from the issue text.

Assessment

Domain
documentation, machine-learning
Issue type
Documentation
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Mostly clear
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.