Sample selection
- Lingua principale
- Jupyter Notebook
- Stelle
- 551
- Fork
- 143
- Metriche di merge delle PR
- Nessuna PR unita negli ultimi 30g
Descrizione
I would like to ask if AutoX has any plans for sample selection?
Now many data sets are so large that the computing power of individuals and small companies cannot afford.
Can a part of the data be selected for training to approximate the effect of full data training?
Guida per i contributori
Nessuna guida per i contributori indicizzata per questo repository
Direzione di ricerca
Look at the AutoX codebase for data loading and preprocessing modules to understand the current pipeline. Investigate existing sampling techniques or if any are implemented. Determine how to integrate a sample selection feature that works with large datasets and evaluate its impact on training performance.
Scritto dal modello di indicizzazione a partire dal testo della issue.
Valutazione
- Stack tecnologico
- jupyter-notebook, machine-learning, python
- Ambito
- data-engineering, machine-learning
- Tipo di issue
- Funzionalità
- Difficoltà
- 4/5
- Tempo stimato
- 3-5 giorni
- Stato di attività
- Ferma
- Chiarezza
- Abbastanza chiara
- Idoneità per principianti
- 35/100