alteryx / alteryx/evalml

Explore Target Encoder vs OHE usage scenarios

Aperta
#1,440 0 commenti 0 reazioni 1 assegnatario Rivendicata da @asniyaz Vedi su GitHub
enhancement needs design spike
Lingua principale
Python
Stelle
850
Fork
96
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

As raised in discussion with @dsherry @freddyaboulton and @rpeck, we want to explore the possibility of creating a `top_n` parameter for Target Encoder and grouping objects outside of this number together into one group, [here](https://github.com/alteryx/evalml/pull/1401#discussion_r523236206). Additionally, we wanted to explore if there were a number of categoricals or a distribution of categoricals where it would be more beneficial to use Target Encoder versus One Hot Encoder. This would be useful in order to add Target Encoder to `AutoMLSearch`.

Guida per i contributori

Apri la guida per i contributori

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.