EpistasisLab / EpistasisLab/tpot

Upper Limit of Data that TPOT Can Handle

Open
#962 1 comment 0 reactions 0 assignees View on GitHub
enhancement question
Dominant language
Jupyter Notebook
Stars
10.1k
Forks
1.6k
PR merge metrics
No merged PRs in 30d

Description

This is less so an issue than a question. I haven't been able to really find populated discussion forums for this package, so I thought I'd ask here.

I'm currently working on datasets for practice that are around 500-1000 rows and 13 features. It takes a really long time to do work on this, which makes sense.

This got me thinking then for larger datasets. I came across another post here that mentioned it being very time consuming and improbably to train a set on 1m records or more.

What do you think is the acceptable dataset size TPOT should be used on typically?

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.