automl / automl/LCBench

clarification required for the validation split

Open
#2 0 comments 2 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
36
Forks
10
PR merge metrics
No merged PRs in 30d

Description

it is unfortunately not clear how the validation split is defined for each dataset. As the list of fields only indicate 'test_split'. Are we to find this information on OpenML website? Also, there appears no information on how to recreate the train/val/test splits.

Contributor guide

No contributing guide indexed for this repository

Research direction

Review how the benchmark currently represents each dataset's fields, especially the test_split value, and check the corresponding OpenML information. Document where the validation split is defined and how the train, validation, and test splits can be recreated, then verify the explanation against the dataset records.

Written by the indexing model from the issue text.

Assessment

Tech stack
jupyter-notebook
Domain
documentation
Issue type
Documentation
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
30/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.