autogluon / autogluon/tabarena

[TechDebt] Improve Getting Datasets Workflow

Open
#179 9 comments 0 reactions 0 assignees View on GitHub
TechDebt
Dominant language
Python
Stars
303
Forks
69
Avg merge
1d 4h
Merged PRs (30d)
49

Description

See https://github.com/TabArena/tabarena_benchmarking_examples/issues/5

I think we could easily add a function or example scripts on how to use the datasets w/o TabArena code to people who want to use them.
How exactly is to be determined.

Contributor guide

Open the contributing guide

Research direction

Start by reading the linked tabarena_benchmarking_examples issue #5 and inspecting how this repository currently exposes dataset usage. The scope and implementation are explicitly undecided; done means providing a clear, usable way to work with the datasets without relying on TabArena code.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.