Consider supporting dasks
Open
enhancement
- Dominant language
- Python
- Stars
- 1k
- Forks
- 131
- PR merge metrics
- No merged PRs in 30d
Description
[pasted from an offline conversation with mrocklin]
I would like to utilize dask to parallelize the computation of some of the ‘into’ ETL arcs/production rules by partitioning (given that there is so much high level type information and metadata accessible)
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue names no files, tests, or entry points. Start by locating the “into” ETL arcs or production rules and determine how Dask partitioning would fit; the issue does not define acceptance criteria for what done looks like.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- data-engineering
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100