apache / apache/datafusion-ballista
Can Ballista do S3 prefetching like Spark?
- Dominant language
- Rust
- Stars
- 2.1k
- Forks
- 320
- Avg merge
- 1d 22h
- Merged PRs (30d)
- 66
Description
**Is your feature request related to a problem or challenge? Please describe what you are trying to do.**
This is to track a comment https://github.com/apache/datafusion-ballista/pull/2161#discussion_r3640798050
**Describe the solution you'd like**
**Describe alternatives you've considered**
**Additional context**
Contributor guide
Research direction
Start by reading the discussion on pull request 2161, especially the linked comment, to understand the S3 prefetching request and how it compares with Spark. The issue does not name files, tests, entry points, or acceptance criteria, so the desired behavior and definition of done need clarification before implementation.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- aws, rust, spark
- Domain
- cloud, distributed-systems
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Quiet
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100