apache / apache/datafusion-ballista

Can Ballista do S3 prefetching like Spark?

Open
#2,162 1 comment 0 reactions 0 assignees View on GitHub
enhancement
Dominant language
Rust
Stars
2.1k
Forks
320
Avg merge
1d 22h
Merged PRs (30d)
66

Description

**Is your feature request related to a problem or challenge? Please describe what you are trying to do.**
This is to track a comment https://github.com/apache/datafusion-ballista/pull/2161#discussion_r3640798050

**Describe the solution you'd like**

**Describe alternatives you've considered**

**Additional context**

Contributor guide

Open the contributing guide

Research direction

Start by reading the discussion on pull request 2161, especially the linked comment, to understand the S3 prefetching request and how it compares with Spark. The issue does not name files, tests, entry points, or acceptance criteria, so the desired behavior and definition of done need clarification before implementation.

Written by the indexing model from the issue text.

Assessment

Tech stack
aws, rust, spark
Domain
cloud, distributed-systems
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Quiet
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.