apache / apache/hudi

Spark: Support PushDown for DataSourceV2 Read

Open
#15,345 0 comments 0 reactions 0 assignees View on GitHub
from-jira priority:high type:improvement
Dominant language
Java
Stars
6.2k
Forks
2.5k
Avg merge
2d 8h
Merged PRs (30d)
111

Description

## JIRA info

- Link: https://issues.apache.org/jira/browse/HUDI-4640
- Type: Improvement
- Epic: https://issues.apache.org/jira/browse/HUDI-4449

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names Spark DataSourceV2 Read and links JIRA HUDI-4640; start by reading that JIRA ticket and locating Hudi's Spark DataSourceV2 read entry point. No files or tests are named, so the ticket should define the implementation scope and validation. Done means DataSourceV2 reads support the requested pushdown behavior.

Written by the indexing model from the issue text.

Assessment

Tech stack
java, spark
Domain
data-engineering
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.