apache / apache/hudi

Spark: Support DataSourceV2 Read

Open
#15,292 1 comment 0 reactions 1 assignee Claimed by @nsivabalan View on GitHub
area:reader area:sql engine:spark from-jira priority:high status:pr-available type:epic
Dominant language
Java
Stars
6.2k
Forks
2.5k
Avg merge
2d 4h
Merged PRs (30d)
112

Description

Introduce v2 reading interface and define {{HoodieBatchScanBuilder}} to provide querying capability.ColumnPrune & PushDown  is follow up.

## JIRA info

- Link: https://issues.apache.org/jira/browse/HUDI-4449
- Type: Epic
- Fix version(s):
- 1.2.0

---

## Comments

30/May/25 06:59;geserdugarov;There is no progress in this task for a while. I wanna try to take an action in this direction, but will start from RFC with design of calls using DSv2.;;;

---

02/Jun/25 09:43;geserdugarov;There is already merged corresponding RFC-38 [https://github.com/apache/hudi/blob/master/rfc/rfc-38/rfc-38.md] .;;;

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.