apache / apache/parquet-java

Decouple Hadoop/Native I/O implementations

Open
#3,279 1 comment 5 reactions 0 assignees View on GitHub
Type: enhancement
Dominant language
Java
Stars
3.1k
Forks
1.6k
Avg merge
3d 12h
Merged PRs (30d)
33

Description

### Describe the enhancement requested

Provide a clean interface to use hadoop/native io classes. Users should be able to use parquet without hadoop on the classpath. This issue tracks the effort

### Component(s)

_No response_

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names no files, tests, or entry points. Start by mapping the Hadoop/native I/O implementations in parquet-java and their classpath dependencies; done means Parquet can use a clean interface without Hadoop on the classpath.

Written by the indexing model from the issue text.

Assessment

Tech stack
java
Domain
backend, data-engineering
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.