Column resolution by ID
- Dominant language
- Java
- Stars
- 3.1k
- Forks
- 1.6k
- Avg merge
- 3d 12h
- Merged PRs (30d)
- 33
Description
Parquet relies on the name. In a lot of usages e.g. schema resolution, this would be a problem. Iceberg uses ID and stored Id/name mappings.
This Jira is to add column ID resolution support.
**Reporter**: [Xinli Shang](https://issues.apache.org/jira/secure/ViewProfile.jspa?name=shangx@uber.com) / @shangxinli
**Assignee**: [Xinli Shang](https://issues.apache.org/jira/secure/ViewProfile.jspa?name=shangx@uber.com) / @shangxinli
#### PRs and other links:
- [GitHub Pull Request #950](https://github.com/apache/parquet-java/pull/950)
- [GitHub Pull Request #950](https://github.com/apache/parquet-mr/pull/950)
**Note**: *This issue was originally created as [PARQUET-2006](https://issues.apache.org/jira/browse/PARQUET-2006). Please see the [migration documentation](https://issues.apache.org/jira/browse/PARQUET-2502) for further details.*
Contributor guide
No contributing guide indexed for this repository
Research direction
Start by reviewing the referenced GitHub Pull Request #950 and the migration documentation linked from the issue, since no source files or tests are named here. The stated goal is column ID resolution support for schema resolution, but the implementation scope and completion checks are not specified.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- java
- Domain
- data-engineering
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 15/100