Reuse hadoop file status and footer in ParquetRecordReader
オープン
Type: enhancement
- 主要言語
- Java
- スター
- 3.1k
- フォーク
- 1.6k
- 平均マージ
- 3日 12時間
- マージ済み PR(30日)
- 33
説明
### Describe the enhancement requested
This is actually [PARQUET-2415](https://issues.apache.org/jira/browse/PARQUET-2415)
_Reuse hadoop file status and footer in ParquetRecordReader_ moved to a github issue.
That has a stale pr #1242 by @wankunde; I've got claude to rebase it and update to junit5/assertj
### Component(s)
Core
コントリビューションガイド
このリポジトリのコントリビューションガイドは索引されていません
調査の方向性
まず core モジュール内の ParquetRecordReader を見つけ、既存の実装として言及されている古い PR #1242 を確認します。関連する core reader tests を実行してカバレッジを比較します。reader が Hadoop の file status と footer を再利用し、リグレッションが発生しないことが完了の条件です。
索引モデルが issue の本文から書いたものです。
評価
- 技術スタック
- hadoop, java
- 領域
- data-engineering
- issue の種類
- 機能追加
- 難易度
- 4/5
- 見積もり時間
- 3〜5日
- 活発さ
- 停滞
- 明瞭さ
- おおむね明確
- 初心者へのやさしさ
- 35/100