apache / apache/parquet-java

Increment a hadoop counter for bytes filtered / skipped

未关闭
#1,393 0 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
Component: Java Component: Parquet Priority: Minor Type: enhancement
主要语言
Java
星标
3.1k
派生
1.6k
平均合并
3 天 12 小时
30 天内合并 PR
33

描述

**Reporter**: [Alex Levenson](https://issues.apache.org/jira/secure/ViewProfile.jspa?name=alexlevenson) / @isnotinvain

**Note**: *This issue was originally created as [PARQUET-45](https://issues.apache.org/jira/browse/PARQUET-45). Please see the [migration documentation](https://issues.apache.org/jira/browse/PARQUET-2502) for further details.*

贡献指南

这个仓库没有索引到贡献指南

调研方向

该 issue 提到了一个用于统计被过滤或跳过字节数的 Hadoop counter,但没有指出任何文件、测试或入口点。首先定位过滤或跳过路径,以及现有的 Hadoop counter 处理逻辑;完成的标准是该 counter 能够报告相关的过滤和跳过字节总数。

由索引模型根据 Issue 内容生成。

评估

技术栈
hadoop, java
领域
data-engineering
Issue 类型
功能
难度
3/5
预计耗时
1-2 天
活跃度
停滞
描述清晰度
需要澄清
新手友好度
25/100

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。