apache / apache/hudi

how to get the root case of the error records

Open
#13,671 5 comments 0 reactions 0 assignees View on GitHub
area:sql type:bug type:community-support
Dominant language
Java
Stars
6.2k
Forks
2.5k
Avg merge
2d 8h
Merged PRs (30d)
111

Description

how to get the root case of the error records? Each file has many errors. How to get the error message for these records?
is there any way to prevent these error records and let the job failed when there is some error messages.

Image

**Environment Description**

* Hudi version : 0.13.1

* Spark version : 3.3.2

* Hive version : 3.1

* Hadoop version : 3.2

* Storage (HDFS/S3/GCS..) : s3

* Running on Docker? (yes/no) : no

Contributor guide

No contributing guide indexed for this repository

Research direction

No source file, test, or entry point is named. Start by reproducing the error-record behavior with Hudi 0.13.1, Spark 3.3.2, Hadoop 3.2, and S3, then inspect the attached error output and issue comments; done means identifying the root cause and confirming how failures should be surfaced.

Written by the indexing model from the issue text.

Assessment

Tech stack
aws, hadoop, java, spark
Domain
data-engineering
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.