upgrade from 0.5.0 to 0.13.0
- Dominant language
- Java
- Stars
- 6.2k
- Forks
- 2.5k
- Avg merge
- 2d 8h
- Merged PRs (30d)
- 111
Description
Our organization is migrating from Hadoop 2.6 to 3.1 and spark 2.3 to 3.1. Our existing data sets (100s of TBs of data) are written using Hudi 0.5.0. We would like to move to Hudi 0.13.0.
The latest version of Hudi has come way since 0.5.0, we are not sure about how to use 0.13.0 directly.
Could someone provide the steps for upgrading from 0.5.0 to 0.13.0?
Thanks,
Selva
Contributor guide
No contributing guide indexed for this repository
Research direction
No files, tests, or entry points are named. Start by reviewing the compatibility and migration differences between Hudi 0.5.0 and 0.13.0 alongside the Hadoop 2.6-to-3.1 and Spark 2.3-to-3.1 changes. Done means documenting reliable upgrade steps for existing datasets and calling out any required intermediate versions or migration risks.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- hadoop, spark
- Domain
- data-engineering, distributed-systems
- Issue type
- Documentation
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100