HDFS as a destination
- Dominant language
- Python
- Stars
- 5.9k
- Forks
- 600
- Avg merge
- 1d 14h
- Merged PRs (30d)
- 38
Description
### Feature description
I would like to request a new feature in DLT that allows users to use HDFS (Hadoop Distributed File System) as a destination for data pipelines. This feature would enable users to easily store and manage large volumes of data in HDFS, which is commonly used in big data and data lake environments
### Are you a dlt user?
I'd consider using dlt, but it's lacking a feature I need.
### Use case
_No response_
### Proposed solution
_No response_
### Related issues
_No response_
Contributor guide
Research direction
No files, tests, or entry points are identified. Start by reviewing dlt's existing destination integrations and the requirements for connecting to HDFS; define the supported HDFS behavior and completion criteria before implementation.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- hadoop, python
- Domain
- data-engineering
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100