alibaba / alibaba/DataX

hbase2hive有抽数瓶颈吗

Open
#2,107 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Java
Stars
17.4k
Forks
5.7k
PR merge metrics
No merged PRs in 30d

Description

单表数据量大概2亿,选用habse reader和hdfs writer,有无更好的解决方案呢

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names only the HBase reader and HDFS writer; no files, tests, configuration, or entry points are provided. Start by locating those components and defining a reproducible benchmark for a roughly 200-million-row table. Done would require a documented bottleneck analysis and a justified alternative or configuration recommendation.

Written by the indexing model from the issue text.

Assessment

Tech stack
java
Domain
data-engineering, databases
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
18/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.