stackabletech / stackabletech/documentation

Deep dive on HDFS HA and disaster recovery

Open
#682 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
CSS
Stars
13
Forks
14
Avg merge
4d 8h
Merged PRs (30d)
10

Description

  • What happens to active clients if different nodes die? (namenode, datanode, journalnode)
    • How long does detection/recovery take?
  • What happens if we permanently lose the data of a node, how can the data be restored? (again, all three)

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by locating the HDFS HA and disaster-recovery documentation entry points in this Antora repository. Research failure and recovery behavior for NameNodes, DataNodes, and JournalNodes, including client impact, detection time, recovery time, and permanent node-data loss. Done means the relevant documentation explains each scenario and restoration path clearly.

Written by the indexing model from the issue text.

Assessment

Tech stack
hadoop
Domain
distributed-systems, documentation
Issue type
Documentation
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.