Deep dive on HDFS HA and disaster recovery

Open
#682 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Assessment

Difficulty
4/5
Estimated time
3-5 days
Newbie friendliness
25/100
Issue type
Documentation
Clarity
Needs clarification
Activity status
Stale
Tech stack
hadoop

Research direction

Start by locating the HDFS HA and disaster-recovery documentation entry points in this Antora repository. Research failure and recovery behavior for NameNodes, DataNodes, and JournalNodes, including client impact, detection time, recovery time, and permanent node-data loss. Done means the relevant documentation explains each scenario and restoration path clearly.

Written by the indexing model from the issue text.

Description

  • What happens to active clients if different nodes die? (namenode, datanode, journalnode)
    • How long does detection/recovery take?
  • What happens if we permanently lose the data of a node, how can the data be restored? (again, all three)
Dominant language
CSS
Stars
13
Forks
14
Avg merge
4d 8h
Merged PRs (30d)
10

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

More from stackabletech/documentation

All issues in stackabletech/documentation

Similar issues

More Distributed Systems issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.