πŸŽ“ Computer Science & Engineering Portal

Master Engineering Disciplines with Structured Notes

Comprehensive academic lecture notes, exam-oriented unit summaries, laboratory manuals, and previous year question papers designed strictly for university students.

πŸ“‘

University Syllabi

AKTU & AICTE aligned semester guidelines.

πŸ“

Exam Question Papers

Previous 5 years solved university papers.

πŸ’‘

Lab Manuals & Viva

Practical codes with outputs & interview Qs.

Core Subjects & Units Hub

Click on any specific unit to immediately view its lecture notes below

Viewing All Lectures

Data Replication

Data Replication :

  • Replication ensures the availability of the data. 
  • Replication is - making a copy of something and the

number of times you make a copy of that particular thing can be expressed as its Replication Factor. 

  • As HDFS stores the data in the form of various blocks at the same time Hadoop is also configured to make a copy of those file blocks. 
  • By default, the Replication Factor for Hadoop is set to 3 which can be configured.
  • We need this replication for our file blocks because for running Hadoop we are using commodity hardware (inexpensive system hardware) which can be crashed at any time.
  • We are not using a supercomputer for our Hadoop setup. 
  • That is why we need such a feature in HDFS that can make copies of that file blocks for backup purposes, this is known as fault tolerance.
  • For the big brand organization, the data is very much important than the storage, so nobody cares about this extra storage.
  • You can configure the Replication factor in your hdfs-site.xml file.


Labels: ,

Discussion & Queries (<$I18NNumComments$>):

<$CommentPager$>
<$I18NAtCommentTimeWithPermalink$>, <$I18NCommentAuthorSaid$>

<$BlogCommentBody$>

<$BlogCommentDeleteIcon$>
<$CommentPager$>