Transcription of HDFS Architecture Guide - Apache Hadoop
{{id}} {{{paragraph}}}
Copyright 2008 The Apache Software Foundation. All rights Architecture Guideby Dhruba BorthakurTable of contents1 introduction .. 32 Assumptions and Goals .. 3 Hardware Failure .. 3 Streaming Data Access ..3 Large Data Sets .. 3 Simple Coherency Model .. 3 Moving Computation is Cheaper than Moving Data .. 4 Portability Across Heterogeneous Hardware and Software Platforms ..43 NameNode and DataNodes ..44 The File System Namespace .. 55 Data Replication ..6 Replica Placement: The First Baby Steps .. 7 Replica Selection .. 7 Safemode .. 86 The Persistence of File System Metadata ..87 The Communication Protocols .. 98 Robustness ..9 Data Disk Failure, Heartbeats and Re-Replication.
HDFS Architecture Guide ... 1 Introduction The Hadoop Distributed File System (HDFS) is a distributed file system designed to run on commodity hardware. It has many similarities with existing distributed file systems. ... replicas of a file do not evenly distribute across the racks. One third of replicas are on one
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document:
{{id}} {{{paragraph}}}
Backward design, ASCD, Introduction, 1 Introduction to Design and the, Design, Community Design Action Team – Strategy Map Introduction, Introduction to Geographic Information Systems, Map design, Graphic design, Introduction to Maps, Value stream mapping, University of Washington, 3 for Business Process Modelling, Introduction to the Quartus II, Introduction to the Quartus ® II