I am new to Hadoop and i know HDFS is 64 mb (min) per block and can increase depending on the system. but as hdfs is installed on top of linux filesystem which is 4kb per block, does hadoop not suffer disk seek? also does hdfs interact with linux filesystem ?
does hadoop not suffer the disk seeks as it sits on top of linux filesystem?
177 Views Asked by Tumpiri Sydney Rockwell At
1
There are 1 best solutions below
Related Questions in HADOOP
- Can I use plone.protect 3.0 with Plone 4.3?
- How do I override the main template in Plone 3?
- Create copy of plone installed onto another server with data
- Where is the main space used up in Plone file upload?
- Create a folder in Plone and set uid
- Plone - Syntax Error when doing hello world tutorial
- collective.googleanalytics report with filter containing tag
- Redirecting Site Root URL to its Language Root Folder
- How to move a content type definition created TTW to the file system
- Multiple contact forms in a Plone website
Related Questions in HDFS
- Can I use plone.protect 3.0 with Plone 4.3?
- How do I override the main template in Plone 3?
- Create copy of plone installed onto another server with data
- Where is the main space used up in Plone file upload?
- Create a folder in Plone and set uid
- Plone - Syntax Error when doing hello world tutorial
- collective.googleanalytics report with filter containing tag
- Redirecting Site Root URL to its Language Root Folder
- How to move a content type definition created TTW to the file system
- Multiple contact forms in a Plone website
Trending Questions
- UIImageView Frame Doesn't Reflect Constraints
- Is it possible to use adb commands to click on a view by finding its ID?
- How to create a new web character symbol recognizable by html/javascript?
- Why isn't my CSS3 animation smooth in Google Chrome (but very smooth on other browsers)?
- Heap Gives Page Fault
- Connect ffmpeg to Visual Studio 2008
- Both Object- and ValueAnimator jumps when Duration is set above API LvL 24
- How to avoid default initialization of objects in std::vector?
- second argument of the command line arguments in a format other than char** argv or char* argv[]
- How to improve efficiency of algorithm which generates next lexicographic permutation?
- Navigating to the another actvity app getting crash in android
- How to read the particular message format in android and store in sqlite database?
- Resetting inventory status after order is cancelled
- Efficiently compute powers of X in SSE/AVX
- Insert into an external database using ajax and php : POST 500 (Internal Server Error)
Popular # Hahtags
Popular Questions
- How do I undo the most recent local commits in Git?
- How can I remove a specific item from an array in JavaScript?
- How do I delete a Git branch locally and remotely?
- Find all files containing a specific text (string) on Linux?
- How do I revert a Git repository to a previous commit?
- How do I create an HTML button that acts like a link?
- How do I check out a remote Git branch?
- How do I force "git pull" to overwrite local files?
- How do I list all files of a directory?
- How to check whether a string contains a substring in JavaScript?
- How do I redirect to another webpage?
- How can I iterate over rows in a Pandas DataFrame?
- How do I convert a String to an int in Java?
- Does Python have a string 'contains' substring method?
- How do I check if a string contains a specific word?
Your thinking is correct to certain extent but look at the bigger picture. When this 64 MB is stored on the Linux file system, it is distributed across many nodes. Consequently, if you want to read 3 blocks (each 4 KB), stored on 3 different Linux file systems (machines), the seek will be for only 1 seek and not 3 seeks as reading will be in parallel.
I think this might help: How are HDFS files getting stored on underlying OS filesystem?