What is the main difference between bucketing and indexing of a table in Hive?
Hive(Bigdata)- difference between bucketing and indexing
2k Views Asked by anusngh At
1
There are 1 best solutions below
Related Questions in HADOOP
- Doctrine batch inserting uses 2GB of Ram
- Persisting other entities inside preUpdate of Doctrine Entity Listener
- doctrine/migrations incompatible with symfony 2.2.*
- Symfony2 - Share Entity Between Bundles with different relationships
- ZF2 / Doctrine Form Multi-select Element for Many-to-Many Relation
- Symfony2 - Doctrine2 Respository - Set Where Condition for All Methods
- symfony many to many orm controller
- Symfony2/Doctrine: Get the field(s) that changed after "Loggable" entity changed
- Call setter method with variable name
- Check If Record Has References
Related Questions in MAPREDUCE
- Doctrine batch inserting uses 2GB of Ram
- Persisting other entities inside preUpdate of Doctrine Entity Listener
- doctrine/migrations incompatible with symfony 2.2.*
- Symfony2 - Share Entity Between Bundles with different relationships
- ZF2 / Doctrine Form Multi-select Element for Many-to-Many Relation
- Symfony2 - Doctrine2 Respository - Set Where Condition for All Methods
- symfony many to many orm controller
- Symfony2/Doctrine: Get the field(s) that changed after "Loggable" entity changed
- Call setter method with variable name
- Check If Record Has References
Related Questions in HIVE
- Doctrine batch inserting uses 2GB of Ram
- Persisting other entities inside preUpdate of Doctrine Entity Listener
- doctrine/migrations incompatible with symfony 2.2.*
- Symfony2 - Share Entity Between Bundles with different relationships
- ZF2 / Doctrine Form Multi-select Element for Many-to-Many Relation
- Symfony2 - Doctrine2 Respository - Set Where Condition for All Methods
- symfony many to many orm controller
- Symfony2/Doctrine: Get the field(s) that changed after "Loggable" entity changed
- Call setter method with variable name
- Check If Record Has References
Related Questions in BIGDATA
- Doctrine batch inserting uses 2GB of Ram
- Persisting other entities inside preUpdate of Doctrine Entity Listener
- doctrine/migrations incompatible with symfony 2.2.*
- Symfony2 - Share Entity Between Bundles with different relationships
- ZF2 / Doctrine Form Multi-select Element for Many-to-Many Relation
- Symfony2 - Doctrine2 Respository - Set Where Condition for All Methods
- symfony many to many orm controller
- Symfony2/Doctrine: Get the field(s) that changed after "Loggable" entity changed
- Call setter method with variable name
- Check If Record Has References
Trending Questions
- UIImageView Frame Doesn't Reflect Constraints
- Is it possible to use adb commands to click on a view by finding its ID?
- How to create a new web character symbol recognizable by html/javascript?
- Why isn't my CSS3 animation smooth in Google Chrome (but very smooth on other browsers)?
- Heap Gives Page Fault
- Connect ffmpeg to Visual Studio 2008
- Both Object- and ValueAnimator jumps when Duration is set above API LvL 24
- How to avoid default initialization of objects in std::vector?
- second argument of the command line arguments in a format other than char** argv or char* argv[]
- How to improve efficiency of algorithm which generates next lexicographic permutation?
- Navigating to the another actvity app getting crash in android
- How to read the particular message format in android and store in sqlite database?
- Resetting inventory status after order is cancelled
- Efficiently compute powers of X in SSE/AVX
- Insert into an external database using ajax and php : POST 500 (Internal Server Error)
Popular # Hahtags
Popular Questions
- How do I undo the most recent local commits in Git?
- How can I remove a specific item from an array in JavaScript?
- How do I delete a Git branch locally and remotely?
- Find all files containing a specific text (string) on Linux?
- How do I revert a Git repository to a previous commit?
- How do I create an HTML button that acts like a link?
- How do I check out a remote Git branch?
- How do I force "git pull" to overwrite local files?
- How do I list all files of a directory?
- How to check whether a string contains a substring in JavaScript?
- How do I redirect to another webpage?
- How can I iterate over rows in a Pandas DataFrame?
- How do I convert a String to an int in Java?
- Does Python have a string 'contains' substring method?
- How do I check if a string contains a specific word?
The main difference is the goal:
Indexes become even more essential when the tables grow extremely large, and as you now undoubtedly know, Hive thrives on large tables.
It is usually used for join operations, because you can optimize joins by bucketing records by a specific 'key' or 'id'. In this way, when you want to do a join operation, records with the same 'key' will be in the same bucket and then the join operation will be faster. You can see this like a technique for decomposing data sets into more manageable parts. This link gives you 5 Tips for efficient Hive queries and one of them is about Bucketing.