Apache Hive Essentials: Essential techniques to help you process, and get unique insights from, big data, 2nd Edition
This book takes you on a fantastic journey to discover the attributes of big data using Apache Hive.
In this book, we prepare you for your journey into big data by frstly introducing you to backgrounds in the big data domain, alongwith the process of setting up and getting familiar with your Hive working environment.
Next, the book guides you through discovering and transforming the values of big data with the help of examples. It also hones your skills in using the Hive language in an effcient manner. Toward the end, the book focuses on advanced topics, such as performance, security, and extensions in Hive, which will guide you on exciting adventures on this worthwhile big data journey.
By the end of the book, you will be familiar with Hive and able to work effeciently to find solutions to big data problems
If you are a data analyst, developer, or simply someone who wants to quickly get started with Hive to explore and analyze Big Data in Hadoop, this is the book for you. Since Hive is an SQL-like language, some previous experience with SQL will be useful to get the most out of this book.
Chapter 1. OVERVIEW OF BIG DATA AND HIVE
Chapter 2. SETTING UP THE HIVE ENVIRONMENT
Chapter 3. DATA DEFINITION AND DESCRIPTION
Chapter 4. Data Correlation and Scope
Chapter 5. DATA MANIPULATION
Chapter 6. DATA AGGREGATION AND SAMPLING
Chapter 7. Extensibility Considerations
Chapter 8. Working with Other Tools
Chapter 9. Performance Considerations
Chapter 10. Security Considerations