Virtualizing Hadoop: How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture

by George J. Trujillo Jr.

Released July 2015

Publisher(s): VMware Press

ISBN: 9780133812350

Start your free trial

Book description

Plan and Implement Hadoop Virtualization for Maximum Performance, Scalability, and Business Agility

Enterprises running Hadoop must absorb rapid changes in big data ecosystems, frameworks, products, and workloads. Virtualized approaches can offer important advantages in speed, flexibility, and elasticity. Now, a world-class team of enterprise virtualization and big data experts guide you through the choices, considerations, and tradeoffs surrounding Hadoop virtualization. The authors help you decide whether to virtualize Hadoop, deploy Hadoop in the cloud, or integrate conventional and virtualized approaches in a blended solution.

First, Virtualizing Hadoop reviews big data and Hadoop from the standpoint of the virtualization specialist. The authors demystify MapReduce, YARN, and HDFS and guide you through each stage of Hadoop data management. Next, they turn the tables, introducing big data experts to modern virtualization concepts and best practices.

Finally, they bring Hadoop and virtualization together, guiding you through the decisions you’ll face in planning, deploying, provisioning, and managing virtualized Hadoop. From security to multitenancy to day-to-day management, you’ll find reliable answers for choosing your best Hadoop strategy and executing it.

Coverage includes the following:

• Reviewing the frameworks, products, distributions, use cases, and roles associated with Hadoop

• Understanding YARN resource management, HDFS storage, and I/O

• Designing data ingestion, movement, and organization for modern enterprise data platforms

• Defining SQL engine strategies to meet strict SLAs

• Considering security, data isolation, and scheduling for multitenant environments

• Deploying Hadoop as a service in the cloud

• Reviewing the essential concepts, capabilities, and terminology of virtualization

• Applying current best practices, guidelines, and key metrics for Hadoop virtualization

• Managing multiple Hadoop frameworks and products as one unified system

• Virtualizing master and worker nodes to maximize availability and performance

• Installing and configuring Linux for a Hadoop environment

Product information

Title: Virtualizing Hadoop: How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture
Author(s): George J. Trujillo Jr.
Release date: July 2015
Publisher(s): VMware Press
ISBN: 9780133812350

book

Moving Hadoop to the Cloud

by Bill Havanki

Until recently, Hadoop deployments existed on hardware owned and run by organizations. Now, of course, you …

book

Hadoop 2 Quick-Start Guide: Learn the Essentials of Big Data Computing in the Apache Hadoop 2 Ecosystem

by Douglas Eadline

Get Started Fast with Apache Hadoop ® 2, YARN, and Today’s Hadoop Ecosystem With Hadoop 2.x …

book

Apache Hadoop 3 Quick Start Guide

by Hrishikesh Vijay Karambelkar

A fast paced guide that will help you learn about Apache Hadoop 3 and its ecosystem …

book

Hadoop Security

by Ben Spivey, Joey Echeverria

As more corporations turn to Hadoop to store and process their most valuable data, the risk …

Virtualizing Hadoop: How to Install, Deploy, and Optimize Hadoop in a Virtualized Architecture

Book description

Table of contents

Product information

You might also like

Moving Hadoop to the Cloud

Hadoop 2 Quick-Start Guide: Learn the Essentials of Big Data Computing in the Apache Hadoop 2 Ecosystem

Apache Hadoop 3 Quick Start Guide

Hadoop Security

Don’t leave empty-handed

It’s yours, free.

Check it out now on O’Reilly