Have a personal or library account? Click to login
OpenStack Sahara Essentials Cover

OpenStack Sahara Essentials

Integrate, deploy, rapidly configure, and successfully manage your own big data-intensive clusters in the cloud using OpenStack Sahara

Paid access
|May 2016
Product purchase options

Integrate, deploy, rapidly configure, and successfully
manage your own big data-intensive clusters in the
cloud using OpenStack Sahara

Key Features

  • A fast paced guide to help you utilize the benefits of Sahara in OpenStack to meet the Big Data world of Hadoop.
  • A step by step approach to simplify the complexity of Hadoop configuration, deployment and maintenance.

Book Description

The Sahara project is a module that aims to simplify the building of data processing capabilities on OpenStack.

The goal of this book is to provide a focused, fast paced guide to installing, configuring, and getting started with integrating Hadoop with OpenStack, using Sahara.

The book should explain to users how to deploy their data-intensive Hadoop and Spark clusters on top of OpenStack. It will also cover how to use the Sahara REST API, how to develop applications for Elastic Data Processing on Openstack, and setting up hadoop or spark clusters on Openstack.

What you will learn

  • Integrate and Install Sahara with OpenStack environment
  • Learn Sahara architecture under the hood
  • Rapidly configure and scale Hadoop clusters on top of OpenStack
  • Explore the Sahara REST API to create, deploy and manage a Hadoop cluster
  • Learn the Elastic Processing Data (EDP) facility to execute jobs in clusters from Sahara
  • Cover other Hadoop stable plugins existing supported by Sahara
  • Discover different features provided by Sahara for Hadoop provisioning and deployment
  • Learn how to troubleshoot OpenStack Sahara issues

Who this book is for

This book targets data scientists, cloud developers and Devops Engineers who would like to become proficient with OpenStack Sahara. Ideally, this book is well suitable for readers who are familiars with databases, Hadoop and Spark solutions. Additionally, a basic prior knowledge of OpenStack is expected. The readers should also be familiar with different Linux boxes, distributions and virtualization technology.

Table of Contents

  1. The essence of Big Data in the Cloud
  2. Integrating OpenStack Sahara
  3. Using OpenStack Sahara
  4. Executing jobs with Sahara
  5. Discovering advanced features with Sahara
  6. Hadoop High Availability using Sahara
  7. Troubleshooting
PDF ISBN: 978-1-78588-014-8
Publisher: Packt Publishing Limited
Copyright owner: © 2016 Packt Publishing Limited
Publication date: 2016
Language: English
Pages: 178
Related subjects: