Book Image

OpenStack Sahara Essentials

By : Omar Khedher
Book Image

OpenStack Sahara Essentials

By: Omar Khedher

Overview of this book

The Sahara project is a module that aims to simplify the building of data processing capabilities on OpenStack. The goal of this book is to provide a focused, fast paced guide to installing, configuring, and getting started with integrating Hadoop with OpenStack, using Sahara. The book should explain to users how to deploy their data-intensive Hadoop and Spark clusters on top of OpenStack. It will also cover how to use the Sahara REST API, how to develop applications for Elastic Data Processing on Openstack, and setting up hadoop or spark clusters on Openstack.
Table of Contents (14 chapters)

Summary


The main reason for Sahara's continuity support in OpenStack is the overwhelming features list implemented. The good modularity of the Sahara project and its integration in OpenStack facilitates the usage of more advanced configurations and performs more customizations. Although the previous chapters have covered the usage of Sahara plugins and how to provision a Hadoop cluster in no time, it misses a trivial feature that should be detailed in this chapter: high availability. The next chapter will discuss which Sahara plugins support the high-availability feature on cluster provisioning and how to configure it.