Sign In Start Free Trial
Account

Add to playlist

Create a Playlist

Modal Close icon
You need to login to use this feature.
  • Book Overview & Buying OpenStack Sahara Essentials
  • Table Of Contents Toc
OpenStack Sahara Essentials

OpenStack Sahara Essentials

By : Omar Khedher
close
close
OpenStack Sahara Essentials

OpenStack Sahara Essentials

By: Omar Khedher

Overview of this book

The Sahara project is a module that aims to simplify the building of data processing capabilities on OpenStack. The goal of this book is to provide a focused, fast paced guide to installing, configuring, and getting started with integrating Hadoop with OpenStack, using Sahara. The book should explain to users how to deploy their data-intensive Hadoop and Spark clusters on top of OpenStack. It will also cover how to use the Sahara REST API, how to develop applications for Elastic Data Processing on Openstack, and setting up hadoop or spark clusters on Openstack.
Table of Contents (9 chapters)
close
close

Chapter 1. The Essence of Big Data in the Cloud

How to quantify data into business value? It's a serious question that we might be prompted to ask when we take a look around and notice the increasing appetite of users for rich media and the content of data across the web. That could generate several challenging points: How to manage the exponential amount of data? Particularly, how to extract from these immense waves of data the most valuable aspects? It is the era of big data! To meet the growing demand of big data and facilitate its analysis, few solutions such as Hadoop and Spark appeared and have become a necessary tool towards making a first successful step into the big data world. However, the first question was not sufficiently answered! It might be needed to introduce a new architecture and cost approach to respond to the scalability of intensive resources consumed when analyzing data. Although Hadoop, for example, is a great solution to run data analysis and processing, there are difficulties with configuration and maintenance. Besides, its complex architecture might require a lot of expertise. In this book, you will learn how to use OpenStack to manage and rapidly configure a Hadoop/Spark cluster. Sahara, the new OpenStack integrated project, offers an elegant self-service to deploy and manage big data clusters. It began as an Apache 2.0 project and now Sahara has joined the OpenStack ecosystem to provide a fast way of provisioning Hadoop clusters in the cloud. In this chapter, we will explore the following points:

  • Introduce briefly the big data groove
  • Understand the success of big data processing when it is combined with the cloud computing paradigm
  • Learn how OpenStack can offer a unique big data management solution
  • Discover Sahara in OpenStack and cover briefly the overall architecture
CONTINUE READING
83
Tech Concepts
36
Programming languages
73
Tech Tools
Icon Unlimited access to the largest independent learning library in tech of over 8,000 expert-authored tech books and videos.
Icon Innovative learning tools, including AI book assistants, code context explainers, and text-to-speech.
Icon 50+ new titles added per month and exclusive early access to books as they are being written.
OpenStack Sahara Essentials
notes
bookmark Notes and Bookmarks search Search in title playlist Add to playlist font-size Font size

Change the font size

margin-width Margin width

Change margin width

day-mode Day/Sepia/Night Modes

Change background colour

Close icon Search
Country selected

Close icon Your notes and bookmarks

Confirmation

Modal Close icon
claim successful

Buy this book with your credits?

Modal Close icon
Are you sure you want to buy this book with one of your credits?
Close
YES, BUY

Submit Your Feedback

Modal Close icon
Modal Close icon
Modal Close icon