In this chapter, you explored the factors behind the success of the emerging technology of data processing and analysis using cloud computing technology. You learned how OpenStack can be a great opportunity to offer the needed scalable and elastic big data on-demand infrastructure. It can be also useful to execute on-demand Elastic Data Processing tasks.
The first chapter exposed the new OpenStack incubated project called Sahara: a rapid, auto-deploy, and scalable solution for Hadoop and Spark clusters. An overall view of the Sahara architecture has been discussed for a fast-paced understanding of the platform and how it works in an OpenStack private cloud environment.
Now it is time to get things running and discover how such an amazing big data management solution can be used by installing OpenStack and integrating Sahara, which will be the topic of the next chapter.