Automated Machine Learning on AWS

By : Trenton Potgieter

Automated Machine Learning on AWS

By: Trenton Potgieter

Overview of this book

AWS provides a wide range of solutions to help automate a machine learning workflow with just a few lines of code. With this practical book, you'll learn how to automate a machine learning pipeline using the various AWS services. Automated Machine Learning on AWS begins with a quick overview of what the machine learning pipeline/process looks like and highlights the typical challenges that you may face when building a pipeline. Throughout the book, you'll become well versed with various AWS solutions such as Amazon SageMaker Autopilot, AutoGluon, and AWS Step Functions to automate an end-to-end ML process with the help of hands-on examples. The book will show you how to build, monitor, and execute a CI/CD pipeline for the ML process and how the various CI/CD services within AWS can be applied to a use case with the Cloud Development Kit (CDK). You'll understand what a data-centric ML process is by working with the Amazon Managed Services for Apache Airflow and then build a managed Airflow environment. You'll also cover the key success criteria for an MLSDLC implementation and the process of creating a self-mutating CI/CD pipeline using AWS CDK from the perspective of the platform engineering team. By the end of this AWS book, you'll be able to effectively automate a complete machine learning pipeline and deploy it to production.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Download the color images

Conventions used

Get in touch

Share Your Thoughts

Section 1: Fundamentals of the Automated Machine Learning Process and AutoML on AWS

Free Chapter

Chapter 1: Getting Started with Automated Machine Learning on AWS

Technical requirements

Overview of the ML process

Complexities in the ML process

An example of the end-to-end ML process

How AWS makes automating the ML development and deployment process easier

Summary

Chapter 2: Automating Machine Learning Model Development Using SageMaker Autopilot

Technical requirements

Introducing the AWS AI and ML landscape

Overview of SageMaker Autopilot

Overcoming automation challenges with SageMaker Autopilot

Using the SageMaker SDK to automate the ML experiment

Summary

Chapter 3: Automating Complicated Model Development with AutoGluon

Technical requirements

Introducing the AutoGluon library

Using AutoGluon for tabular data

Using AutoGluon for image data

Summary

Section 2: Automating the Machine Learning Process with Continuous Integration and Continuous Delivery (CI/CD)

Chapter 4: Continuous Integration and Continuous Delivery (CI/CD) for Machine Learning

Technical requirements

Introducing the CI/CD methodology

Automating ML with CI/CD

Creating a CI/CD pipeline on AWS

Summary

Chapter 5: Continuous Deployment of a Production ML Model

Technical requirements

Deploying the CI/CD pipeline

Building the ML model artifacts

Executing the automated ML model deployment

Summary

Section 3: Optimizing a Source Code-Centric Approach to Automated Machine Learning

Chapter 6: Automating the Machine Learning Process Using AWS Step Functions

Technical requirements

Introducing AWS Step Functions

Using the Step Functions Data Science SDK for CI/CD

Building the CI/CD pipeline resources

Summary

Chapter 7: Building the ML Workflow Using AWS Step Functions

Technical requirements

Building the state machine workflow

Performing the integration test

Monitoring the pipeline's progress

Summary

Section 4: Optimizing a Data-Centric Approach to Automated Machine Learning

Chapter 8: Automating the Machine Learning Process Using Apache Airflow

Technical requirements

Introducing Apache Airflow

Introducing Amazon MWAA

Using Airflow to process the abalone dataset

Configuring the MWAA prerequisites

Configuring the MWAA environment

Summary

Chapter 9: Building the ML Workflow Using Amazon Managed Workflows for Apache Airflow

Technical requirements

Developing the data-centric workflow

Creating synthetic Abalone survey data

Executing the data-centric workflow

Summary

Section 5: Automating the End-to-End Production Application on AWS

Chapter 10: An Introduction to the Machine Learning Software Development Life Cycle (MLSDLC)

Technical requirements

Introducing the MLSDLC

Building the application platform

Examining ML and data engineering roles

Understanding the security lens

Summary

Chapter 11: Continuous Integration, Deployment, and Training for the MLSDLC

Technical requirements

Codifying the continuous integration stage

Managing the continuous deployment stage

Managing continuous training

Summary

Examining ML and data engineering roles

In previous chapters, we have used the term ML practitioner as a blanket term for any person responsible for automating the ML process. Within the context of the MLSDLC process, we typically see this role split into two distinct functions, namely the following:

Data scientist: The data scientist is primarily responsible for building, training, and tuning an ML model that meets the business requirements of the use case.
ML engineer: Among numerous responsibilities, the ML engineer is primarily responsible for designing the overall ML system to support the model, managing the appropriate datasets for model training, and ensuring the final ML application addresses the business requirements for the use case.

However, for the sake of the ACME application example, we will group these two functions under the banner of the ML team, with the following diagram highlighting how this team fits into the MLSDLC process:

...

Automated Machine Learning on AWS

By : Trenton Potgieter

Automated Machine Learning on AWS

By: Trenton Potgieter

Overview of this book

Related Content you might be interested in

Current Title:

Automated Machine Learning on AWS

Getting Started with Amazon SageMaker Studio

Applied Machine Learning and High-Performance Computing on AWS

Machine Learning Engineering on AWS

Examining ML and data engineering roles