AWS Certified Machine Learning Specialty: MLS-C01 Certification Guide

By : Somanath Nanda, Weslley Moura

AWS Certified Machine Learning Specialty: MLS-C01 Certification Guide

By: Somanath Nanda, Weslley Moura

Overview of this book

The AWS Certified Machine Learning Specialty exam tests your competency to perform machine learning (ML) on AWS infrastructure. This book covers the entire exam syllabus using practical examples to help you with your real-world machine learning projects on AWS. Starting with an introduction to machine learning on AWS, you'll learn the fundamentals of machine learning and explore important AWS services for artificial intelligence (AI). You'll then see how to prepare data for machine learning and discover a wide variety of techniques for data manipulation and transformation for different types of variables. The book also shows you how to handle missing data and outliers and takes you through various machine learning tasks such as classification, regression, clustering, forecasting, anomaly detection, text mining, and image processing, along with the specific ML algorithms you need to know to pass the exam. Finally, you'll explore model evaluation, optimization, and deployment and get to grips with deploying models in a production environment and monitoring them. By the end of this book, you'll have gained knowledge of the key challenges in machine learning and the solutions that AWS has released for each of them, along with the tools, methods, and techniques commonly used in each domain of AWS ML.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Download the color images

Conventions used

Get in touch

Reviews

Section 1: Introduction to Machine Learning

Free Chapter

Chapter 1: Machine Learning Fundamentals

Comparing AI, ML, and DL

Classifying supervised, unsupervised, and reinforcement learning

The CRISP-DM modeling life cycle

Data splitting

Modeling expectations

Introducing ML frameworks

ML in the cloud

Summary

Questions

Chapter 2: AWS Application Services for AI/ML

Technical requirements

Analyzing images and videos with Amazon Rekognition

Text to speech with Amazon Polly

Speech to text with Amazon Transcribe

Implementing natural language processing with Amazon Comprehend

Translating documents with Amazon Translate

Extracting text from documents with Amazon Textract

Creating chatbots on Amazon Lex

Summary

Section 2: Data Engineering and Exploratory Data Analysis

Chapter 3: Data Preparation and Transformation

Identifying types of features

Dealing with categorical features

Dealing with numerical features

Understanding data distributions

Handling missing values

Dealing with outliers

Dealing with unbalanced datasets

Dealing with text data

Summary

Questions

Chapter 4: Understanding and Visualizing Data

Visualizing relationships in your data

Visualizing comparisons in your data

Visualizing distributions in your data

Visualizing compositions in your data

Building key performance indicators

Introducing Quick Sight

Summary

Questions

Chapter 5: AWS Services for Data Storing

Technical requirements

Storing data on Amazon S3

Controlling access to buckets and objects on Amazon S3

Protecting data on Amazon S3

Securing S3 objects at rest and in transit

Using other types of data stores

Relational Database Services (RDSes)

Managing failover in Amazon RDS

Taking automatic backup, RDS snapshots, and restore and read replicas

Writing to Amazon Aurora with multi-master capabilities

Storing columnar data on Amazon Redshift

Amazon DynamoDB for NoSQL database as a service

Summary

Chapter 6: AWS Services for Data Processing

Technical requirements

Creating ETL jobs on AWS Glue

Querying S3 data using Athena

Processing real-time data using Kinesis data streams

Storing and transforming real-time data using Kinesis Data Firehose

Different ways of ingesting data from on-premises into AWS

Processing stored data on AWS

Summary

Section 3: Data Modeling

Chapter 7: Applying Machine Learning Algorithms

Introducing this chapter

Storing the training data

A word about ensemble models

Supervised learning

Unsupervised learning

Textual analysis

Image processing

Summary

Questions

Chapter 8: Evaluating and Optimizing Models

Introducing model evaluation

Evaluating classification models

Evaluating regression models

Model optimization

Summary

Questions

Chapter 9: Amazon SageMaker Modeling

Technical requirements

Creating notebooks in Amazon SageMaker

Model tuning

Choosing instance types in Amazon SageMaker

Securing SageMaker notebooks

Creating alternative pipelines with Lambda Functions

Working with Step Functions

Summary

Why subscribe?

Other Books You May Enjoy

Packt is searching for authors like you

Leave a review - let other readers know what you think

Customer Reviews

5 star

4 star

3 star

2 star

1 star

Dealing with unbalanced datasets

At this point, I hope you have realized why data preparation is probably the longest part of our work. We have learned about data transformation, missing data values, and outliers, but the list of problems goes on. Don't worry – bear with me and let's master this topic together!

Another well-known problem with ML models, specifically with binary classification problems, is unbalanced classes. In a binary classification model, we say that a dataset is unbalanced when most of its observations belong to the same class (target variable).

This is very common in fraud identification systems, for example, where most of the events belong to a regular operation, while a very small number of events belong to a fraudulent operation. In this case, we can also say that fraud is a rare event.

There is no strong rule for defining whether a dataset is unbalanced or not, in the sense of it being necessary to worry about it. Most challenge problems...

AWS Certified Machine Learning Specialty: MLS-C01 Certification Guide

By : Somanath Nanda, Weslley Moura

AWS Certified Machine Learning Specialty: MLS-C01 Certification Guide

By: Somanath Nanda, Weslley Moura

Overview of this book

Related Content you might be interested in

Current Title:

AWS Certified Machine Learning Specialty: MLS-C01 Certification Guide

AWS Certified Cloud Practitioner Exam Guide

Hands-On Artificial Intelligence on Amazon Web Services

Data Wrangling on AWS

Dealing with unbalanced datasets