Production-Ready Applied Deep Learning

By : Tomasz Palczewski, Jaejun (Brandon) Lee, Lenin Mookiah

Production-Ready Applied Deep Learning

By: Tomasz Palczewski, Jaejun (Brandon) Lee, Lenin Mookiah

Overview of this book

Machine learning engineers, deep learning specialists, and data engineers encounter various problems when moving deep learning models to a production environment. The main objective of this book is to close the gap between theory and applications by providing a thorough explanation of how to transform various models for deployment and efficiently distribute them with a full understanding of the alternatives. First, you will learn how to construct complex deep learning models in PyTorch and TensorFlow. Next, you will acquire the knowledge you need to transform your models from one framework to the other and learn how to tailor them for specific requirements that deployment environments introduce. The book also provides concrete implementations and associated methodologies that will help you apply the knowledge you gain right away. You will get hands-on experience with commonly used deep learning frameworks and popular cloud services designed for data analytics at scale. Additionally, you will get to grips with the authors’ collective knowledge of deploying hundreds of AI-based services at a large scale. By the end of this book, you will have understood how to convert a model developed for proof of concept into a production-ready application optimized for a particular production setting.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Download the color images

Conventions used

Get in touch

Share Your Thoughts

Part 1 – Building a Minimum Viable Product

Free Chapter

Chapter 1: Effective Planning of Deep Learning-Driven Projects

Technical requirements

What is DL?

Understanding the role of DL in our daily lives

Overview of DL projects

Planning a DL project

Summary

Further reading

Chapter 2: Data Preparation for Deep Learning Projects

Technical requirements

Setting up notebook environments

Data collection, data cleaning, and data preprocessing

Extracting features from data

Performing data visualization

Introduction to Docker

Summary

Chapter 3: Developing a Powerful Deep Learning Model

Technical requirements

Going through the basic theory of DL

Components of DL frameworks

Implementing and training a model in PyTorch

Implementing and training a model in TF

Decomposing a complex, state-of-the-art model implementation

Summary

Chapter 4: Experiment Tracking, Model Management, and Dataset Versioning

Technical requirements

Overview of DL project tracking

DL project tracking with Weights & Biases

DL project tracking with MLflow and DVC

Dataset versioning – beyond Weights & Biases, MLflow, and DVC

Summary

Part 2 – Building a Fully Featured Product

Chapter 5: Data Preparation in the Cloud

Technical requirements

Data processing in the cloud

Introduction to Apache Spark

Setting up a single-node EC2 instance for ETL

Setting up an EMR cluster for ETL

Creating a Glue job for ETL

Utilizing SageMaker for ETL

Comparing the ETL solutions in AWS

Summary

Chapter 6: Efficient Model Training

Technical requirements

Training a model on a single machine

Training a model on a cluster

Training a model using SageMaker

Training a model using Horovod

Training a model using Ray

Training a model using Kubeflow

Summary

Chapter 7: Revealing the Secret of Deep Learning Models

Technical requirements

Obtaining the best performing model using hyperparameter tuning

Understanding the behavior of the model with Explainable AI

Summary

Part 3 – Deployment and Maintenance

Chapter 8: Simplifying Deep Learning Model Deployment

Technical requirements

Introduction to ONNX

Conversion between TensorFlow and ONNX

Conversion between PyTorch and ONNX

Summary

Chapter 9: Scaling a Deep Learning Pipeline

Technical requirements

Inferencing using Elastic Kubernetes Service

Inferencing using SageMaker

Summary

Chapter 10: Improving Inference Efficiency

Technical requirements

Network quantization – reducing the number of bits used for model parameters

Weight sharing – reducing the number of distinct weight values

Network pruning – eliminating unnecessary connections within the network

Knowledge distillation – obtaining a smaller network by mimicking the prediction

Network Architecture Search – finding the most efficient network architecture

Summary

Chapter 11: Deep Learning on Mobile Devices

Preparing DL models for mobile devices

Creating iOS apps with a DL model

Creating Android apps with a DL model

Summary

Chapter 12: Monitoring Deep Learning Endpoints in Production

Technical requirements

Introduction to DL endpoint monitoring in production

Monitoring using CloudWatch

Monitoring a SageMaker endpoint using CloudWatch

Monitoring an EKS endpoint using CloudWatch

Summary

Chapter 13: Reviewing the Completed Deep Learning Project

Reviewing a DL project

Gathering the reusable knowledge, concepts, and artifacts for future projects

Summary

Index

Why subscribe?

Other Books You May Enjoy

Packt is searching for authors like you

Share Your Thoughts

Customer Reviews

5 star

4 star

3 star

2 star

1 star

Overview of DL projects

While DL and other software engineering projects have a lot in common, DL projects emphasize planning, due to the extensive need for resources, mainly coming from the complexity of the models and the high volume of data involved. In general, DL projects can be split into the following phases:

Project planning
Building MVPs
Building FFPs
Deployment and maintenance
Project evaluation

In this section, we provide high-level overviews of these phases. The following sections cover each phase in detail.

Project planning

As the first step, the project lead must clearly define what needs to be achieved by the project and understand groups that can affect or be affected by the project. The evaluation metrics need to be defined and agreed upon, as they will be revisited during project evaluation. Then, the team members group together to discuss individual responsibilities and achieve business objectives using available resources. This process naturally leads to a timeline, an estimate of how long the project would take. Overall, project planning should result in the generation of a document called a playbook, which includes a thorough description of how the project will be carried out and evaluated.

Building minimum viable products

Once the direction is clear for everyone, the next step is to build an MVP, a simplistic version of the target deliverable that showcases the project’s value. Another important aspect of this phase is to understand the project’s difficulties and reject paths with greater risks or less promising outcomes by following the fail fast, fail often philosophy. Therefore, data scientists and engineers typically work with partially sampled datasets in development settings and ignore insignificant optimizations.

Building fully featured products

Once the feasibility of the project has been confirmed by the MVP, it must be packaged into an FFP. This phase aims to polish up the MVP to build a production-ready deliverable with various optimizations. In the case of DL projects, additional data preparation techniques are introduced to improve the quality of input data, or the model pipeline gets augmented slightly for greater model performance. Additionally, the data preparation pipeline and model training pipeline may be migrated to the cloud, exploiting various web services for higher throughput and quality. In this case, the whole pipeline often gets reimplemented using different tools and services. This book focuses on Amazon Web Services (AWS), the most popular web service for handling high volumes of data and expensive computations.

Deployment and maintenance

In many cases, the deployment settings are different from the development settings. Therefore, different sets of tools are often involved when moving an FFP to production. Furthermore, deployment may introduce problems that weren’t visible during development, which mainly arise as a result of limited computational resources. Consequently, many engineers and scientists spend additional time improving the user experience during this phase. Most people believe that deployment is the last step. However, there is one more step: maintenance. The quality of data and model performance needs to be monitored consistently to provide stable services to targeted users.

Project evaluation

In the last phase, project evaluation, the team should revisit the discussions made during project planning to evaluate whether the project has been carried out successfully or not. Furthermore, the details of the project need to be recorded, and potential improvements must be discussed so that the next projects can be achieved more efficiently.

Things to remember

a. The phases within DL projects are split into project planning, building MVPs, building FFPs, deployment and maintenance, and project evaluation.

b. During the project planning phase, the project goal and evaluation metrics are defined, and the team discusses an individual's responsibility, available resources, and the timeline for the project.

c. The purpose of building an MVP is to understand the difficulties of the project and reject paths that pose greater risks or offer less promising outcomes.

d. The FFP is a production-ready deliverable that is an optimized version of the MVP. The data preparation pipeline and model training pipeline may be migrated to the cloud, exploiting various web services for higher throughput and quality.

e. Deployment settings often provide limited computational resources. In this case, the system needs to be tuned to provide stable services to target users.

f. Upon the completion of the project, the team needs to revisit the timeline, assigned responsibilities, and business requirements to evaluate the success of the project.

In the following section, we will walk you through how to plan a DL project properly.

Production-Ready Applied Deep Learning

By : Tomasz Palczewski, Jaejun (Brandon) Lee, Lenin Mookiah

Production-Ready Applied Deep Learning

By: Tomasz Palczewski, Jaejun (Brandon) Lee, Lenin Mookiah

Overview of this book

Related Content you might be interested in

Current Title:

Production-Ready Applied Deep Learning

Overview of DL projects

Project planning

Building minimum viable products

Building fully featured products

Deployment and maintenance

Project evaluation