Python Deep Learning - Third Edition

By : Ivan Vasilev

4 (1)

Buy this Book

Python Deep Learning - Third Edition

4 (1)

By: Ivan Vasilev

Buy this Book

Overview of this book

The field of deep learning has developed rapidly recently and today covers a broad range of applications. This makes it challenging to navigate and hard to understand without solid foundations. This book will guide you from the basics of neural networks to the state-of-the-art large language models in use today. The first part of the book introduces the main machine learning concepts and paradigms. It covers the mathematical foundations, the structure, and the training algorithms of neural networks and dives into the essence of deep learning. The second part of the book introduces convolutional networks for computer vision. We’ll learn how to solve image classification, object detection, instance segmentation, and image generation tasks. The third part focuses on the attention mechanism and transformers – the core network architecture of large language models. We’ll discuss new types of advanced tasks they can solve, such as chatbots and text-to-image generation. By the end of this book, you’ll have a thorough understanding of the inner workings of deep neural networks. You'll have the ability to develop new models and adapt existing ones to solve your tasks. You’ll also have sufficient understanding to continue your research and stay up to date with the latest advancements in the field.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Conventions used

Get in touch

Share Your Thoughts

Download a free PDF copy of this book

Part 1:Introduction to Neural Networks

Free Chapter

Chapter 1: Machine Learning – an Introduction

Technical requirements

Introduction to ML

Different ML approaches

Summary

Chapter 2: Neural Networks

Technical requirements

The need for NNs

The math of NNs

An introduction to NNs

Training NNs

Summary

Chapter 3: Deep Learning Fundamentals

Technical requirements

Introduction to DL

Fundamental DL concepts

Deep neural networks

Training deep neural networks

Applications of DL

Introducing popular DL libraries

Summary

Part 2: Deep Neural Networks for Computer Vision

Chapter 4: Computer Vision with Convolutional Networks

Technical requirements

Intuition and justification for CNNs

Convolutional layers

Pooling layers

The structure of a convolutional network

Classifying images with PyTorch and Keras

Advanced types of convolutions

Advanced CNN models

Summary

Chapter 5: Advanced Computer Vision Applications

Technical requirements

Transfer learning (TL)

Object detection

Introducing image segmentation

Image generation with diffusion models

Summary

Part 3: Natural Language Processing and Transformers

Chapter 6: Natural Language Processing and Recurrent Neural Networks

Technical requirements

Natural language processing

Introducing RNNs

Implementing text classification

Summary

Chapter 7: The Attention Mechanism and Transformers

Technical requirements

Introducing seq2seq models

Understanding the attention mechanism

Building transformers with attention

Summary

Chapter 8: Exploring Large Language Models in Depth

Technical requirements

Introducing LLMs

LLM architecture

Training LLMs

Emergent abilities of LLMs

Introducing Hugging Face Transformers

Summary

Chapter 9: Advanced Applications of Large Language Models

Technical requirements

Classifying images with Vision Transformer

Understanding the DEtection TRansformer

Generating images with stable diffusion

Exploring fine-tuning transformers

Harnessing the power of LLMs with LangChain

Summary

Part 4: Developing and Deploying Deep Neural Networks

Chapter 10: Machine Learning Operations (MLOps)

Technical requirements

Understanding model development

Exploring model deployment

Summary

Index

Why subscribe?

Other Books You May Enjoy

Packt is searching for authors like you

Share Your Thoughts

Download a free PDF copy of this book

Customer Reviews

4 (1)

5 star

4 star

100%

3 star

2 star

1 star

Training NNs

The NN function <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>θ</mml:mi></mml:mrow></mml:msub><mml:mfenced separators="|"><mml:mrow><mml:mi mathvariant="bold">x</mml:mi></mml:mrow></mml:mfenced></mml:math> approximates the function <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:mi>g</mml:mi><mml:mfenced separators="|"><mml:mrow><mml:mi mathvariant="bold">x</mml:mi></mml:mrow></mml:mfenced></mml:math> : . The goal of the training is to find parameters, θ, such that <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>θ</mml:mi></mml:mrow></mml:msub><mml:mfenced separators="|"><mml:mrow><mml:mi mathvariant="bold">x</mml:mi></mml:mrow></mml:mfenced></mml:math> will best approximate <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:mi>g</mml:mi><mml:mfenced separators="|"><mml:mrow><mml:mi mathvariant="bold">x</mml:mi></mml:mrow></mml:mfenced></mml:math> . First, we’ll see how to do that for a
single-layer network, using an optimization algorithm called GD. Then, we’ll extend it to a deep feedforward network with the help of BP.

Note

We should note that an NN and its training algorithm are two separate things. This means we can adjust the weights of a network in some way other than GD and BP, but this is the most popular and efficient way to do so and is, ostensibly, the only way that is currently used in practice.

GD

For the purposes of this section, we’ll train a simple NN using the mean square error (MSE) cost function. It measures the difference (known as error) between the network output and the training data labels of all training samples:

$<mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math" display="block"><mml:mi>J</mml:mi><mml:mfenced separators="|"><mml:mrow><mml:mi>θ</mml:mi></mml:mrow></mml:mfenced><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn><mml:mi>n</mml:mi></mml:mrow></mml:mfrac><mml:mrow><mml:munderover><mml:mo stretchy="false">∑</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover><mml:mrow><mml:msup><mml:mrow><mml:mfenced separators="|"><mml:mrow><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>θ</mml:mi></mml:mrow></mml:msub><mml:mfenced separators="|"><mml:mrow><mml:msup><mml:mrow><mml:mi mathvariant="bold">x</mml:mi></mml:mrow><mml:mrow><mml:mfenced separators="|"><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:msup></mml:mrow></mml:mfenced><mml:mo>-</mml:mo><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mfenced separators="|"><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:mfenced></mml:mrow></mml:msup></mml:mrow></mml:mfenced></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mrow></mml:math>$

At first, this might look scary, but fear not! Behind the scenes, it’s very simple and straightforward...

Python Deep Learning - Third Edition

By : Ivan Vasilev

Python Deep Learning - Third Edition

By: Ivan Vasilev

Overview of this book

Related Content you might be interested in

Current Title:

Python Deep Learning - Third Edition

Mastering Transformers

Hands-On Convolutional Neural Networks with TensorFlow

Deep Learning with TensorFlow and Keras

Training NNs

GD