Python Deep Learning - Third Edition

By : Ivan Vasilev

4 (1)

Buy this Book

Python Deep Learning - Third Edition

4 (1)

By: Ivan Vasilev

Buy this Book

Overview of this book

The field of deep learning has developed rapidly recently and today covers a broad range of applications. This makes it challenging to navigate and hard to understand without solid foundations. This book will guide you from the basics of neural networks to the state-of-the-art large language models in use today. The first part of the book introduces the main machine learning concepts and paradigms. It covers the mathematical foundations, the structure, and the training algorithms of neural networks and dives into the essence of deep learning. The second part of the book introduces convolutional networks for computer vision. We’ll learn how to solve image classification, object detection, instance segmentation, and image generation tasks. The third part focuses on the attention mechanism and transformers – the core network architecture of large language models. We’ll discuss new types of advanced tasks they can solve, such as chatbots and text-to-image generation. By the end of this book, you’ll have a thorough understanding of the inner workings of deep neural networks. You'll have the ability to develop new models and adapt existing ones to solve your tasks. You’ll also have sufficient understanding to continue your research and stay up to date with the latest advancements in the field.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Conventions used

Get in touch

Share Your Thoughts

Download a free PDF copy of this book

Part 1:Introduction to Neural Networks

Free Chapter

Chapter 1: Machine Learning – an Introduction

Technical requirements

Introduction to ML

Different ML approaches

Summary

Chapter 2: Neural Networks

Technical requirements

The need for NNs

The math of NNs

An introduction to NNs

Training NNs

Summary

Chapter 3: Deep Learning Fundamentals

Technical requirements

Introduction to DL

Fundamental DL concepts

Deep neural networks

Training deep neural networks

Applications of DL

Introducing popular DL libraries

Summary

Part 2: Deep Neural Networks for Computer Vision

Chapter 4: Computer Vision with Convolutional Networks

Technical requirements

Intuition and justification for CNNs

Convolutional layers

Pooling layers

The structure of a convolutional network

Classifying images with PyTorch and Keras

Advanced types of convolutions

Advanced CNN models

Summary

Chapter 5: Advanced Computer Vision Applications

Technical requirements

Transfer learning (TL)

Object detection

Introducing image segmentation

Image generation with diffusion models

Summary

Part 3: Natural Language Processing and Transformers

Chapter 6: Natural Language Processing and Recurrent Neural Networks

Technical requirements

Natural language processing

Introducing RNNs

Implementing text classification

Summary

Chapter 7: The Attention Mechanism and Transformers

Technical requirements

Introducing seq2seq models

Understanding the attention mechanism

Building transformers with attention

Summary

Chapter 8: Exploring Large Language Models in Depth

Technical requirements

Introducing LLMs

LLM architecture

Training LLMs

Emergent abilities of LLMs

Introducing Hugging Face Transformers

Summary

Chapter 9: Advanced Applications of Large Language Models

Technical requirements

Classifying images with Vision Transformer

Understanding the DEtection TRansformer

Generating images with stable diffusion

Exploring fine-tuning transformers

Harnessing the power of LLMs with LangChain

Summary

Part 4: Developing and Deploying Deep Neural Networks

Chapter 10: Machine Learning Operations (MLOps)

Technical requirements

Understanding model development

Exploring model deployment

Summary

Index

Why subscribe?

Other Books You May Enjoy

Packt is searching for authors like you

Share Your Thoughts

Download a free PDF copy of this book

Customer Reviews

4 (1)

5 star

4 star

100%

3 star

2 star

1 star

Introducing RNNs

An RNN is a type of NN that can process sequential data with variable length. Examples of such data include text sequences or the price of a stock at various moments in time. By using the word sequential, we imply that the sequence elements are related to each other and their order matters. For example, if we take a book and randomly shuffle all the words in it, the text will lose its meaning, even though we’ll still know the individual words.

RNNs get their name because they apply the same function over a sequence recurrently. We can define an RNN as a recurrence relation:

Here, f is a differentiable function, <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi mathvariant="bold">s</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:math> is a vector of values called internal RNN state (at step t), and <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi mathvariant="bold">x</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:math> is the network input at step t. Unlike regular NNs, where the state only depends on the current input (and RNN weights), here, <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi mathvariant="bold">s</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:math> is a function of both the current input, as well as the previous state, <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi mathvariant="bold">s</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>-</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:math> . You can think of <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi mathvariant="bold">s</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>-</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:math> as the RNN’s summary of all previous inputs. The...