Python Deep Learning - Third Edition

By : Ivan Vasilev

4 (1)

Buy this Book

Python Deep Learning - Third Edition

4 (1)

By: Ivan Vasilev

Buy this Book

Overview of this book

The field of deep learning has developed rapidly recently and today covers a broad range of applications. This makes it challenging to navigate and hard to understand without solid foundations. This book will guide you from the basics of neural networks to the state-of-the-art large language models in use today. The first part of the book introduces the main machine learning concepts and paradigms. It covers the mathematical foundations, the structure, and the training algorithms of neural networks and dives into the essence of deep learning. The second part of the book introduces convolutional networks for computer vision. We’ll learn how to solve image classification, object detection, instance segmentation, and image generation tasks. The third part focuses on the attention mechanism and transformers – the core network architecture of large language models. We’ll discuss new types of advanced tasks they can solve, such as chatbots and text-to-image generation. By the end of this book, you’ll have a thorough understanding of the inner workings of deep neural networks. You'll have the ability to develop new models and adapt existing ones to solve your tasks. You’ll also have sufficient understanding to continue your research and stay up to date with the latest advancements in the field.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Conventions used

Get in touch

Share Your Thoughts

Download a free PDF copy of this book

Part 1:Introduction to Neural Networks

Free Chapter

Chapter 1: Machine Learning – an Introduction

Technical requirements

Introduction to ML

Different ML approaches

Summary

Chapter 2: Neural Networks

Technical requirements

The need for NNs

The math of NNs

An introduction to NNs

Training NNs

Summary

Chapter 3: Deep Learning Fundamentals

Technical requirements

Introduction to DL

Fundamental DL concepts

Deep neural networks

Training deep neural networks

Applications of DL

Introducing popular DL libraries

Summary

Part 2: Deep Neural Networks for Computer Vision

Chapter 4: Computer Vision with Convolutional Networks

Technical requirements

Intuition and justification for CNNs

Convolutional layers

Pooling layers

The structure of a convolutional network

Classifying images with PyTorch and Keras

Advanced types of convolutions

Advanced CNN models

Summary

Chapter 5: Advanced Computer Vision Applications

Technical requirements

Transfer learning (TL)

Object detection

Introducing image segmentation

Image generation with diffusion models

Summary

Part 3: Natural Language Processing and Transformers

Chapter 6: Natural Language Processing and Recurrent Neural Networks

Technical requirements

Natural language processing

Introducing RNNs

Implementing text classification

Summary

Chapter 7: The Attention Mechanism and Transformers

Technical requirements

Introducing seq2seq models

Understanding the attention mechanism

Building transformers with attention

Summary

Chapter 8: Exploring Large Language Models in Depth

Technical requirements

Introducing LLMs

LLM architecture

Training LLMs

Emergent abilities of LLMs

Introducing Hugging Face Transformers

Summary

Chapter 9: Advanced Applications of Large Language Models

Technical requirements

Classifying images with Vision Transformer

Understanding the DEtection TRansformer

Generating images with stable diffusion

Exploring fine-tuning transformers

Harnessing the power of LLMs with LangChain

Summary

Part 4: Developing and Deploying Deep Neural Networks

Chapter 10: Machine Learning Operations (MLOps)

Technical requirements

Understanding model development

Exploring model deployment

Summary

Index

Why subscribe?

Other Books You May Enjoy

Packt is searching for authors like you

Share Your Thoughts

Download a free PDF copy of this book

Customer Reviews

4 (1)

5 star

4 star

100%

3 star

2 star

1 star

LLM architecture

In Chapter 7, we introduced the multi-head attention (MHA) mechanism and the three major transformer variants—encoder-decoder, encoder-only, and decoder-only (we used BERT and GPT as prototypical encoder and decoder models). In this section, we’ll discuss various bits and pieces of the LLM architecture. Let’s start by focusing our attention (yes—it’s the same old joke) on the attention mechanism.

LLM attention variants

The attention we discussed so far is known as global attention. The following diagram displays the connectivity matrix of a bidirectional global self-attention mechanism (context window with size n=8):

Figure 8.1 – Global self-attention with a context window with size n=8

Each row and column represent the full input token sequence, . The dotted colored diagonal cells represent the current input token (query), <mml:math xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:m="http://schemas.openxmlformats.org/officeDocument/2006/math"><mml:msub><mml:mrow><mml:mi mathvariant="bold">t</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:math> . The uninterrupted colored cells of each column represent all tokens...

Python Deep Learning - Third Edition

By : Ivan Vasilev

Python Deep Learning - Third Edition

By: Ivan Vasilev

Overview of this book

Related Content you might be interested in

Current Title:

Python Deep Learning - Third Edition

Mastering Transformers

Hands-On Convolutional Neural Networks with TensorFlow

Deep Learning with TensorFlow and Keras

LLM architecture

LLM attention variants