Python Deep Learning - Third Edition

By : Ivan Vasilev

4 (1)

Buy this Book

Python Deep Learning - Third Edition

4 (1)

By: Ivan Vasilev

Buy this Book

Overview of this book

The field of deep learning has developed rapidly recently and today covers a broad range of applications. This makes it challenging to navigate and hard to understand without solid foundations. This book will guide you from the basics of neural networks to the state-of-the-art large language models in use today. The first part of the book introduces the main machine learning concepts and paradigms. It covers the mathematical foundations, the structure, and the training algorithms of neural networks and dives into the essence of deep learning. The second part of the book introduces convolutional networks for computer vision. We’ll learn how to solve image classification, object detection, instance segmentation, and image generation tasks. The third part focuses on the attention mechanism and transformers – the core network architecture of large language models. We’ll discuss new types of advanced tasks they can solve, such as chatbots and text-to-image generation. By the end of this book, you’ll have a thorough understanding of the inner workings of deep neural networks. You'll have the ability to develop new models and adapt existing ones to solve your tasks. You’ll also have sufficient understanding to continue your research and stay up to date with the latest advancements in the field.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Conventions used

Get in touch

Share Your Thoughts

Download a free PDF copy of this book

Part 1:Introduction to Neural Networks

Free Chapter

Chapter 1: Machine Learning – an Introduction

Technical requirements

Introduction to ML

Different ML approaches

Summary

Chapter 2: Neural Networks

Technical requirements

The need for NNs

The math of NNs

An introduction to NNs

Training NNs

Summary

Chapter 3: Deep Learning Fundamentals

Technical requirements

Introduction to DL

Fundamental DL concepts

Deep neural networks

Training deep neural networks

Applications of DL

Introducing popular DL libraries

Summary

Part 2: Deep Neural Networks for Computer Vision

Chapter 4: Computer Vision with Convolutional Networks

Technical requirements

Intuition and justification for CNNs

Convolutional layers

Pooling layers

The structure of a convolutional network

Classifying images with PyTorch and Keras

Advanced types of convolutions

Advanced CNN models

Summary

Chapter 5: Advanced Computer Vision Applications

Technical requirements

Transfer learning (TL)

Object detection

Introducing image segmentation

Image generation with diffusion models

Summary

Part 3: Natural Language Processing and Transformers

Chapter 6: Natural Language Processing and Recurrent Neural Networks

Technical requirements

Natural language processing

Introducing RNNs

Implementing text classification

Summary

Chapter 7: The Attention Mechanism and Transformers

Technical requirements

Introducing seq2seq models

Understanding the attention mechanism

Building transformers with attention

Summary

Chapter 8: Exploring Large Language Models in Depth

Technical requirements

Introducing LLMs

LLM architecture

Training LLMs

Emergent abilities of LLMs

Introducing Hugging Face Transformers

Summary

Chapter 9: Advanced Applications of Large Language Models

Technical requirements

Classifying images with Vision Transformer

Understanding the DEtection TRansformer

Generating images with stable diffusion

Exploring fine-tuning transformers

Harnessing the power of LLMs with LangChain

Summary

Part 4: Developing and Deploying Deep Neural Networks

Chapter 10: Machine Learning Operations (MLOps)

Technical requirements

Understanding model development

Exploring model deployment

Summary

Index

Why subscribe?

Other Books You May Enjoy

Packt is searching for authors like you

Share Your Thoughts

Download a free PDF copy of this book

Customer Reviews

4 (1)

5 star

4 star

100%

3 star

2 star

1 star

Summary

LLMs are very large transformers with various modifications to accommodate the large size. In this chapter, we discussed these modifications, as well as the qualitative differences between LLMs and regular transformers. First, we focused on their architecture, including more efficient attention mechanisms such as sparse attention and prefix decoders. We also discussed the nuts and bolts of the LLM architecture. Next, we surveyed the latest LLM architectures with special attention given to the GPT and LlaMa series of models. Then, we discussed LLM training, including training datasets, the Adam optimization algorithm, and various performance improvements. We also discussed the RLHF technique and the emergent abilities of LLMs. Finally, we introduced the Hugging Face Transformers library.

In the next chapter, we’ll discuss transformers for computer vision (CV), multimodal transformers, and we’ll continue our introduction to the Transformers library.

Python Deep Learning - Third Edition

By : Ivan Vasilev

Python Deep Learning - Third Edition

By: Ivan Vasilev

Overview of this book

Related Content you might be interested in

Current Title:

Python Deep Learning - Third Edition

Mastering Transformers

Hands-On Convolutional Neural Networks with TensorFlow

Deep Learning with TensorFlow and Keras

Summary