Deep Learning with Theano

Book Image

Deep Learning with Theano

By : Christopher Bourez

Book Image

Deep Learning with Theano

By: Christopher Bourez

Overview of this book

This book offers a complete overview of Deep Learning with Theano, a Python-based library that makes optimizing numerical expressions and deep learning models easy on CPU or GPU. The book provides some practical code examples that help the beginner understand how easy it is to build complex neural networks, while more experimented data scientists will appreciate the reach of the book, addressing supervised and unsupervised learning, generative models, reinforcement learning in the fields of image recognition, natural language processing, or game strategy. The book also discusses image recognition tasks that range from simple digit recognition, image classification, object localization, image segmentation, to image captioning. Natural language processing examples include text generation, chatbots, machine translation, and question answering. The last example deals with generating random data that looks real and solving games such as in the Open-AI gym. At the end, this book sums up the best -performing nets for each task. While early research results were based on deep stacks of neural layers, in particular, convolutional layers, the book presents the principles that improved the efficiency of these architectures, in order to help the reader build new custom nets.

Deep Learning with Theano

Deep Learning with Theano

Credits

About the Author

About the Author

Acknowledgments

Acknowledgments

About the Reviewers

About the Reviewers

www.PacktPub.com

www.PacktPub.com

Customer Feedback

Customer Feedback

Preface

Free Chapter

Theano Basics

The need for tensors

Installing and loading Theano

Graphs and symbolic computing

Operations on tensors

Memory and variables

Functions and automatic differentiation

Loops in symbolic computing

Configuration, profiling and debugging

Classifying Handwritten Digits with a Feedforward Network

Classifying Handwritten Digits with a Feedforward Network

The MNIST dataset

Structure of a training program

Classification loss function

Single-layer linear model

Cost function and errors

Backpropagation and stochastic gradient descent

Multiple layer model

Convolutions and max layers

Optimization and other update rules

Related articles

Encoding Word into Vector

Encoding Word into Vector

Encoding and embedding

Continuous Bag of Words model

Training the model

Visualizing the learned embeddings

Evaluating embeddings – analogical reasoning

Evaluating embeddings – quantitative analysis

Application of word embeddings

Further reading

Generating Text with a Recurrent Neural Net

Generating Text with a Recurrent Neural Net

A dataset for natural language

Simple recurrent network

Metrics for natural language performance

Training loss comparison

Example of predictions

Applications of RNN

Related articles

Analyzing Sentiment with a Bidirectional LSTM

Analyzing Sentiment with a Bidirectional LSTM

Installing and configuring Keras

Preprocessing text data

Designing the architecture for the model

Compiling and training the model

Evaluating the model

Saving and loading the model

Running the example

Further reading

Locating with Spatial Transformer Networks

Locating with Spatial Transformer Networks

MNIST CNN model with Lasagne

A localization network

Unsupervised learning with co-localization

Region-based localization networks

Further reading

Classifying Images with Residual Networks

Classifying Images with Residual Networks

Natural image datasets

Residual connections

Stochastic depth

Dense connections

Data augmentation

Further reading

Translating and Explaining with Encoding – decoding Networks

Translating and Explaining with Encoding – decoding Networks

Sequence-to-sequence networks for natural language processing

Seq2seq for translation

Seq2seq for chatbots

Improving efficiency of sequence-to-sequence network

Deconvolutions for images

Multimodal deep learning

Further reading

Selecting Relevant Inputs or Memories with the Mechanism of Attention

Selecting Relevant Inputs or Memories with the Mechanism of Attention

Differentiable mechanism of attention

Store and retrieve information in Neural Turing Machines

Memory networks

Further reading

Predicting Times Sequences with Advanced RNN

Predicting Times Sequences with Advanced RNN

Dropout for RNN

Deep approaches for RNN

Stacked recurrent networks

Deep transition recurrent network

Highway networks design principle

Recurrent Highway Networks

Further reading

Learning from the Environment with Reinforcement

Learning from the Environment with Reinforcement

Reinforcement learning tasks

Simulation environments

Training stability

Policy gradients with REINFORCE algorithms

Related articles

Learning Features with Unsupervised Generative Networks

Learning Features with Unsupervised Generative Networks

Generative models

Semi-supervised learning

Further reading

Extending Deep Learning with Theano

Extending Deep Learning with Theano

Theano Op in Python for CPU

Theano Op in Python for the GPU

Theano Op in C for CPU

Theano Op in C for GPU

Coalesced transpose via shared memory, NVIDIA parallel for all

The future of artificial intelligence

Further reading

Index

Customer Reviews

5 star

0

4 star

0

3 star

0

2 star

0

1 star

0

Related articles

You can refer to the following articles:

Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning, Ronald J. Williams, 1992
Policy Gradient Methods for Reinforcement Learning with Function Approximation, Richard S. Sutton, David McAllester, Satinder Singh, Yishay Mansour, 1999
Playing Atari with Deep Reinforcement Learning, Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, Martin Riedmiller, 2013
Mastering the Game of Go with Deep Neural Networks and Tree Search, David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel & Demis Hassabis, 2016
Asynchronous Methods for Deep Reinforcement Learning, Volodymyr Mnih, Adrià Puigdomènech...