Mastering Machine Learning Algorithms

Book Image

Mastering Machine Learning Algorithms

Book Image

Mastering Machine Learning Algorithms

Overview of this book

Machine learning is a subset of AI that aims to make modern-day computer systems smarter and more intelligent. The real power of machine learning resides in its algorithms, which make even the most difficult things capable of being handled by machines. However, with the advancement in the technology and requirements of data, machines will have to be smarter than they are today to meet the overwhelming data needs; mastering these algorithms and using them optimally is the need of the hour. Mastering Machine Learning Algorithms is your complete guide to quickly getting to grips with popular machine learning algorithms. You will be introduced to the most widely used algorithms in supervised, unsupervised, and semi-supervised machine learning, and will learn how to use them in the best possible manner. Ranging from Bayesian models to the MCMC algorithm to Hidden Markov models, this book will teach you how to extract features from your dataset and perform dimensionality reduction by making use of Python-based libraries such as scikit-learn v0.19.1. You will also learn how to use Keras and TensorFlow 1.x to train effective neural networks. If you are looking for a single resource to study, implement, and solve end-to-end machine learning problems and use-cases, this is the book you need.

Title Page

Dedication

Packt Upsell

Contributors

Preface

Free Chapter

Machine Learning Model Fundamentals

Machine Learning Model Fundamentals

Models and data

Features of a machine learning model

Loss and cost functions

Introduction to Semi-Supervised Learning

Introduction to Semi-Supervised Learning

Semi-supervised scenario

Generative Gaussian mixtures

Contrastive pessimistic likelihood estimation

Semi-supervised Support Vector Machines (S3VM)

Transductive Support Vector Machines (TSVM)

Graph-Based Semi-Supervised Learning

Graph-Based Semi-Supervised Learning

Label propagation

Label spreading

Label propagation based on Markov random walks

Manifold learning

Bayesian Networks and Hidden Markov Models

Bayesian Networks and Hidden Markov Models

Conditional probabilities and Bayes' theorem

Bayesian networks

Hidden Markov Models (HMMs)

EM Algorithm and Applications

EM Algorithm and Applications

MLE and MAP learning

Gaussian mixture

Factor analysis

Principal Component Analysis

Independent component analysis

Addendum to HMMs

Hebbian Learning and Self-Organizing Maps

Hebbian Learning and Self-Organizing Maps

Sanger's network

Rubner-Tavan's network

Self-organizing maps

Clustering Algorithms

Clustering Algorithms

k-Nearest Neighbors

Spectral clustering

Ensemble Learning

Ensemble Learning

Ensemble learning fundamentals

Gradient boosting

Ensembles of voting classifiers

Ensemble learning as model selection

Neural Networks for Machine Learning

Neural Networks for Machine Learning

The basic artificial neuron

Multilayer perceptrons

Optimization algorithms

Regularization and dropout

Batch normalization

Advanced Neural Models

Advanced Neural Models

Deep convolutional networks

Recurrent networks

Transfer learning

Autoencoders

Variational autoencoders

Generative Adversarial Networks

Generative Adversarial Networks

Adversarial training

Wasserstein GAN (WGAN)

Deep Belief Networks

Deep Belief Networks

Introduction to Reinforcement Learning

Introduction to Reinforcement Learning

Reinforcement Learning fundamentals

Policy iteration

Value iteration

TD(0) algorithm

Advanced Policy Estimation Algorithms

Advanced Policy Estimation Algorithms

TD(λ) algorithm

SARSA algorithm

Other Books You May Enjoy

Other Books You May Enjoy

Leave a review - let other readers know what you think

Index

Customer Reviews

5 star

0

4 star

0

3 star

0

2 star

0

1 star

0

AdaBoost

In the previous section, we have seen that sampling with a replacement leads to datasets where the samples are randomly reweighted. However, if M is very large, most of the samples will appear only once and, moreover, all the choices are totally random. AdaBoost is an algorithm proposed by Schapire and Freund that tries to maximize the efficiency of each weak learner by employing adaptive boosting (the name derives from this). In particular, the ensemble is grown sequentially and the data distribution is recomputed at each step so as to increase the weight of those samples that were misclassified and reduce the weight of the ones that were correctly classified. In this way, every new learner is forced to focus on those regions that were more problematic for the previous estimators. The reader can immediately understand that, contrary to random forests and other bagging methods, boosting doesn't rely on randomness to reduce the variance and improve the accuracy. Rather, it works...