Mastering Java Machine Learning

Mastering Java Machine Learning

By : Uday Kamath, Krishna Choppella

Buy this Book

Mastering Java Machine Learning

By: Uday Kamath, Krishna Choppella

Buy this Book

Overview of this book

Java is one of the main languages used by practicing data scientists; much of the Hadoop ecosystem is Java-based, and it is certainly the language that most production systems in Data Science are written in. If you know Java, Mastering Machine Learning with Java is your next step on the path to becoming an advanced practitioner in Data Science. This book aims to introduce you to an array of advanced techniques in machine learning, including classification, clustering, anomaly detection, stream learning, active learning, semi-supervised learning, probabilistic graph modeling, text mining, deep learning, and big data batch and stream machine learning. Accompanying each chapter are illustrative examples and real-world case studies that show how to apply the newly learned techniques using sound methodologies and the best Java-based tools available today. On completing this book, you will have an understanding of the tools and techniques for building powerful machine learning models to solve data science problems in just about any domain.

Mastering Java Machine Learning

Credits

Foreword

About the Authors

About the Reviewers

www.PacktPub.com

Customer Feedback

Preface

Free Chapter

Machine Learning Review

Machine learning – history and definition

What is not machine learning?

Machine learning – concepts and terminology

Machine learning – types and subtypes

Datasets used in machine learning

Machine learning applications

Practical issues in machine learning

Machine learning – roles and process

Machine learning – tools and datasets

Summary

Practical Approach to Real-World Supervised Learning

Formal description and notation

Data transformation and preprocessing

Feature relevance analysis and dimensionality reduction

Model building

Model assessment, evaluation, and comparisons

Case Study – Horse Colic Classification

Summary

References

Unsupervised Machine Learning Techniques

Issues in common with supervised learning

Issues specific to unsupervised learning

Feature analysis and dimensionality reduction

Clustering

Outlier or anomaly detection

Real-world case study

Summary

References

Semi-Supervised and Active Learning

Semi-supervised learning

Active learning

Case study in active learning

Summary

References

Real-Time Stream Machine Learning

Assumptions and mathematical notations

Basic stream processing and computational techniques

Concept drift and drift detection

Incremental supervised learning

Incremental unsupervised learning using clustering

Unsupervised learning using outlier detection

Case study in stream learning

Summary

References

Probabilistic Graph Modeling

Probability revisited

Graph concepts

Bayesian networks

Markov networks and conditional random fields

Summary

Deep Learning

Multi-layer feed-forward neural network

Limitations of neural networks

Deep learning

Case study

Summary

References

Text Mining and Natural Language Processing

NLP, subfields, and tasks

Issues with mining unstructured data

Text processing components and transformations

Topics in text mining

Tools and usage

Summary

References

Big Data Machine Learning – The Final Frontier

What are the characteristics of Big Data?

Big Data Machine Learning

Batch Big Data Machine Learning

Case study

Linear Algebra

Vector

Matrix

Probability

Axioms of probability

Bayes' theorem

Index

Customer Reviews

5 star

4 star

3 star

2 star

1 star

Semi-supervised learning

The idea behind semi-supervised learning is to learn from labeled and unlabeled data to improve the predictive power of the models. The notion is explained with a simple illustration, Figure 1, which shows that when a large amount of unlabeled data is available, for example, HTML documents on the web, the expert can classify a few of them into known categories such as sports, news, entertainment, and so on. This small set of labeled data together with the large unlabeled dataset can then be used by semi-supervised learning techniques to learn models. Thus, using the knowledge of both labeled and unlabeled data, the model can classify unseen documents in the future. In contrast, supervised learning uses labeled data only:

Figure 1. Semi-Supervised Learning process (bottom) contrasted with Supervised Learning (top) using classification of web documents as an example. The main difference is the amount of labeled data available for learning, highlighted by the qualifier...

Mastering Java Machine Learning

By : Uday Kamath, Krishna Choppella

Mastering Java Machine Learning

By: Uday Kamath, Krishna Choppella

Overview of this book

Related Content you might be interested in

Current Title:

Mastering Java Machine Learning

Machine Learning in Java

Deep Learning with Hadoop

Mastering Machine Learning Algorithms.

Semi-supervised learning