Machine Learning with Spark

Apache Spark is a framework for distributed computing that is designed from the ground up to be optimized for low latency tasks and in-memory data storage. It is one of the few frameworks for parallel computing that combines speed, scalability, in-memory processing, and fault tolerance with ease of programming and a flexible, expressive, and powerful API design.
This book guides you through the basics of Spark's API used to load and process data and prepare the data to use as input to the various machine learning models. There are detailed examples and real-world use cases for you to explore common machine learning models including recommender systems, classification, regression, clustering, and dimensionality reduction. You will cover advanced topics such as working with large-scale text data, and methods for online machine learning and model evaluation using Spark Streaming.

439 бумажных страниц

Год выхода издания: 2015
Издательство: Packt Publishing

На полках

Oleg Danilchenko
Machine learning
- 15
- 4
Отписаться
Cheshire Rabbit
Data Engineering
- 19
- 3
Отписаться

Machine Learning with Spark

Похожие книгиВсе

На полках