Loading
Cover of Spark Cookbook

Book guide and evaluation

Spark Cookbook

Rishi Yadav

English Unordered Data Science
4.8 / 5

0 reviews

2015

Published

226

pages

199

views

Over 60 recipes on Spark, covering Spark Core, Spark SQL, Spark Streaming, MLlib, and GraphX librariesAbout This Book Become an expert at graph processing using GraphX Use Apache Spark as your single big data compute platform and master its libraries Learn with recipes tha

Before you read

What will you get from this book?

Over 60 recipes on Spark, covering Spark Core, Spark SQL, Spark Streaming, MLlib, and GraphX librariesAbout This Book Become an expert at graph processing using GraphX Use Apache Spark as your single big data compute platform and master its libraries Learn with recipes that can be run on a single machine as well as on a production cluster of thousands of machines Who This Book Is ForIf you are a data engineer, an application developer, or a data scientist who would like to leverage the power of Apache Spark to get better insights from big data, then this is the book for you.What You Will Learn Install and configure Apache Spark with various cluster managers Set up development environments Perform interactive queries using Spark SQL Get to grips with real-time streaming analytics using Spark Streaming Master supervised learning and unsupervised learning using MLlib Build a recommendation engine using MLlib Develop a set of[...]common applications or project types, and solutions that solve complex big data problems Use Apache Spark as your single big data compute platform and master its libraries In DetailBy introducing in-memory persistent storage, Apache Spark eliminates the need to store intermediate data in filesystems, thereby increasing processing speed by up to 100 times.This book will focus on how to analyze large and complex sets of data. Starting with installing and configuring Apache Spark with various cluster managers, you will cover setting up development environments. You will then cover various recipes to perform interactive queries using Spark SQL and real-time streaming with various sources such as Twitter Stream and Apache Kafka. You will then focus on machine learning, including supervised learning, unsupervised learning, and recommendation engine algorithms. After mastering graph processing using GraphX, you will cover various recipes for cluster optimization and troubleshooting.

Ask this book

Your question is answered in the context of this title and author. Each answer uses 2 points.

Sign in to ask the book assistant.

Reader reviews

0 reviews, 4.8 average out of 5

No reviews yet

If you have read this book, help the next reader with your experience.

Write a review

Sign in to publish a review.

Reader questions and answers

Ask a focused question and learn from the community.

Sign in to ask or answer a question.

No questions yet

Be the first to ask a clear, useful question.