Guide to High Performance Distributed Computing

Name: Guide to High Performance Distributed Computing
Price: 65.99 USD
Availability: InStock
Author: K. G. Srinivasa; Anil Kumar Muppalla
ISBN: 9783319134970

Case Studies with Hadoop, Scalding and Spark

By:	K. G. Srinivasa; Anil Kumar Muppalla
Publisher:	Springer Nature
Print ISBN:	9783319134963
eText ISBN:	9783319134970
Edition:	0
Copyright:	2015
Format:	Page Fidelity

Select an option

eBook Features

Instant Access

Purchase and read your book immediately

Read Offline

Access your eTextbook anytime and anywhere

Study Tools

Built-in study tools like highlights and more

Read Aloud

Listen and follow along as Bookshelf reads to you

Details

Table of Contents

This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.