Guide to High Performance Distributed Computing: Case Studies with Hadoop, Scalding and Spark (Computer Communications and Networks)

by K.G. Srinivasa and Anil Kumar Muppalla

0 ratings • 0 reviews • 0 shelved

Book cover for Guide to High Performance Distributed Computing

Shelve It

Bookhype may earn a small commission from qualifying purchases. Full disclosure.

Guide to High Performance Distributed Computing: Case Studies with Hadoop, Scalding and Spark (Computer Communications and Networks)

by K.G. Srinivasa and Anil Kumar Muppalla

0 ratings • 0 reviews • 0 shelved

This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.

This Edition
Other Editions

ISBN13 9783319383477
Publish Date 6 October 2016 (first published 9 March 2015)
Publish Status Active
Publish Country CH
Imprint Springer International Publishing AG

Edition Softcover reprint of the original 1st ed. 2015
Format Paperback
Pages 304
Language English