Guide to High Performance Distributed Computing

Guide to High Performance Distributed Computing Case Studies With Hadoop, Scalding and Spark - Computer Communications and Networks

2015

Hardback (09 Mar 2015)

Save $19.87

  • RRP $80.55
  • $60.68
Add to basket

Includes delivery to the United States

10+ copies available online - Usually dispatched within 7 days

Publisher's Synopsis

This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.

Book information

ISBN: 9783319134963
Publisher: Springer International Publishing
Imprint: Springer
Pub date:
Edition: 2015
DEWEY: 004.36
DEWEY edition: 23
Language: English
Number of pages: 304
Weight: 632g
Height: 247mm
Width: 166mm
Spine width: 23mm