Guide to High Performance Distributed Computing

Case Studies with Hadoop, Scalding and Spark, Computer Communications and Networks

Srinivasa, K G/Muppalla, Anil Kumar

Springer Verlag GmbH

Informatik, EDV/Datenkommunikation, Netzwerke

Erschienen am 09.03.2015, 1. Auflage 2015

53,49 €

(inkl. MwSt.)

excl. Versandkosten

Lieferbar innerhalb 1 - 2 Wochen

In den Warenkorb

Auf Wunschliste

Bibliografische Daten

ISBN/EAN: 9783319134963

Sprache: Englisch

Umfang: xvii, 304 S., 43 s/w Illustr., 304 p. 43 illus.

Format (T/L/B): 2.2 x 24.2 x 16.3 cm

Einband: gebundenes Buch

Beschreibung

This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.