Guide to High Performance Distributed Computing
60,80 €*
Sofort verfügbar, Lieferzeit: 1-3 Tage
Produktnummer:
9783319134963
This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.
Autor: | Srinivasa, K. G. Muppalla, Anil Kumar |
---|---|
EAN: | 9783319134963 |
Sprache: | Englisch |
Seitenzahl: | 304 |
Produktart: | Gebunden |
Verlag: | Springer Springer, Berlin Springer International Publishing |
Untertitel: | Case Studies with Hadoop, Scalding and Spark |
Schlagworte: | Hadoop Spark (Apache) |
Größe: | 25 × 158 × 243 |
Gewicht: | 659 g |