Guide to High Performance Distributed Computing: Case Studies with Hadoop, Scalding and Spark (Computer Communications and Networks)
K.G. Srinivasa, Anil Kumar Muppalla
- 出版商: Springer
- 出版日期: 2015-03-09
- 售價: $2,330
- 貴賓價: 9.5 折 $2,214
- 語言: 英文
- 頁數: 304
- 裝訂: Hardcover
- ISBN: 3319134965
- ISBN-13: 9783319134963
-
相關分類:
Hadoop、Spark
-
相關翻譯:
高性能分佈式計算系統開發與實現:基於Hadoop、Scalding和Spark (簡中版)
商品描述
This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.