Guide to High Performance Distributed Computing

This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.

Verwandte Artikel

Guide to High Performance Distributed Computing Muppalla, Anil Kumar, Srinivasa, K. G.

53,49 €*

Weitere Produkte vom selben Autor

Practical Social Network Analysis with Python Raj P. M., Krishna, Srinivasa, K. G., Mohan, Ankith

149,79 €*
Network Data Analytics Srinivasa, K. G., H., Srinidhi, G. M., Siddesh

117,69 €*
The Internet of Educational Things Srinivasa, K. G., Kurni, Muralidhar

160,49 €*
Network Data Analytics Srinivasa, K. G., H., Srinidhi, G. M., Siddesh

117,69 €*