,

Guide to High Performance Distributed Computing

Case Studies with Hadoop, Scalding and Spark

Gebonden Engels 2015 2014e druk 9783319134963
Verwachte levertijd ongeveer 9 werkdagen

Samenvatting

This timely text/reference describes the development and implementation of large-scale distributed processing systems using open source tools and technologies. Comprehensive in scope, the book presents state-of-the-art material on building high performance distributed computing systems, providing practical guidance and best practices as well as describing theoretical software frameworks. Features: describes the fundamentals of building scalable software systems for large-scale data processing in the new paradigm of high performance distributed computing; presents an overview of the Hadoop ecosystem, followed by step-by-step instruction on its installation, programming and execution; Reviews the basics of Spark, including resilient distributed datasets, and examines Hadoop streaming and working with Scalding; Provides detailed case studies on approaches to clustering, data classification and regression analysis; Explains the process of creating a working recommender system using Scalding and Spark.

Specificaties

ISBN13:9783319134963
Taal:Engels
Bindwijze:gebonden
Aantal pagina's:327
Uitgever:Springer International Publishing
Druk:2014

Lezersrecensies

Wees de eerste die een lezersrecensie schrijft!

Inhoudsopgave

<p>Part I: Programming Fundamentals of High Performance Distributed Computing</p><p>Introduction</p><p>Getting Started with Hadoop</p><p>Getting Started with Spark</p><p>Programming Internals of Scalding and Spark</p><p>Part II: Case studies using Hadoop, Scalding and Spark</p><p>Case Study I: Data Clustering using Scalding and Spark</p><p>Case Study II: Data Classification using Scalding and Spark</p><p>Case Study III: Regression Analysis using Scalding and Spark</p><p>Case Study IV: Recommender System using Scalding and Spark</p>

Managementboek Top 100

Rubrieken

    Personen

      Trefwoorden

        Guide to High Performance Distributed Computing