Apache Spark: A Unified Engine for Big Data Processing.

The article discusses the open source computing framework, Apache Spark, which unifies streaming, batch, and interactive big data workloads to unlock new applications. Topics include Spark's use of an RDD programming model, the use of Spark in diverse applications such as batch processing and image...

Descripción completa

Detalles Bibliográficos
Publicado en:Communications of the ACM Vol. 59; no. 11; pp. 56 - 66
Autores principales: ZAHARIA, MATEI, XIN, REYNOLD S., WENDELL, PATRICK, DAS, TATHAGATA, ARMBRUST, MICHAEL, DAVE, ANKUR, XIANGRUI MENG, ROSEN, JOSH, VENKATARAMAN, SHIVARAM, FRANKLIN, MICHAEL J., GHODSI, ALI, GONZALEZ, JOSEPH, SHENKER, SCOTT, STOICA, ION
Formato: Artículo
Publicado: Association for Computing Machinery Nov2016
Materias:
Acceso en línea:Ver este registro en EBSCOhost
Descripción
Sumario:The article discusses the open source computing framework, Apache Spark, which unifies streaming, batch, and interactive big data workloads to unlock new applications. Topics include Spark's use of an RDD programming model, the use of Spark in diverse applications such as batch processing and image processing, and the additional cost of Spark over other specialized systems due to fault tolerance.