Beyond Hadoop.
The article focuses on open source processing frameworks and real-time computing systems. It comments on the Apache Hadoop software system which uses computer clusters to process large datasets and talks about how online music service Pandora uses Hadoop to analyze data on skipped songs and listener...
| Publicado en: | Communications of the ACM Vol. 56; no. 1; pp. 22 - 25 |
|---|---|
| Autor principal: | |
| Formato: | Artículo |
| Publicado: |
Association for Computing Machinery
Jan2013
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=84631534&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 84631534 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 00010782 ACM jtl: Communications of the ACM issn: 00010782 maglogo: N pubinfo: dt: Jan2013 vid: 56 iid: 1 pid: 68 pub: Association for Computing Machinery artinfo: ui: 84631534 10.1145/2398356.2398364 ppf: 22 ppct: 3 formats: tig: atl: Beyond Hadoop. aug: au: Mone, Gregory su: Real-time computing Parallel computers Distributed computing Electronic file management Filtering software Data distribution sug: subj: Real-time computing Parallel computers Distributed computing Electronic file management Filtering software Data distribution ab: The article focuses on open source processing frameworks and real-time computing systems. It comments on the Apache Hadoop software system which uses computer clusters to process large datasets and talks about how online music service Pandora uses Hadoop to analyze data on skipped songs and listener ratings through machine learning, collective intelligence tasks and collaborative filtering. It mentions search engine company Google's distributed data system, Google File System (GFS), and MapReduce which allows jobs to be broken into smaller pieces and sent to different computers. It states that Hadoop uses two software modules, one a file system similar to GFS which disperses large datasets on multiple computers, and Hadoop MapReduce which splits up and analyzes mined data. pubtype: Periodical doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y dt: @attributes: year: 2013 holdings: @attributes: islocal: N |
|---|