The World Wide Web as Complex Data Set: Expanding the Digital Humanities into the Twentieth Century and Beyond through Internet Research.
While intellectual property protections effectively frame digital humanities text mining as a field primarily for the study of the nineteenth century, the Internet offers an intriguing object of study for humanists working in later periods. As a complex data source, the World Wide Web presents its o...
| Publicado en: | International Journal of Humanities & Arts Computing: A Journal of Digital Humanities Vol. 10; no. 1; pp. 95 - 110 |
|---|---|
| Autor principal: | |
| Formato: | Artículo |
| Publicado: |
Edinburgh University Press
Mar2016
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=113576270&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 113576270 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 17538548 2QD7 jtl: International Journal of Humanities & Arts Computing: A Journal of Digital Humanities issn: 17538548 maglogo: N pubinfo: dt: Mar2016 vid: 10 iid: 1 pid: 2327 pub: Edinburgh University Press artinfo: ui: 113576270 10.3366/ijhac.2016.0162 ppf: 95 ppct: 15 formats: fmt: @attributes: type: P size: 96KB tig: atl: The World Wide Web as Complex Data Set: Expanding the Digital Humanities into the Twentieth Century and Beyond through Internet Research. aug: au: Black, Michael L. su: Intellectual property World Wide Web -- Research Text mining Data mining Information retrieval sug: subj: Intellectual property World Wide Web -- Research Text mining Data mining Information retrieval keyword: 20th century intellectual property text mining webscraping world wide web ab: While intellectual property protections effectively frame digital humanities text mining as a field primarily for the study of the nineteenth century, the Internet offers an intriguing object of study for humanists working in later periods. As a complex data source, the World Wide Web presents its own methodological challenges for digital humanists, but lessons learned from projects studying large nineteenth century corpora offer helpful starting points. Complicating matters further, legal and ethical questions surrounding web scraping, or the practice of large scale data retrieval over the Internet, will require humanists to frame their research to distinguish it from commercial and malicious activities. This essay reviews relevant research in the digital humanities and new media studies in order to show how web scraping might contribute to humanities research questions. In addition to recommendations for addressing the complex concerns surrounding web scraping this essay also provides a basic overview of the process and some recommendations for resources. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y custom: Copyright of International Journal of Humanities & Arts Computing: A Journal of Digital Humanities is the property of Edinburgh University Press and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. item: International Journal of Humanities & Arts Computing: A Journal of Digital Humanities holder: Edinburgh University Press dt: @attributes: year: 2016 holdings: @attributes: islocal: N |
|---|