Objective Vision: Confusing the Subject of Computer Vision.

Convolutional neural networks (CNNs) are a key technology powering the automated technologies of seeing known as computer vision. CNNs have been especially successful in systems that perform object recognition from visual data. This article examines the persistence of a mid‐twentieth‐century ontolog...

Full description

Bibliographic Details
Published in:Social Text Vol. 41; no. 3; pp. 35 - 56
Main Author: Dobson, James E.
Format: Article
Published: Duke University Press Sep2023
Subjects:
Online Access:View this record in EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=172364945&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 172364945
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        01642472
        6Z2
      jtl: Social Text
      issn: 01642472
      maglogo: N
    pubinfo:
      dt: Sep2023
      vid: 41
      iid: 3
      pid: 154
      pub: Duke University Press
    artinfo:
      ui:
        172364945
        10.1215/01642472-10613653
      ppf: 35
      ppct: 21
      formats:
      tig:
        atl: Objective Vision: Confusing the Subject of Computer Vision.
      aug:
        au: Dobson, James E.
      su:
        Computer vision
        Convolutional neural networks
      sug:
        subj:
          Computer vision
          Convolutional neural networks
      keyword:
        computer vision
        convolutional neural networks
        explainability
        ontology
      ab: Convolutional neural networks (CNNs) are a key technology powering the automated technologies of seeing known as computer vision. CNNs have been especially successful in systems that perform object recognition from visual data. This article examines the persistence of a mid‐twentieth‐century ontology of the digital image in these contemporary technologies. While CNNs are multidimensional, their ontology flattens distinctions between background and foreground, between subjects and objects, and even the relations established among the categories of information used to organize and train these models. This ontology enables the introduction and amplification of bias and troubling correlations and the transfer or slippage of learned associations between humans and objects found in the training image archives. Inspecting and interpreting what CNNs learn and index through their complex architectures can be difficult if not impossible because of how they encode and obfuscate quite human ways of seeing the world and the image repertoires used to train these algorithms that are rife with residues of prior representations.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      dt:
        @attributes:
          year: 2023
    holdings:
      @attributes:
        islocal: N