Batch loading collections into DSpace: using Perl scripts for automation and quality control.

This paper describes batch loading workflows developed for the Knowledge Bank, The Ohio State University's institutional repository. In the five years since the inception of the repository approximately 80 percent of the items added to the Knowledge Bank, a DSpace repository, have been batch loaded....

Full description

Bibliographic Details
Published in:Information Technology & Libraries Vol. 29; no. 3; pp. 117 - 128
Main Author: Walsh MP
Format: computer program pictorial tables/charts Journal Article
Published: American Library Association Sep2010
Online Access:View this record in EBSCOhost
Description
Summary:This paper describes batch loading workflows developed for the Knowledge Bank, The Ohio State University's institutional repository. In the five years since the inception of the repository approximately 80 percent of the items added to the Knowledge Bank, a DSpace repository, have been batch loaded. Most of the batch loads utilized Perl scripts to automate the process of importing metadata and content files. Custom Perl scripts were used to migrate data from spreadsheets or comma-separated values files into the DSpace archive directory format to build collections and tables of contents and to provide data quality control. Two projects are described to illustrate the process and workflows.