Abstract
Data curation projects frequently deal with data that were not created for the purposes of long-term preservation and re-use. How can curation of such legacy data be improved by supplying necessary metadata? In this report, we address this and other questions by creating robust metadata for twenty legacy research datasets. We report on quantitative and qualitative metrics of creating domain-specific metadata and propose a four-prong framework of metadata creation for legacy research data. Our findings indicate that there is a steep learning curve in encoding metadata using the FGDC content standard for digital geospatial metadata. Our project also demonstrates that data curators who are handed research data “as is” and are tasked with incorporating such data into a data sharing environment can be very successful in creating descriptive metadata -- particularly, in conducting subject analysis and assigning keywords based on controlled vocabularies and thesauri. At the same time, they need to be aware of limitations in their efforts when it comes to structural and administrative metadata.
The full text of this article is available as a PDF.
Download PDFArticle details
- Published
- Section
- Project Reports
- Published in
- DC-2013--The Lisbon Proceedings
- License
- CC BY 4.0 · open access
- Download
- Download PDF
Indexed in
Described in Dublin Core
This article's metadata, in the vocabulary these proceedings are about.
- dcterms:title
- Collaborate, Automate, Prepare, Prioritize: Creating Metadata for Legacy Research Data
- dcterms:creator
- Kouper, Inna
- Konkiel, Stacy R.
- Liss, Jennifer A.
- Hardesty, Juliet L.
- dcterms:date
- 2013-09-02
- dcterms:identifier
- doi:10.23106/dcmi.952136231
- dcterms:subject
- metadata
- research data
- metadata quality
- legacy data
- dcterms:isPartOf
- DC-2013--The Lisbon Proceedings
- dcterms:publisher
- Dublin Core Metadata Initiative
- dcterms:type
- Text
- dcterms:language
- en
- dcterms:rights
- CC BY 4.0