Catégories
EN

The role of metadata in reproducible computational research

Authors : Jeremy Leipzig, Daniel Nüst, Charles Tapley Hoyt, Stian Soiland-Reyes, Karthik Ram, Jane Greenberg

Reproducible computational research (RCR) is the keystone of the scientific method for in silico analyses, packaging the transformation of raw data to published results.

In addition to its role in research integrity, RCR has the capacity to significantly accelerate evaluation and reuse. This potential and wide-support for the FAIR principles have motivated interest in metadata standards supporting RCR.

Metadata provides context and provenance to raw data and methods and is essential to both discovery and validation. Despite this shared connection with scientific data, few studies have explicitly described the relationship between metadata and RCR.

This article employs a functional content analysis to identify metadata standards that support RCR functions across an analytic stack consisting of input data, tools, notebooks, pipelines, and publications.

Our article provides background context, explores gaps, and discovers component trends of embeddedness and methodology weight from which we derive recommendations for future work.

URL : The role of metadata in reproducible computational research

Original location : https://arxiv.org/abs/2006.08589

Catégories
Non classé

SPI-Hub™: a gateway to scholarly publishing information

Authors : Taneya Y. Koonce, Mallory N. Blasingame, Jerry Zhao, Annette M. Williams, Jing Su, Spencer J. DesAutels, Dario A. Giuse, John D. Clark, Zachary E. Fox, Nunzia Bettinsoli Giuse

Background

Advances in the health sciences rely on sharing research and data through publication. As information professionals are often asked to contribute their knowledge to assist clinicians and researchers in selecting journals for publication, the authors recognized an opportunity to build a decision support tool, SPI-Hub: Scholarly Publishing Information Hub™, to capture the team’s collective publishing industry knowledge, while carefully retaining the quality of service.

Case Presentation

SPI-Hub’s decision support functionality relies on a data framework that describes journal publication policies and practices through a newly designed metadata structure, the Knowledge Management Journal Record™.

Metadata fields are populated through a semi-automated process that uses custom programming to access content from multiple sources. Each record includes 25 metadata fields representing best publishing practices. Currently, the database includes more than 24,000 health sciences journal records.

To correctly capture the resources needed for both completion and future maintenance of the project, the team conducted an internal study to assess time requirements for completing records through different stages of automation.

Conclusions

The journal decision support tool, SPI-Hub, provides an opportunity to assess publication practices by compiling data from a variety of sources in a single location.

Automated and semi-automated approaches have effectively reduced the time needed for data collection.

Through a comprehensive knowledge management framework and the incorporation of multiple quality points specific to each journal, SPI-Hub provides prospective users with both recommendations for publication and holistic assessment of the trustworthiness of journals in which to publish research and acquire trusted knowledge.

URL : SPI-Hub™: a gateway to scholarly publishing information

Original location : http://jmla.pitt.edu/ojs/jmla/article/view/815

Catégories
EN

Monitoring agreements with open access elements: why article-level metadata are important

Authors : Mafalda Marques, Saskia Woutersen-Windhouwer, Arja Tuuliniemi

Agreements with open access (OA) elements (e.g. agreements with APC discounts, offsetting agreements, read and publish agreements) have been increasing in number in the last few years.

With more agreements including some form of OA, consortia and academic institutions need to monitor the number of OA publications, the costs and the value of these agreements. Publishers are therefore required to account for the articles published OA to consortia, academic institutions and research funders.

One way publishers can do so is by providing regular reports with article-level metadata. This article uses the Knowledge Exchange (KE) and the Efficiency and Standards for Article Charges (ESAC) initiative recommendations as a check-list to assess what article-level metadata consortia request from publishers and what metadata publishers deliver to consortia.

KE countries’ agreements with major publishers were analysed to assess how far consortia and publishers are from requesting and providing article-level metadata. The results from this research can be used as a benchmark to determine how major publishers were performing until early 2019 and prior to Plan S coming into effect in 2021.

A recommendation is made that publishers use the article-level metadata check-list as a template to provide the metadata recommended by KE and ESAC.

URL : Monitoring agreements with open access elements: why article-level metadata are important

DOI : http://doi.org/10.1629/uksg.489

Catégories
EN

Open Social Knowledge Creation and Library and Archival Metadata

Authors: Dean Seeman, Heather Dean

Standardization both reflects and facilitates the collaborative and networked approach to metadata creation within the fields of librarianship and archival studies.

These standards—such as Resource Description and Access and Rules for Archival Description—and the theoretical frameworks they embody enable professionals to work more effectively together.

Yet such guidelines also determine who is qualified to undertake the work of cataloging and processing in libraries and archives. Both fields are empathetic to facilitating user-generated metadata and have taken steps towards collaborating with their research communities (as illustrated, for example, by social tagging and folksonomies) but these initial experiments cannot yet be regarded as widely adopted and radically open and social.

This paper explores the recent histories of descriptive work in libraries and archives and the challenges involved in departing from deeply established models of metadata creation.

URL : Open Social Knowledge Creation and Library and Archival Metadata

Catégories
EN

Harvesting the Academic Landscape: Streamlining the Ingestion of Professional Scholarship Metadata into the Institutional Repository

Authors : Jonathan Bull, Teresa Auch Schultz

INTRODUCTION

Although librarians initially hoped institutional repositories (IRs) would grow through researcher self-archiving, practice shows that growth is much more likely through library-directed deposit. Libraries must then find efficient ways to ingest material into their IR to ensure growth and relevance.

DESCRIPTION OF PROGRAM

Valparaiso University developed and implemented a workflow that was semiautomated to help cut down on the time needed to ingest articles into its IR, ValpoScholar. The workflow, which continues to be refined, makes use of practices and ideas used by other repositories to more efficiently collect metadata for items and upload them to the repository.

NEXT STEPS

The article discusses the pros and cons of this workflow and areas of ingesting that still need to be addressed, including adding full-text items, checking copyright policies, managing student staffing, and dealing with hurdles created by the repository’s software.

URL : Harvesting the Academic Landscape: Streamlining the Ingestion of Professional Scholarship Metadata into the Institutional Repository

DOI : http://doi.org/10.7710/2162-3309.2201

Catégories
EN

Big Metadata, Smart Metadata, and Metadata Capital: Toward Greater Synergy Between Data Science and Metadata

Author : Jane Greenberg

Purpose

The purpose of the paper is to provide a framework for addressing the disconnect between metadata and data science. Data science cannot progress without metadata research. This paper takes steps toward advancing the synergy between metadata and data science, and identifies pathways for developing a more cohesive metadata research agenda in data science.

Design/methodology/approach

This paper identifies factors that challenge metadata research in the digital ecosystem, defines metadata and data science, and presents the concepts big metadata, smart metadata, and metadata capital as part of a metadata lingua franca connecting to data science.

Findings

The “utilitarian nature” and “historical and traditional views” of metadata are identified as two intersecting factors that have inhibited metadata research. Big metadata, smart metadata, and metadata capital are presented as part of a metadata lingua franca to help frame research in the data science research space.

Research limitations

There are additional, intersecting factors to consider that likely inhibit metadata research, and other significant metadata concepts to explore.

Practical implications

The immediate contribution of this work is that it may elicit response, critique, revision, or, more significantly, motivate research. The work presented can encourage more researchers to consider the significance of metadata as a research worthy topic within data science and the larger digital ecosystem.

Originality/value

Although metadata research has not kept pace with other data science topics, there is little attention directed to this problem. This is surprising, given that metadata is essential for data science endeavors. This examination synthesizes original and prior scholarship to provide new grounding for metadata research in data science.

URL : Big Metadata, Smart Metadata, and Metadata Capital: Toward Greater Synergy Between Data Science and Metadata

DOI : https://doi.org/10.1515/jdis-2017-0012

Catégories
EN

High-Quality Metadata and Repository Staffing: Perceptions of United States–Based OpenDOAR Participants

Digital repositories require good metadata, created according to community-based principles that include provisions for interoperability. When metadata is of high quality, digital objects become sharable and metadata can be harvested and reused outside of the local system.

A sample of U.S.-based repository administrators from the OpenDOAR initiative were surveyed to understand aspects of the quality and creation of their metadata, and how their metadata could improve.

Most respondents (65%) thought their metadata was of average quality; none thought their metadata was high quality or poor quality. The discussion argues that increased strategic staffing will alleviate many perceived issues with metadata quality.

URL : http://tandfonline.com/doi/full/10.1080/01639374.2015.1116480