Catégories
EN

Assessing Researcher Data Sharing Practices Using Publicly Available Dataset Metadata: A Reproducible Workflow for Institutional Stakeholders

Authors : Danielle R. Kirsch, Isaac Wink

Introduction

Scholarly communities are experiencing increased emphasis on research data sharing and reuse. Given the range of available data repositories and variability in the use of persistent identifiers (PIDs) for individuals and institutions, aggregating datasets by researchers at a specific institution is a significant challenge. We developed a reproducible workflow to evaluate 1) data repository use by researchers at our institutions, 2) metrics of reuse, and 3) quality of metadata records.

Methods

We used the DataCite REST API to locate metadata records for datasets with creator affiliation names that matched our institutions. We performed substantial cleaning and deduplication before comparing repository use, citations, usage metrics, and metadata completeness.

Results

The most common data repositories were Dryad, Harvard Dataverse, figshare, Zenodo, ICPSR, and one institution’s institutional repository. View, download, and citation counts were available from a limited number of repositories, with some discrepancies between different citation reporting methods. PIDs were more frequently used for authors and affiliations than funders, and the inclusion of PIDs varied across repositories, with Dryad being the most consistent.

Discussion & Conclusion

Although broad trends in repository and PID use were similar between institutions, our analysis also surfaced examples that illustrate inconsistencies in metadata across repositories. Variable implementation of the DataCite metadata schema requires a significant amount of data cleaning to obtain meaningful results. Even then, incomplete metadata makes some datasets impossible to locate. Repositories, funders, researchers, and institutional open data advocates must coordinate to create complete and usable metadata that integrates datasets into the scholarship ecosystem.

URL : Assessing Researcher Data Sharing Practices Using Publicly Available Dataset Metadata: A Reproducible Workflow for Institutional Stakeholders

DOI : https://doi.org/10.31274/jlsc.22913

Catégories
EN

Evaluating Multilingual Metadata Quality in Crossref

Authors : Dennis Donathan, Mike Nason, Marco Tullney, Julie Shi, Juan Pablo Alperin

Introduction

Scholarly research spans multiple languages, making multilingual metadata crucial for organizing and accessing knowledge across linguistic boundaries. These multilingual metadata already exist and are propagated throughout the scholarly publishing infrastructure, but the extent to which they are correctly recorded, or how they affect metadata quality more broadly, is little understood.

Methods

Our study quantifies the prevalence of multilingual records across a sample of publisher metadata and offers an understanding of their completeness, quality, and alignment with metadata standards.

Utilizing the Crossref API to generate a random sample of 519,665 journal article records, we categorize each record into four distinct language types: English monolingual, non-English monolingual, multilingual, and uncategorized. We then investigate the prevalence of programmatically detectable errors and the prevalence of multilingual records within the sample to determine whether multilingualism influences the quality of article metadata.

Results

We find that English-only records are still in the vast majority among metadata found in Crossref, but that, while non-English and multilingual records present unique challenges, they are not a source of significant metadata quality issues and, in a few instances, are more complete or correct than English monolingual records.

Discussion & Conclusion

Our findings contribute to discussions surrounding multilingualism in scholarly communication, serving as a resource for researchers, publishers, and information professionals seeking to enhance the global dissemination of knowledge and foster inclusivity in the academic landscape.

URL : Evaluating Multilingual Metadata Quality in Crossref

DOI : https://doi.org/10.31274/jlsc.19779

Catégories
EN

Repository (R)evolution: Metadata, Interoperability, and Sustainability

Authors : Linda Eells, Julia Kelly, Shannon Farrell

Introduction

Successfully managing an open-access repository requires constant attention to user community priorities in order to inform the development or selection of a platform that fulfills constantly evolving functional demands in an increasingly complex operational environment.

This paper uses AgEcon Search (AES) as an example of the way that varying platforms address the metadata and other platform needs of a repository. AES is a successful subject repository with an international scope that has resided on several different platforms in its 25-year lifespan.

Elements and Considerations

Critical among the technical requirements of a repository is interoperability with other information sources and the ability to accommodate and describe different types of objects, including data. Experienced in the use of easy and widely used Dublin Core (DC), as well as Machine-Readable Cataloging 21 (MARC 21)-based repository platforms, we discuss both metadata schemas from administrative and user perspectives.

Reconsidering underlying metadata issues might positively impact both technical and administrative issues that are currently restricting the development of robust, interoperable systems. As managers of AES, we are uniquely placed to discuss both technical and sustainability issues.

Conclusions

Although many institutional and subject repositories are on platforms that use DC for their metadata, other options are available. MARC, the well-established library standard, can provide the wide range of fields needed to fully and accurately describe the variety of document and data types that are included in repositories.

URL : Repository (R)evolution: Metadata, Interoperability, and Sustainability

DOI : https://doi.org/10.31274/ jlsc.16890 

Catégories
EN

Evaluation and impact of descriptive metadata on academic event management in Ukraine: A quantitative study

Author : Sabina Auhunas

Objective

The study aims to understand the impact of descriptive metadata in academic events. It focuses on the need for analytical frameworks that take into account the characteristics of the events and the interests of the participants.

Design/Methodology/Approach

The article focuses on academic event management and metadata quality based on user
preferences and feedback. It conducted a survey among Ukrainian organizers and scholars between August and October 2022, analyzing the responses of 1,270 participants using descriptive statistics and qualitative analysis in RStudio.

Results/Discussion

The survey showed that most (over 84%) of organizers and academics are dissatisfied with the quality of metadata, with a third rating it as very bad. Frequent errors in metadata emphasized the need for better management, including a preference for using identifiers like ORCID and DOI and a preference for open access to information about academic events.

Conclusions

The results highlight the importance of developing specialized tools for metadata management and standardization of metadata elements in Ukraine to facilitate organization and participation in academic events at national and international levels.

Originality/Value

The study makes an important contribution to the understanding of descriptive metadata management in academic events in Ukraine, suggesting ways to improve efficiency in this area.

URL : Evaluation and impact of descriptive metadata on academic event management in Ukraine: A quantitative study

DOI : https://osf.io/preprints/socarxiv/fxb74

Catégories
EN

Connecting Repositories to the Global Research Community: A Re-Curation Process

Author : Ted Habermann

Over the last decade, significant changes have affected the work that data repositories of all kinds do. First, the emergence of globally unique and persistent identifiers (PIDs) has created new opportunities for repositories to engage with the global research community by connecting existing repository resources to the global research infrastructure. Second, repository use cases have evolved from data discovery to data discovery and reuse, significantly increasing metadata requirements.

To respond to these evolving requirements, we need retrospective and on-going curation, i.e. re-curation, processes that 1) find identifiers and add them to existing metadata to connect datasets to a wider range of communities, and 2) add elements that support reuse to globally connected metadata.

The goal of this work is to introduce the concept of re-curation with representative examples that are generally applicable to many repositories: 1) increasing completeness of affiliations and identifiers for organizations and funders in the Dryad Repository and 2) measuring and increasing FAIRness of DataCite metadata beyond required fields for institutional repositories.

These re-curation efforts are a critical part of reshaping existing metadata and repository processes so they can take advantage of new connections, engage with global research communities, and facilitate data reuse.

URL : Connecting Repositories to the Global Research Community: A Re-Curation Process

DOI : https://doi.org/10.7191/jeslib.739

Catégories
EN

Challenges of Open Science: Problems of Metadata Research and Analysis of Scientometric Indicators of Scientists (Experience of the Scientific Library of Yaroslav Mudryi National Law University)

Authors : Margarita Kulyk

Objective

The scientific publication aims to analyze the development of open science based on international and national experience, as well as to show efforts of the National Law University in the direction of promotion of open science and management of metadata related to scientometrics indicators of scientists.

Methods

To attain the indicated objective, theoretical methods of scientific research were used: literature analysis and systematic approach.

Results

The result of this research is systematization of scientometrics services developed by the Scientific Library of Yaroslav Mudryi National Law University to inform scientists about their scientometric indicators that reflect publication activity in the two main international scientometric databases Scopus and Web of Science.

Conclusions

It is important to promote open science in university repositories, which are one of the services that manage metadata.

URL : Challenges of Open Science: Problems of Metadata Research and Analysis of Scientometric Indicators of Scientists (Experience of the Scientific Library of Yaroslav Mudryi National Law University)

Catégories
FR

Modèles et outils pour la publication de métadonnées d’archives géographiques et de leurs données dérivées

Auteur.ices/Authors : Hersent Melvin, Abadie Nathalie, Duménieu Bertrand, Perret Julien

L’interopérabilité des données dans un projet pluridisciplinaire est primordiale. Prenant l’exemple d’un projet de recherche en histoire spatiale, nous comparerons dans un premier temps les standards et vocabulaires à notre disposition pour décrire des données géographiques et des documents d’archives.

Nous proposons ensuite un alignement entre les standards retenus : l’ISO 19115 et RiC-O. Enfin, nous proposons une architecture de microservices pour la saisie, le stockage, la publication sur le Web et l’interrogation unifiée des métadonnées de nos sources.

URL : Modèles et outils pour la publication de métadonnées d’archives géographiques et de leurs données dérivées

Original location : https://hal.science/hal-04110787