Catégories
EN

Evaluating Multilingual Metadata Quality in Crossref

Authors : Dennis Donathan, Mike Nason, Marco Tullney, Julie Shi, Juan Pablo Alperin

Introduction

Scholarly research spans multiple languages, making multilingual metadata crucial for organizing and accessing knowledge across linguistic boundaries. These multilingual metadata already exist and are propagated throughout the scholarly publishing infrastructure, but the extent to which they are correctly recorded, or how they affect metadata quality more broadly, is little understood.

Methods

Our study quantifies the prevalence of multilingual records across a sample of publisher metadata and offers an understanding of their completeness, quality, and alignment with metadata standards.

Utilizing the Crossref API to generate a random sample of 519,665 journal article records, we categorize each record into four distinct language types: English monolingual, non-English monolingual, multilingual, and uncategorized. We then investigate the prevalence of programmatically detectable errors and the prevalence of multilingual records within the sample to determine whether multilingualism influences the quality of article metadata.

Results

We find that English-only records are still in the vast majority among metadata found in Crossref, but that, while non-English and multilingual records present unique challenges, they are not a source of significant metadata quality issues and, in a few instances, are more complete or correct than English monolingual records.

Discussion & Conclusion

Our findings contribute to discussions surrounding multilingualism in scholarly communication, serving as a resource for researchers, publishers, and information professionals seeking to enhance the global dissemination of knowledge and foster inclusivity in the academic landscape.

URL : Evaluating Multilingual Metadata Quality in Crossref

DOI : https://doi.org/10.31274/jlsc.19779

Catégories
EN

Repository (R)evolution: Metadata, Interoperability, and Sustainability

Authors : Linda Eells, Julia Kelly, Shannon Farrell

Introduction

Successfully managing an open-access repository requires constant attention to user community priorities in order to inform the development or selection of a platform that fulfills constantly evolving functional demands in an increasingly complex operational environment.

This paper uses AgEcon Search (AES) as an example of the way that varying platforms address the metadata and other platform needs of a repository. AES is a successful subject repository with an international scope that has resided on several different platforms in its 25-year lifespan.

Elements and Considerations

Critical among the technical requirements of a repository is interoperability with other information sources and the ability to accommodate and describe different types of objects, including data. Experienced in the use of easy and widely used Dublin Core (DC), as well as Machine-Readable Cataloging 21 (MARC 21)-based repository platforms, we discuss both metadata schemas from administrative and user perspectives.

Reconsidering underlying metadata issues might positively impact both technical and administrative issues that are currently restricting the development of robust, interoperable systems. As managers of AES, we are uniquely placed to discuss both technical and sustainability issues.

Conclusions

Although many institutional and subject repositories are on platforms that use DC for their metadata, other options are available. MARC, the well-established library standard, can provide the wide range of fields needed to fully and accurately describe the variety of document and data types that are included in repositories.

URL : Repository (R)evolution: Metadata, Interoperability, and Sustainability

DOI : https://doi.org/10.31274/ jlsc.16890 

Catégories
EN

Evaluation and impact of descriptive metadata on academic event management in Ukraine: A quantitative study

Author : Sabina Auhunas

Objective

The study aims to understand the impact of descriptive metadata in academic events. It focuses on the need for analytical frameworks that take into account the characteristics of the events and the interests of the participants.

Design/Methodology/Approach

The article focuses on academic event management and metadata quality based on user
preferences and feedback. It conducted a survey among Ukrainian organizers and scholars between August and October 2022, analyzing the responses of 1,270 participants using descriptive statistics and qualitative analysis in RStudio.

Results/Discussion

The survey showed that most (over 84%) of organizers and academics are dissatisfied with the quality of metadata, with a third rating it as very bad. Frequent errors in metadata emphasized the need for better management, including a preference for using identifiers like ORCID and DOI and a preference for open access to information about academic events.

Conclusions

The results highlight the importance of developing specialized tools for metadata management and standardization of metadata elements in Ukraine to facilitate organization and participation in academic events at national and international levels.

Originality/Value

The study makes an important contribution to the understanding of descriptive metadata management in academic events in Ukraine, suggesting ways to improve efficiency in this area.

URL : Evaluation and impact of descriptive metadata on academic event management in Ukraine: A quantitative study

DOI : https://osf.io/preprints/socarxiv/fxb74

Catégories
EN

Connecting Repositories to the Global Research Community: A Re-Curation Process

Author : Ted Habermann

Over the last decade, significant changes have affected the work that data repositories of all kinds do. First, the emergence of globally unique and persistent identifiers (PIDs) has created new opportunities for repositories to engage with the global research community by connecting existing repository resources to the global research infrastructure. Second, repository use cases have evolved from data discovery to data discovery and reuse, significantly increasing metadata requirements.

To respond to these evolving requirements, we need retrospective and on-going curation, i.e. re-curation, processes that 1) find identifiers and add them to existing metadata to connect datasets to a wider range of communities, and 2) add elements that support reuse to globally connected metadata.

The goal of this work is to introduce the concept of re-curation with representative examples that are generally applicable to many repositories: 1) increasing completeness of affiliations and identifiers for organizations and funders in the Dryad Repository and 2) measuring and increasing FAIRness of DataCite metadata beyond required fields for institutional repositories.

These re-curation efforts are a critical part of reshaping existing metadata and repository processes so they can take advantage of new connections, engage with global research communities, and facilitate data reuse.

URL : Connecting Repositories to the Global Research Community: A Re-Curation Process

DOI : https://doi.org/10.7191/jeslib.739

Catégories
EN

Challenges of Open Science: Problems of Metadata Research and Analysis of Scientometric Indicators of Scientists (Experience of the Scientific Library of Yaroslav Mudryi National Law University)

Authors : Margarita Kulyk

Objective

The scientific publication aims to analyze the development of open science based on international and national experience, as well as to show efforts of the National Law University in the direction of promotion of open science and management of metadata related to scientometrics indicators of scientists.

Methods

To attain the indicated objective, theoretical methods of scientific research were used: literature analysis and systematic approach.

Results

The result of this research is systematization of scientometrics services developed by the Scientific Library of Yaroslav Mudryi National Law University to inform scientists about their scientometric indicators that reflect publication activity in the two main international scientometric databases Scopus and Web of Science.

Conclusions

It is important to promote open science in university repositories, which are one of the services that manage metadata.

URL : Challenges of Open Science: Problems of Metadata Research and Analysis of Scientometric Indicators of Scientists (Experience of the Scientific Library of Yaroslav Mudryi National Law University)

Catégories
FR

Modèles et outils pour la publication de métadonnées d’archives géographiques et de leurs données dérivées

Auteur.ices/Authors : Hersent Melvin, Abadie Nathalie, Duménieu Bertrand, Perret Julien

L’interopérabilité des données dans un projet pluridisciplinaire est primordiale. Prenant l’exemple d’un projet de recherche en histoire spatiale, nous comparerons dans un premier temps les standards et vocabulaires à notre disposition pour décrire des données géographiques et des documents d’archives.

Nous proposons ensuite un alignement entre les standards retenus : l’ISO 19115 et RiC-O. Enfin, nous proposons une architecture de microservices pour la saisie, le stockage, la publication sur le Web et l’interrogation unifiée des métadonnées de nos sources.

URL : Modèles et outils pour la publication de métadonnées d’archives géographiques et de leurs données dérivées

Original location : https://hal.science/hal-04110787

Catégories
EN

What are Researchers’ Needs in Data Discovery? Analysis and Ranking of a Large-Scale Collection of Crowdsourced Use Cases

Authors : Brigitte Mathiak, Nick Juty, Alessia Bardi, Julien Colomb, Peter Kraker

Data discovery is important to facilitate data re-use. In order to help frame the development and improvement of data discovery tools, we collected a list of requirements and users’ wishes.

This paper presents the analysis of these 101 use cases to examine data discovery requirements; these cases were collected between 2019 and 2020. We categorized the information across 12 ‘topics’ and eight types of users.

While the availability of metadata was an expected topic of importance, users were also keen on receiving more information on data citation and a better overview of their field. We conducted and analysed a survey among data infrastructure specialists in a first attempt at ranking the requirements.

Between these data professionals, these rankings were very different, excepting the availability of metadata and data quality assessment.

URL : What are Researchers’ Needs in Data Discovery? Analysis and Ranking of a Large-Scale Collection of Crowdsourced Use Cases

DOI : http://doi.org/10.5334/dsj-2023-003