Catégories
EN

Openness in Big Data and Data Repositories. The Application of an Ethics Framework for Big Data in Healthand Research

Authors : Vicki Xafis, Markus K. Labude

There is a growing expectation, or even requirement, for researchers to deposit a variety of research data in data repositories as a condition of funding or publication. This expectation recognizes the enormous benefits of data collected and created for research purposes being made available for secondary uses, as open science gains increasing support.

This is particularly so in the context of big data, especially where health data is involved. There are, however, also challenges relating to the collection, storage, and re-use of research data.

This paper gives a brief overview of the landscape of data sharing via data repositories and discusses some of the key ethical issues raised by the sharing of health-related research data, including expectations of privacy and confidentiality, the transparency of repository governance structures, access restrictions, as well as data ownership and the fair attribution of credit.

To consider these issues and the values that are pertinent, the paper applies the deliberative balancing approach articulated in the Ethics Framework for Big Data in Health and Research (Xafis et al. 2019) to the domain of Openness in Big Data and Data Repositories.

Please refer to that article for more information on how this framework is to be used, including a full explanation of the key values involved and the balancing approach used in the case study at the end.

URL : Openness in Big Data and Data Repositories. The Application of an Ethics Framework for Big Data in Healthand Research

DOI : https://doi.org/10.1007/s41649-019-00097-z

Catégories
FR

Ouverture des données de recherche dans le domaine académique suisse : outils pour le choix d’une stratégie institutionnelle en matière de dépôt de données

Auteur/Author : Marielle Guirlet

Le contexte actuel de l’Open Science se traduit par des exigences d’ouverture des données de recherche. Le dépôt de données est un instrument crucial pour partager publiquement ces données.

Néanmoins, l’offre actuelle pléthorique et très diverse rend la sélection du dépôt difficile pour les chercheurs et les chercheuses. Pour les aider, leurs institutions d’affiliation émettent des recommandations pour le choix du meilleur dépôt. Elles proposent parfois aussi leur propre dépôt de données ou envisagent de le créer.

Cette étude, basée sur un travail de Master en sciences de l’information, s’intéresse à la démarche que les institutions académiques suisses peuvent suivre pour définir leur stratégie de soutien aux chercheurs et aux chercheuses en termes de dépôt.

Elle identifie aussi les informations qui vont aider ces institutions à choisir entre orienter ces chercheurs et ces chercheuses vers un dépôt existant (et lequel) et créer un nouveau dépôt, et aux spécifications que ce dépôt doit remplir.

Après avoir défini les concepts des données de recherche et des dépôts ouverts, les fonctionnalités, les outils et les services nécessaires à un dépôt pour mettre en œuvre le partage public de données sont discutés.

A partir des critères utilisés par la certification CoreTrustSeal pour évaluer la qualité d’un dépôt, et en tenant compte de ces fonctionnalités, de ces outils et ces services, un modèle de description d’un dépôt de données de recherche ouvertes est élaboré. Ce modèle peut être utilisé pour l’évaluation d’un dépôt existant ou pour la conception d’un nouveau dépôt.

Les stratégies de neuf institutions académiques suisses en matière de dépôt de données de recherche, dépôts utilisés et dépôts recommandés, sont analysées. Des recommandations sont formulées, sur la base des bonnes pratiques observées.

Des outils développés pour le choix de la meilleure stratégie en termes de dépôt de données de recherche ouvertes sont alors présentés. Un vade-mecum se présentant comme une liste de questions permet de collecter certaines informations utiles.

Un guide décisionnel accompagne l’institution dans sa réflexion et lui permet de choisir sa stratégie de façon éclairée, avec les informations collectées précédemment. Une fois cette stratégie choisie, des informations complémentaires et des recommandations sont disponibles pour sa mise en pratique.

Une version prototype de ces outils pour navigateur Internet est aussi présentée. Elle est adaptable à une évolution du contexte et transposable à d’autres pays.

URL : http://www.ressi.ch/num21/article182

Catégories
EN

Repository Approaches to Improving the Quality of Shared Data and Code

Authors : Ana Trisovic, Katherine Mika, Ceilyn Boyd, Sebastian Feger, Mercè Crosas

Sharing data and code for reuse has become increasingly important in scientific work over the past decade. However, in practice, shared data and code may be unusable, or published results obtained from them may be irreproducible.

Data repository features and services contribute significantly to the quality, longevity, and reusability of datasets.

This paper presents a combination of original and secondary data analysis studies focusing on computational reproducibility, data curation, and gamified design elements that can be employed to indicate and improve the quality of shared data and code.

The findings of these studies are sorted into three approaches that can be valuable to data repositories, archives, and other research dissemination platforms.

URL : Repository Approaches to Improving the Quality of Shared Data and Code

DOI : https://doi.org/10.3390/data6020015

Catégories
EN

Improving Opportunities for New Value of Open Data: Assessing and Certifying Research Data Repositories

Author : Robert R. Downs

Investments in research that produce scientific and scholarly data can be leveraged by enabling the resulting research data products and services to be used by broader communities and for new purposes, extending reuse beyond the initial users and purposes for which the data were originally collected.

Submitting research data to a data repository offers opportunities for the data to be used in the future, providing ways for new benefits to be realized from data reuse. Improvements to data repositories that facilitate new uses of data increase the potential for data reuse and for gains in the value of open data products and services that are associated with such reuse.

Assessing and certifying the capabilities and services offered by data repositories provides opportunities for improving the repositories and for realizing the value to be attained from new uses of data.

The evolution of data repository certification instruments is described and discussed in terms of the implications for the curation and continuing use of research data.

URL : Improving Opportunities for New Value of Open Data: Assessing and Certifying Research Data Repositories

DOI : http://doi.org/10.5334/dsj-2021-001

Catégories
FR

Entrepôts de données de recherche : mesurer l’impact de l’Open Science à l’aune de la consultation des jeux de données déposés

Auteur/Author  : Violaine Rebouillat

Les décennies 2000 et 2010 ont vu se développer un nombre croissant de e-infrastructures de recherche, rendant plus aisés le partage et l’accès aux données scientifiques. Cette tendance s’est vue renforcée par l’essor de politiques d’ouverture des données, lesquelles ont donné lieu à une multiplication de réservoirs de données – aussi appelés « entrepôts de données ». Quantifier et qualifier l’utilisation des données rendues publiques constitue un élément essentiel pour évaluer l’impact des politiques d’ouverture des données.

Dans cet article, nous questionnons l’utilisation des données déposées dans les entrepôts. Dans quelle mesure ces données sont-elles consultées et téléchargées ?

L’article présente les premiers résultats d’une enquête quantitative auprès de 20 entrepôts. Il esquisse deux tendances, qui restent à ce stade propres à l’échantillon étudié, à savoir : (1) l’augmentation globale du nombre de consultations, de téléchargements et de données disponibles dans les entrepôts sur la période étudiée (2015-2020), et (2) la concentration des téléchargements sur une proportion relativement faible des données de l’entrepôt (de l’ordre de 10% à 30%).

URL : https://hal.archives-ouvertes.fr/hal-02928817/

Catégories
EN

Data journals: incentivizing data access and documentation within the scholarly communication system

Authors : William H. Walters

Data journals provide strong incentives for data creators to verify, document and disseminate their data. They also bring data access and documentation into the mainstream of scholarly communication, rewarding data creators through existing mechanisms of peer-reviewed publication and citation tracking.

These same advantages are not generally associated with data repositories, or with conventional journals’ data-sharing mandates. This article describes the unique advantages of data journals. It also examines the data journal landscape, presenting the characteristics of 13 data journals in the fields of biology, environmental science, chemistry, medicine and health sciences.

These journals vary considerably in size, scope, publisher characteristics, length of data reports, data hosting policies, time from submission to first decision, article processing charges, bibliographic index coverage and citation impact. They are similar, however, in their peer review criteria, their open access license terms and the characteristics of their editorial boards.

URL : Data journals: incentivizing data access and documentation within the scholarly communication system

DOI : http://doi.org/10.1629/uksg.510

Catégories
EN

Research Data Management as an Integral Part of the Research Process of Empirical Disciplines Using Landscape Ecology as an Example

Authors : Winfried Schröder, Stefan Nickel

Research Data Management (RDM) is regarded as an elementary component of empirical disciplines. Taking Landscape Ecology in Germany as an example the article demonstrates how to integrate RDM into the research design as a complement of the classic quality control and assurance in empirical research that has, so far, generally been limited to data production.

Sharing and reuse of empirical data by scientists as well as thorough peer reviews of knowledge produced by empirical research requires that the problem of the research in question, the operationalized definitions of the objects of investigation and their representative selection are documented and archived as well as the methods of data production including indicators for data quality and all data collected and produced.

On this basis, the extent to which this complemented design of research processes has already been realized is demonstrated by research projects of the Chair of Landscape Ecology at the University of Vechta, Germany.

This study is part of a joined research project on Research Data Management funded by the German Federal Ministry of Education and Research.

URL : Research Data Management as an Integral Part of the Research Process of Empirical Disciplines Using Landscape Ecology as an Example

DOI : http://doi.org/10.5334/dsj-2020-026