22 papers · ranked by Valyu relevance
Dawei Lin, Matthew McAuliffe, Kim D. Pruitt, Anupama Gururaj + 3 more
'Christine Melchior' 'Charles Schmitt' 'Susan N. Wright'] The demand for open data and open science is on the rise, fueled by expectations from the scientific community, calls to increase transparency and reproducibility in research findings, and developments such as the Final Data Management and Sharing Policy from…
Heinz Pampel, Paul Vierkant, Frank Scholze, Roland Bertelmann + 7 more
'Maxi Kindling' 'Jens Klump' 'Hans-Jürgen Goebelbecker' 'Jens Gundlach' 'Peter Schirmbacher' 'Uwe Dierolf' 'Hussein Suleman'] Researchers require infrastructures that ensure a maximum of accessibility, stability and reliability to facilitate working with and sharing of research data. Such infrastructures are being…
Matthew J. Harvey, Andrew McLean, Henry S. Rzepa
The design and use of a metadata-driven data repository for research data management is described. Metadata is collected automatically during the submission process whenever possible and is registered with DataCite in accordance with their current metadata schema, in exchange for a persistent digital object identifier.…
Francesco Napolitano
Background Reproducibility in Data Analysis research has long been a significant concern, particularly in the areas of Bioinformatics and Computational Biology. Towards the aim of developing reproducible and reusable processes, Data Analysis management tools can help giving structure and coherence to complex data…
Pierre Tremouilhac, Chia‐Lin Lin, Pei‐Chi Huang, Yu‐Chieh Huang + 7 more
'An Nguyen' 'Nicole Jung' 'Felix Bach' 'Robert Ulrich' 'Bernhard Neumair' 'Achim Streit' 'Stefan Bräse'] The repository Chemotion provides solutions for current challenges to store research data in a feasible manner. A main advantage of Chemotion is the comprehensive functionality, offering options to collect, prepare…
Henry Rzepa
The history of the development of two tools for managing research resources and the data produced from them is summarised. These tools are a portal or electronic laboratory notebook for computational chemistry interfaced in one direction to a high-performance computing resource and in the other direction to a modern…
Philippa C. Griffin, Jyoti Khadake, Kate S. LeMay, Suzanna E. Lewis + 23 more
Throughout history, the life sciences have been revolutionised by technological advances; in our era this is manifested by advances in instrumentation for data generation, and consequently researchers now routinely handle large amounts of heterogeneous data in digital formats. The simultaneous transitions towards…
Tom Hanika, Robert Jäschke
Data is always at the center of the theoretical development and investigation of the applicability of formal concept analysis. It is therefore not surprising that a large number of data sets are repeatedly used in scholarly articles and software tools, acting as de facto standard data sets. However, the distribution of…
Yufeng Luo, Roland Haas, Qian Zhang, Gabrielle Allen
Data sharing is essential in the numerical simulations research. We introduce a data repository, DataVault, that is designed for data sharing, search and analysis. A comparative study of existing repositories is performed to analyze features that are critical to a data repository. We describe the architecture…
Yi Shen
Nowadays, we have the emergence and abundance of many different data repositories and archival systems for scientific data discovery, use, and analysis. With the burgeoning data sharing platforms available, this study addresses how natural resources and environmental scientists navigate these diverse data sources, what…
David Herrmann, Patrick Hodapp, Martin Starman, Pei-Chi Huang + 16 more
Analytical data in chemistry and other disciplines is usually generated in different formats and lacks common data and metadata standards that are necessary for a FAIR handling of research data. In the work presented herein, we describe a workflow that uses non-standardized, in some cases proprietary data formats from…
Bohdan B. Khomtchouk, Kasra A. Vand, Thor Wahlestedt, Kelly Khomtchouk + 2 more
We propose a search engine and file retrieval system for all bioinformatics databases worldwide. PubData searches biomedical data in a user-friendly fashion similar to how PubMed searches biomedical literature. PubData is built on novel network programming, natural language processing, and artificial intelligence…
Chia-Lin Lin, Pei-Chi Huang, Simone Graessle, Christoph Grathwol + 20 more
Results of scientific work in chemistry can usually be obtained in the form of materials and data. A big step towards transparency and reproducibility of the scientific work can be gained if scientists publish their data in a FAIR (Findable, Accessible, Interoperable, Reusable) manner in research data repositories.…
Fritz Lekschas, Nils Gehlenborg
The number of data sets in biomedical repositories has grown rapidly over the past decade, providing scientists in fields like genomics and other areas of high-throughput biology with tremendous opportunities to re-use data. Scientists are able to test hypotheses computationally instead of generating their own data, to…
Marta Teperek, Rhys Morgan, Michelle Ellefson, Danny Kingsley
Repository managers can never be one hundred percent sure of the security ofhosted research data. Even assuming that human errors and technical faults will never happen, repositories can be subject to hacking attacks. Therefore, repositories accepting personal/sensitive data (or other forms of restricted data) should…
Birol Tilki, Thomas Schulenberg, Steve Canham, Rita Banzi + 2 more
'Wolfgang Kuchinke' 'Christian Ohmann'] Background: Given the increasing number and heterogeneity of data repositories, an improvement and harmonisation of practice within repositories for clinical trial data is urgently needed. The objective of the study was to develop and evaluate a demonstrator repository, using a…
Dorothea Strecker, Heinz Pampel, Rouven Schabinger, Nina Leonie Weisweiler
'Nina Leonie Weisweiler'] Currently, there is limited research investigating the phenomenon of research data repositories being shut down, and the impact this has on the long-term availability of data. This paper takes an infrastructure perspective on the preservation of research data by using a registry to identify…
Yasin El Abiead, Michael Strobel, Thomas Payne, Eoin Fahy + 13 more
Public untargeted metabolomics data is a growing resource for metabolite and phenotype discovery; however, accessing and utilizing these data across repositories pose significant challenges. Therefore, we've developed pan-repository universal identifiers and harmonized cross-repository metadata. This novel ecosystem…
Authors not listed
The discoverability and reusability of data is critical for machine learning to drive new discovery in the chemical sciences, and the ‘FAIR Guiding Principles for scientific data management and stewardship’ provide a measurable set of guidelines that can be used to ensure the accessibility of reusable data. We…
George Macgregor, Joy Carol Davidson
This article seeks to determine the extent to which the principle of persistence is observed by repositories and the organizations that operate them. We also evaluate the impact that negative repository persistence levels may be having on the scholarly record. We do this by interrogating and combining data about…
Dimitri Yatsenko, Jacob Reimer, Alexander S. Ecker, Edgar Y. Walker + 6 more
The rise of big data in modern research poses serious challenges for data management: Large and intricate datasets from diverse instrumentation must be precisely aligned, annotated, and processed in a variety of ways to extract new insights. While high levels of data integrity are expected, research teams have diverse…
Heidi J. Imker
Online resources enable unfettered access to and analysis of scientific data and are considered crucial for the advancement of modern science. Despite the clear power of online data resources, including web-available databases, proliferation can be problematic due to challenges in sustainability and long-term…