26 papers · ranked by Valyu relevance
Greenland, Sander
The study of associations and their causal explanations is a central research activity whose methodology varies tremendously across fields. Even within specialized subfields, comparisons across textbooks and journals reveals that the basics are subject to considerable variation and controversy. This variation is often…
Sean Simmons
Measuring changes in cell type composition between conditions (disease vs not, knockout vs wild type, treated vs not, etc) is fast becoming a standard step in single cell RNA-Seq analysis. Despite that, there is no agreement on the best approach for this type of analysis. As such, we decided to test numerous methods…
Laura C. Guglielmetti, Fabio Faber-Castell, Lukas Fink, Raphael N. Vuille-dit-Bille
'Raphael N. Vuille-dit-Bille'] Background Statistic scripts are often made by mathematicians and cryptic for clinicians or non-mathematician scientists. Nevertheless, almost all research projects necessitate the application of some statistical tests or at least an understanding thereof. The present review aims on…
Nicole M. White, Thirunavukarasu Balasubramaniam, Richi Nayak, Adrian G. Barnett + 1 more
'Adrian G. Barnett' 'Benedikt Ley'] Appropriate descriptions of statistical methods are essential for evaluating research quality and reproducibility. Despite continued efforts to improve reporting in publications, inadequate descriptions of statistical methods persist. At times, reading statistical methods sections…
Authors not listed
We developed OpenStats, a user-friendly web application that brings the power of the R language to researchers through a high-level interface and broad support for statistical methods such as t-tests and ANOVA. OpenStats was integrated into our electronic lab notebook Chemotion ELN via its third-party API, enabling…
Charles W. Champ, Andrew V. Sills
A case is made that researchers are interested in studying processes. Often the inferences they are interested in making are about the process and its associated population. On other occasions, a researcher may be interested in making an inference about the collection of individuals the process has generated. We will…
David J. Hand
We examine the role of trustworthiness and trust in statistical inference, arguing that it is the extent of trustworthiness in inferential statistical tools which enables trust in the conclusions. Certain tools, such as the p‐value and significance test, have recently come under renewed criticism, with some arguing…
Estevao Alves da Silva
About 15 years ago, statistical softwares were rarely employed to conduct statistical analyses, and biologists/ecologists used roughly 100 statistical procedures in research. This number has likely increased substantially with the development of new analytical methods and the widespread adoption of computers and…
Célia Landmann Szwarcwald
This article aimed to present an overview of national health surveys, sampling techniques, and components of statistical analysis of data collected using complex sampling designs. Briefly, surveys aimed at assessing the nutritional status of Brazilians and maternal and child health care were described. Surveys aimed at…
Edsaúl Emilio Pérez-Guerrero, Miryam Rosario Guillén-Medina, Fabiola Márquez-Sandoval, José María Vera-Cruz + 5 more
'Fabiola Márquez-Sandoval' 'José María Vera-Cruz' 'Martha Patricia Gallegos-Arreola' 'Manuel Alejandro Rico-Méndez' 'José Alonso Aguilar-Velázquez' 'Itzae Adonai Gutiérrez-Hurtado' 'Gabriel Chodick'] Epidemiological studies are essential in medicine and public health as they help identify risk factors and causes of…
Yishan Wang, Chenxuan Zang, Ziyi Li, Charles C. Guo + 2 more
Spatial transcriptomics (ST) provides unprecedented insights into gene expression patterns while retaining spatial context, making it a valuable tool for understanding complex tissue architectures, such as those found in cancers. Seurat, by far the most popular tool for analyzing ST data, uses the Wilcoxon rank-sum…
Carol Ting
It has long been a puzzle why, despite sustained reform efforts, many applied scientific fields remain dominated by Null Hypothesis Significance Testing (NHST), a framework that dichotomizes study results and privileges "statistically significant" findings. This paper examines that puzzle by situating the development…
Kalina Hristova, William C. Wimley
We present a simple, spreadsheet-based method to determine the statistical significance of the difference between any two arbitrary curves. This modified Chi-squared approach includes a critical correction for the deviation from normality in measurements with small sample size, which are typical in biomedical sciences.…
Thomas B. Bertelsen, Asle Hoffart, Sondre Sverd Rekdal, Rune Zahl-Olsen
Background: Statistical methods are a cornerstone of research in clinical psychology and are used in clinical trials and reviews to determine the best available evidence. The most widespread statistical framework, frequentist statistics, is often misunderstood and misused. Even when properly applied, this framework can…
Authors not listed
The rapid growth of worldwide computing power has transformed in silico chemistry into a discipline that is integrated into the daily work of many chemists. Nowadays, researchers find it increasingly straightforward to predict a wide range of molecular properties and chemi- cal processes at reasonable computational…
Christopher Brydges, Xiaoyu Che, W. Ian Lipkin, Oliver Fiehn
Univariate analyses of metabolomics data currently follow a frequentist approach, using p-values to reject a null-hypothesis. However, the usability of p-values is plagued by many misconceptions and inherent pitfalls. We here propose the use of Bayesian statistics to quantify evidence supporting different hypotheses…
Wolfgang Rolke
Univariate Data Authors: ['Wolfgang Rolke'] We present the results of a large number of simulation studies regarding the power of various goodnessof-fit as well as nonparametric two-sample tests for univariate data. This includes both continuous and discrete data. In general no single method can be relied upon to…
Salamiah A. Jamal
Environment Authors: ['Salamiah A. Jamal'] Abstract—This study utilises the R programming language for statistical data analysis to understand Tourism dynamics in Poland. It focuses on methods for data visualisation, multivariate statistics, and hypothesis testing. To investigate the expenditure behavior of tourist…
Paul Vos
This paper examines the foundational concept of random variables in probability theory and statistical inference, demonstrating that their mathematical definition requires no reference to randomization or hypothetical repeated sampling. We show how measure-theoretic probability provides a framework for modeling…
Wenqisi Pan, Zeyu Lu, Wei Jiang, Johan Lim + 2 more
In meta-analyses of continuous outcomes, the sample mean and standard deviation (SD) are essential for synthesizing effect sizes across studies. However, clinical studies frequently report alternative summary statistics, such as the median, quartiles, and range. To enable inclusion of such studies, various methods have…
Zinan Lu, Jonathan Anns, Yishan Mai, Rou Zhang + 10 more
Data analysis in experimental science mainly relies on null-hypothesis significance testing, despite its well-known limitations. A powerful alternative is estimation statistics, which focuses on effect-size quantification. However, current estimation tools struggle with the complex, multi-group comparisons common in…
Mingrui Xu, Zhiming Li, Keyi Mou, Kalakani Mohammad Shuaib + 1 more
'Richard D. Gill'] Gwet’s first-order agreement coefficient (AC $_{1}$) is widely used to assess the agreement between raters. This paper proposes several asymptotic statistics for a homogeneity test of stratified AC $_{1}$ in large sample sizes. These statistics may have unsatisfactory performance, especially for…
Conrad Hübler
A novel application to determine stability constants from supramolecular titration experiments is presented. The focus lies on NMR titration and ITC experiments for pure 1:1 systems, as well as mixed 2:1/1:1, 1:1/1:2 and 2:1/1:1/1:2 systems. SupraFit provides global and local fitting and a global search tool.…
Authors not listed
Plastic mechanical recycling is the conventional technological step towards circularity. In such aspects, complex mixtures of polyolefin blends are often fed into mechanical recycling systems, resulting in moulded products with uncertain quality. To add to the difficulty of heterogeneous feedstocks, the testing of…
Van N. T. La, Stanley Nicholson, Amna Haneef, Lulu Kang + 1 more
Some data are just underappreciated. Maybe they look different or come from a different background than most other data. Maybe they don't fit neatly into common notions of what data on a ``curve'' should look like. Whatever the case, they are pigeonholed into a restricted role that limits their contributions. But if…
Yuji Kaiya, Ryo Tamura, Koji Tsuda
Kinetic models are widely used in simulating the relationship between the input space and the outcome space of a chemical process. Ignoring the computational cost, complete profiling, i.e., performing simulation at all grid points in the input space, would be the best way to understand the model, because it provides us…