Search · four archives
Search · four archives
26 papers · ranked by Valyu relevance
Rui Portocarrero Sarmento, Vera L. Costa
The use of statistical software in academia and enterprises has been evolving over the last years. More often than not, students, professors, workers and users in general have all had, at some point, exposure to statistical software. Sometimes, difficulties are felt when dealing with such type of software. Very few…
Shraddha Parab, Supriya Bhalerao
Statistical tests are mathematical tools for analyzing quantitative data generated in a research study. The multitude of statistical tests makes a researcher difficult to remember which statistical test to use in which condition. There are various points which one needs to ponder upon while choosing a statistical test.…
Henk van Elst
These lecture notes were written with the aim to provide an accessible though technically solid introduction to the logic of systematical analyses of statistical data to undergraduate and to postgraduate students, in particular in the Social Sciences and in Economics. They may also serve as a general reference for the…
Laura C. Guglielmetti, Fabio Faber-Castell, Lukas Fink, Raphael N. Vuille-dit-Bille
'Raphael N. Vuille-dit-Bille'] Background Statistic scripts are often made by mathematicians and cryptic for clinicians or non-mathematician scientists. Nevertheless, almost all research projects necessitate the application of some statistical tests or at least an understanding thereof. The present review aims on…
Robert E. Kass, Brian S. Caffo, Marie Davidian, Xiao-Li Meng + 3 more
Grappling with variability is central to the discipline of statistics. Variability comes in many forms. In some cases variability is good, because we need variability in predictors to explain variability in outcomes. For example, to determine if smoking is associated with lung cancer, we need variability in smoking…
Rising Odegua
A large amount of data is produced every second from modern information systems such as mobile devices, the world wide web, Internet of Things, social media, and so on. Analysis and mining of these massive data require a lot of advanced tools and techniques. Therefore, big data analytics and mining is currently an…
Xuming He, David Madigan, Bin Yu, Jon Wellner
| EXECUTIVE SUMMARY | 4 | | --- | --- | | SECTION 1: ROLE/VALUE OF STATISTICS AND DATA SCIENCE | 6 | | SECTION 2: CHALLENGES IN SCIENTIFIC AND SOCIAL APPLICATIONS | 10 | | SECTION 3: FOUNDATIONAL RESEARCH | 16 | | SECTION 4: PROFESSIONAL CULTURE & COMMUNITY RESPONSIBILITIES | 20 | | SECTION 5: DOCTORAL EDUCATION | 23 |…
Authors not listed
We developed OpenStats, a user-friendly web application that brings the power of the R language to researchers through a high-level interface and broad support for statistical methods such as t-tests and ANOVA. OpenStats was integrated into our electronic lab notebook Chemotion ELN via its third-party API, enabling…
Atef F. Hashem, M. A. Abdelkawy, Abdisalam Hassan Muse, Haitham M. Yousof
'Haitham M. Yousof'] The current study introduces and examines copula-coupled probability distributions. It explains their mathematical features and shows how they work with real datasets. Researchers, statisticians, and practitioners can use this study’s findings to build models that capture complex multivariate data…
Won-Seok Choi
A testing method to identify statistically significant differences by comparing the significance level and the probability value based on the Null Hypothesis Significance Test (NHST) has been used in food research. However, problems with this testing method have been discussed. Several alternatives to the NHST and the…
Michael B Brimacombe
The development of graduate education in biostatistics and medical statistics is discussed in the context of training within a medical center setting. The need for medical researchers to employ a wide variety of statistical designs in clinical, genetic, basic science and translational settings justifies the ongoing…
Joses Ho, Tayfun Tumkaya, Sameer Aryal, Hyungwon Choi + 1 more
As interventions always have effects, the analyst’s appropriate task is to quantify the effect size and assess its precision. For this purpose, the difference axis is more productively used to compare the two means, represented as the difference of means, Δ (Figure 1). Around Δ, the analyst plots an indicator of…
Wenqisi Pan, Zeyu Lu, Wei Jiang, Johan Lim + 2 more
In meta-analyses of continuous outcomes, the sample mean and standard deviation (SD) are essential for synthesizing effect sizes across studies. However, clinical studies frequently report alternative summary statistics, such as the median, quartiles, and range. To enable inclusion of such studies, various methods have…
Branimir K. Hackenberger
Meta-analysis is a statistical tool that allows the analysis of results from various scientific studies, which are often not performed in the same place or using the same method. The data used in meta-analysis may be proprietary or may be obtained from literature or various databases. Meta-analysis is a crucial part of…
Jean-Bernard Martens
While the applied psychology community relies on statistics to assist drawing conclusions from quantitative data, the methods being used mostly today do not reflect several of the advances in statistics that have been realized over the past decades. We show in this paper how a number of issues with how statistical…
Authors not listed
Plastic mechanical recycling is the conventional technological step towards circularity. In such aspects, complex mixtures of polyolefin blends are often fed into mechanical recycling systems, resulting in moulded products with uncertain quality. To add to the difficulty of heterogeneous feedstocks, the testing of…
Danny Salem, Anuradha Surendra, Graeme SV McDowell, Miroslava Čuperlović-Culf
Unsupervised data projection for the determination of trends in the data, visualization of multidimensional data in a reduced dimension space or feature space reduction through combination of data is a major step in data mining. Methods such as Principal Component Analysis or t-Distribution Stochastic Neighbor…
Ringyao Jajo, Shivani Kansal, Sonia Balyan, Saurabh Raghuvanshi
Data visualisation technique has greatly improved as technology has advanced. While representing the data through graph, it has made the underlying data structure become more transparent and interpretable. However, the informational scope of freely available generic visualisation tools is still limited since they only…
Rafael Ramos, Igor Nelson
Gene expression regulates several complex traits observed. In this study, datasets comprising of transcriptome information and clinical traits regarding fat composition and vitals were analyzed via several statistical methods in order to find relations between genes and clinical outcomes. Biological big data is diverse…
Roger W. Hoerl, Ronald D. Snee
Several authors, including the American Statistical Association (ASA), have noted the challenges facing statisticians when attacking large, complex and unstructured problems, as opposed to well-defined textbook problems. Clearly, the standard paradigm of selecting the one "correct" statistical method for such problems…
Stephanie C. Hicks, Roger D. Peng
The data revolution has led to an increased interest in the practice of data analysis. For a given problem, there can be significant or subtle differences in how a data analyst constructs or creates a data analysis, including differences in the choice of methods, tooling, and workflow. In addition, data analysts can…
Lucy D’Agostino McGowan, Roger D. Peng, Stephanie C. Hicks
The data science revolution has led to an increased interest in the practice of data analysis. While much has been written about statistical thinking, a complementary form of thinking that appears in the practice of data analysis is design thinking – the problem-solving process to understand the people for whom a…
Linda A. Auker, Erika L. Barthelmess
Ecology requires training in data management and analysis. In this paper, we present data from the last 10 years demonstrating the increase in the use of R, an open-source programming environment, in ecology and its prevalence as a required skill in job descriptions. Because of its transparent and flexible nature, R is…
Ramsey Issa, Robert Sorenson, Taylor D. Sparks
The discovery of new dental materials is typically a slow process due to high-dimensionality of the formulation space as well as the multiple competing objectives which must be optimized for a given application. Here, we lay out a strategy using active learning and Bayesian optimization that has led to the discovery of…
Amelia Carolina Sparavigna
Previous studies (Sparavigna, 2023) have demonstrated the Tsallis q-Gaussian functions suitable for the analysis of Raman spectra. Here we consider two asymmetric forms of them, with the aim of further improving the fitting of Raman spectra. The first asymmetric form, that we are here proposing for the first time, is a…
Rebecca Brunk, Kriti Shukla, Bryant Hutson, Yue Wang + 7 more
Genomic sequencing and other big biological data is unquestionably of paramount value, however the success in recruiting highly skilled individuals with diverse backgrounds has been limited. A main reason for this deficiency could be due to the lack of educational resources and early exposure to the field. With the…