14 papers · ranked by Valyu relevance
Timothy Crawley, Arthur G. Palmer III
The ability to make robust inferences about the dynamics of biological macromolecules using NMR spectroscopy depends heavily on the application of appropriate theoretical models for nuclear spin relaxation. Data analysis for NMR laboratory-frame relaxation experiments typically involves selecting one of several…
Reem Salman, Ayman Alzaatreh, Hana Sulieman, Shaimaa Faisal + 1 more
'Mohamed Medhat Gaber'] In the past decade, big data has become increasingly prevalent in a large number of applications. As a result, datasets suffering from noise and redundancy issues have necessitated the use of feature selection across multiple domains. However, a common concern in feature selection is that…
TT Vu, UM Braga-Neto
There has been considerable interest recently in the application of bagging in the classification of both gene-expression data and protein-abundance mass spectrometry data. The approach is often justified by the improvement it produces on the performance of unstable, overfitting classification rules under small-sample…
Kęstutis Baltakys, Juho Kanniainen, Frank Emmert-Streib
Multilayer networks are attracting growing attention in many fields, including finance. In this paper, we develop a new tractable procedure for multilayer aggregation based on statistical validation, which we apply to investor networks. Moreover, we propose two other improvements to their analysis: transaction…
Susmita Datta, Vasyl Pihur, Somnath Datta
Background Generally speaking, different classifiers tend to work well for certain types of data and conversely, it is usually not known a priori which algorithm will be optimal in any given classification application. In addition, for most classification problems, selecting the best performing classification algorithm…
Leonard Roth, Matthias Studer, Emilie Zuercher, Isabelle Peytremann-Bridevaux
'Isabelle Peytremann-Bridevaux'] Background In standard Sequence Analysis, similar trajectories are clustered together to create a typology of trajectories, which is then often used to evaluate the association between sequence patterns and covariates inside regression models. The sampling uncertainty, which affects…
Juan Botella, Desirée Blázquez, Manuel Suero, James F. Juola
Assessing significant change (or reliable change) in a person often involve comparing the responses of that person in two administrations of a test or scale. Several procedures have been proposed to determine if a difference between two observed scores is statistically significant or rather is within the range of mere…
Cheng-Jian Xu, Huub CJ Hoefsloot, Age K Smilde
Background High-throughput functional genomics technologies generate large amount of data with hundreds or thousands of measurements per sample. The number of sample is usually much smaller in the order of ten or hundred. This poses statistical challenges and calls for appropriate solutions for the analysis of this…
Elias Chaibub Neto, Kay Hamacher
In this paper we propose a vectorized implementation of the non-parametric bootstrap for statistics based on sample moments. Basically, we adopt the multinomial sampling formulation of the non-parametric bootstrap, and compute bootstrap replications of sample moment statistics by simply weighting the observed data…
Kwanghee Jung, Jaehoon Lee, Vibhuti Gupta, Gyeongcheol Cho
Generalized structured component analysis (GSCA) is a theoretically well-founded approach to component-based structural equation modeling (SEM). This approach utilizes the bootstrap method to estimate the confidence intervals of its parameter estimates without recourse to distributional assumptions, such as…
Vasyl Pihur, Susmita Datta, Somnath Datta
Background Researchers in the field of bioinformatics often face a challenge of combining several ordered lists in a proper and efficient manner. Rank aggregation techniques offer a general and flexible framework that allows one to objectively perform the necessary aggregation. With the rapid growth of high-throughput…
Peter C. Austin
Background Healthcare provider profiling involves the comparison of outcomes between patients cared for by different healthcare providers. An important component of provider profiling is risk-adjustment so that providers that care for sicker patients are not unfairly penalized. One method for provider profiling entails…
Haikady N. Nagaraja, Shane Sanders, Alan D Hutson
The relationship between social choice aggregation rules and non-parametric statistical tests has been established for several cases. An outstanding, general question at this intersection is whether there exists a non-parametric test that is consistent upon aggregation of data sets (not subject to Yule-Simpson…
Michał Bałchanowski, Urszula Boryczka, Jiayi Ma
The aim of a recommender system is to suggest to the user certain products or services that most likely will interest them. Within the context of personalized recommender systems, a number of algorithms have been suggested to generate a ranking of items tailored to individual user preferences. However, these algorithms…