15 papers · ranked by Valyu relevance
M Harshvardhan, Pritam Ranjan
Pearson correlation is the default measure of association in most statistical software, yet it is only appropriate for pairs of continuous variables with a linear relationship. When variables are binary, ordinal, or categorical, specialized methods (e.g., point-biserial, polychoric, tetrachoric, and Cramér's~$V$) may…
Alessandro Fontana
Methods to find correlation between variables are of interest to many disciplines, including statistics, machine learning, (big) data mining and neurosciences. Parameters that measure correlation between two variables are of limited utility when used with multiple variables. In this work, I propose a simple criterion…
Rami Mahdi
In this paper, a robust non-parametric measure of statistical dependence, or correlation, between two random variables is presented. The proposed coefficient is a permutation-like statistic that quantifies how much the observed sample Sn : {(Xi , Yi), i = 1 . . . n} is discriminable from the permutated sample Sˆ n×n …
Peter Richmond, Bertrand M. Roehner
Cheating in examinations is acknowledged by an increasing number of organizations to be widespread. We examine two different approaches to assess their effectiveness at detecting anomalous results, suggestive of collusion, using data taken from a number of multiple-choice examinations organized by the UK Radio…
Adam Lehavi, Seong‐Tae Kim
—In the realm of cybersecurity, intrusion detection systems (IDS) detect and prevent attacks based on collected computer and network data. In recent research, IDS models have been constructed using machine learning (ML) and deep learning (DL) methods such as Random Forest (RF) and deep neural networks (DNN). Feature…
Yue Gong, Raul Castro Fernandez
As hypothesis generation becomes increasingly automated, a new bottleneck has emerged: hypothesis assessment. Modern systems can surface thousands of statistical relationships—correlations, trends, causal links—but offer little guidance on which ones are novel, non-trivial, or worthy of expert attention. In this work…
Ze Lin, Bo Li, Jinyao Shen
Pearson's correlation coefficient is commonly used as a single-number summary of association between two responses. In many applications, however, the strength of association is itself heterogeneous and may vary with demographic, biological, experimental, or environmental covariates. The regcorr package implements…
Fatih Dikbaş
Correlation remains to be one of the most widely used statistical tools for assessing the strength of relationships between data series. This paper presents a novel compositional correlation method for detecting linear and nonlinear relationships by considering the averages of all parts of all possible compositions of…
Mustafa Attallah
Pearson's correlation to select predictor variables for linear models Authors: ['Mustafa Attallah'] This article examines the limitations of Pearson's correlation in selecting predictor variables for linear models. Using mtcars and iris datasets from R, this paper demonstrates the limitation of this correlation measure…
Ming Luo, S. Radhakrishnan, Sagar Kamarthi
Study on Surface Roughness in Finish Turning Authors: ['Ming Luo' 'S. Radhakrishnan' 'Sagar Kamarthi'] | 1 | INTRODUCTION | 3 | | --- | --- | --- | | 2 | THE CONCEPT OF CORRELATION BETWEEN TWO VARIABLES | 3 | | | 2.1 Pearson Correlation Coefficient | 8 | | | 2.2 Spearman's Rank Correlation Coefficient | 9 | | | 2.3…
Rudy Arthur
Comparing spatial data sets is a ubiquitous task in data analysis, however the presence of spatial autocorrelation means that standard estimates of variance will be wrong and tend to over-estimate the statistical significance of correlations and other observations. While there are a number of existing approaches to…
Xiatong Cai, Guangpeng Pei, Yuen Zhu, Donggang Guo + 1 more
With the establishment of global biological monitor network and development of remote sensing technology, data won't be a limitation, but the variance brought by spatial heterogeneous and fractal will influence correlation coefficient significantly with the enlarged sample scale. Those impede us to find more intrinsic…
Chang Liu, Yi Xu
In machine learning and pattern recognition, feature selection has been a hot topic in the literature. Unsupervised feature selection is challenging due to the loss of labels which would supply the related information.How to define an appropriate metric is the key for feature selection. We propose a filter method for…
Zenon Gniazdowski
The article investigates the possibility of measuring the strength of a linear correlation relationship between nominal data and numerical data. Correlation coefficients for variables coded with real numbers as well as for variables coded with complex numbers were studied. For variables coded with real numbers…
Trevor J. Hefley, Kristin Broms, Brian M. Brost, Frances E. Buderman + 5 more
'Shannon L. Kay' 'Henry R. Scharf' 'John Tipton' 'Perry J. Williams' 'Mevin B. Hooten'] Analyzing ecological data often requires modeling the autocorrelation created by spatial and temporal processes. Many of the statistical methods used to account for autocorrelation can be viewed as regression models that include…