21 papers · ranked by Valyu relevance
Benjamin M Peter
Many questions about human genetic history can be addressed by examining the patterns of shared genetic variation between sets of populations. A useful methodological framework for this purpose are F-statistics, that measure shared genetic drift between sets of two, three and four populations, and can be used to test…
Lauren E. Frankel, Cécile Ané
F-statistics are commonly used to assess hybridization, admixture or introgression between populations or deeper evolutionary lineages. Their fast calculation from allele frequencies allows for rapid downstream admixture graph inference. One frequently overlooked assumption of the f_4_-test is a constant substitution…
Gözde Atağ, Mehmet Somel
Patterson’s f-statistics are among the most heavily utilized tools for analysing genome-wide allele frequency data for demographic inference. Beyond studying admixture, f_3_ and f_4_ statistics are also used for clustering populations to identify groups with similar histories. However, previous studies have noted an…
Mehreen R. Mughal, Michael DeGiorgio
The Patterson F- and D-statistics are commonly-used measures for quantifying population relationships and for testing hypotheses about demographic history. These statistics make use of allele frequency information across populations to infer different aspects of population history, such as population structure and…
Kalle Leppälä
The f_4_-ratio estimation is a popular technique for estimating gene flow proportion as a ratio between two f_4_-statistics. We point out the underappreciated fact that the root is not important in f_4_-ratio estimation, making the technique highly versatile. We study robustness and the bidirectional case. We introduce…
Samuele Soraggi, Carsten Wiuf, Anders Albrechtsen
The detection of ancient gene flow between human populations is an important issue in population genetics. A common tool for detecting ancient admixture events is the D-statistic. The D-statistic is based on the hypothesis of a genetic relationship that involves four populations, whose correctness is assessed by…
Yire Shin, Piyapatr Busababodhin, Jeong‐Soo Park
The generalized extreme value distribution (GEVD) has been widely used to model the extreme events in many areas. It is however limited to using only block maxima, which motivated to model the GEVD dealing with r-largest order statistics (rGEVD). The rGEVD which uses more than one extreme per block can significantly…
María J. Blanca, Rafael Alarcón, Roser Bono, Jaume Arnau + 2 more
Introduction Data analysis of split-plot designs requires normality, homogeneity of variance, and multisample sphericity. Adjusted F-tests and the bootstrap-F procedure have been proposed as alternatives to the F-statistic when these assumptions are violated. However, simulation studies have not identified the…
Korbinian Strimmer
Background False discovery rate (FDR) methods play an important role in analyzing high-dimensional data. There are two types of FDR, tail area-based FDR and local FDR, as well as numerous statistical algorithms for estimating or controlling FDR. These differ in terms of underlying test statistics and procedures…
María J. Blanca, Jaume Arnau, F. Javier García-Castro, Rafael Alarcón + 1 more
The aim of this study was to analyze the effect of sphericity violation, irrespective of non-normality, on the F-statistic, F-GG, and F-HF, examining the performance of each in terms of type I error and power under a variety of conditions that may be encountered in real research practice. Normal data were generated…
R. M. Dünki, M. Dressel
Reducing a feature vector to an optimized dimensionality is a common problem in biomedical signal analysis. This analysis retrieves the characteristics of the time series and its associated measures with an adequate methodology followed by an appropriate statistical assessment of these measures (e.g., spectral power or…
David S. Lee, Justin McCrary, Marcelo J. Moreira, Jack Porter
In the single IV model, current practice relies on the first-stage F exceeding some threshold (e.g., 10) as a criterion for trusting t-ratio inferences, even though this yields an anti-conservative test. We show that a true 5 percent test instead requires an F greater than 104.7. Maintaining 10 as a threshold requires…
Authors not listed
The precision of thermodynamic modeling for ionic liquid (IL)–solute systems is fundamentally reliant on the quality of experimental data. However, prevalent databases such as ILThermo frequently exhibit conflicting measurements for the same systems under identical temperature and pressure conditions. These disparities…
Megan H. Murray, Jeffrey D. Blume
False discovery rates (FDR) are an essential component of statistical inference, representing the propensity for an observed result to be mistaken. FDR estimates should accompany observed results to help the user contextualize the relevance and potential impact of findings. This paper introduces a new user-friendly R…
Henk van Elst
These lecture notes were written with the aim to provide an accessible though technically solid introduction to the logic of systematical analyses of statistical data to undergraduate and to postgraduate students, in particular in the Social Sciences and in Economics. They may also serve as a general reference for the…
Avinash Vaidheeswaran, Cheng Li, Huda Ashfaq, Xiongjun Wu + 2 more
Bubbling fluidization experiments were performed in three cylindrical columns having internal diameters of 2.5, 4 and 6 inches. Glass particles having a sauter mean diameter of 332 microns were used, and the operating conditions were held constant in all the units. Statistics of differential pressure and interface…
Joong‐Ho Won, Johan Lim, Donghyeon Yu, Byung‐Soo Kim + 1 more
The advance of modern high-throughput technologies in many scientific disciplines such as genomics and brain imaging has dramatically increased both the size and the dimension of the data and made data analysis a major challenge. In particular, it is often required to test thousands or millions of hypotheses…
Olivia McGough, Daniela Witten, Daniel Kessler
Suppose that a data analyst wishes to report the results of a least squares linear regression only if the overall null hypothesis—H 1:p 0 : β1 = β2 = . . . = βp = 0—is rejected. This practice, which we refer to as F-screening (since the overall null hypothesis is typically tested using an F-statistic), is in fact…
Riko Kelter
Hypothesis testing is a central statistical method in psychology and the cognitive sciences. However, the problems of null hypothesis significance testing (NHST) and p-values have been debated widely, but few attractive alternatives exist. This article introduces the fbst R package, which implements the Full Bayesian…
Merga Abdissa Aga, Fucai Lin
We propose the Fréchet-Power Function (FPF) distribution, a novel two-parameter model that combines the bounded support of the Power Function distribution with the heavy-tailed flexibility of a Fréchet-type generator. This combination enables the FPF to capture complex features such as skewness, heavy tails, and…
Authors not listed
To rationalize and improve the performance of high energy density electrode materials for all-solid-state fluoride ion batteries and abstract the increasing demands of energy storage, it is important to gain a fundamental understanding of structure/phase evolution during cell operation, which is closely correlated to…