15 papers · ranked by Valyu relevance
Joël L. Horowitz
The bootstrap is a method for estimating the distribution of an estimator or test statistic by resampling one's data or a model estimated from the data. Under conditions that hold in a wide variety of econometric applications, the bootstrap provides approximations to distributions of statistics, coverage probabilities…
Thomas Pitschel
An algorithm is described that enables efficient deterministic approximate computation of the bootstrap distribution for any linear bootstrap method T ∗ n , alleviating the need for repeated resampling from observations (resp. input-derived data). In essence, the algorithm computes the distribution function from a…
Jared Clark, Richard L. Warr
Bootstrapping was designed to randomly resample data from a fixed sample using Monte Carlo techniques. However, the original sample itself defines a discrete distribution. Convolutional methods are well suited for discrete distributions, and we show the advantages of utilizing these techniques for bootstrapping. The…
Zhonglei Wang, Jae Kwang Kim, Liuhua Peng
Bootstrap is a useful tool for making statistical inference, but it may provide erroneous results under complex survey sampling. Most studies about bootstrap-based inference are developed under simple random sampling and stratified random sampling. In this paper, we propose a unified bootstrap method applicable to some…
Giles Hooker, Lucas Mentch
This paper examines the use of a residual bootstrap for bias correction in machine learning regression methods. Accounting for bias is an important obstacle in recent efforts to develop statistical inference for machine learning methods. We demonstrate empirically that the proposed bootstrap bias correction can lead to…
Jinyuan Chang, Peter Hall
We show that, when the double bootstrap is used to improve performance of bootstrap methods for bias correction, techniques based on using a single double-bootstrap sample for each singlebootstrap sample can be particularly effective. In particular, they produce third-order accuracy for much less computational expense…
Diego Didona, Paolo Romano
—Performance modeling typically relies on two antithetic methodologies: white box models, which exploit knowledge on system's internals and capture its dynamics using analytical approaches, and black box techniques, which infer relations among the input and output variables of a system based on the evidences gathered…
Behnam Yousefimehr, Mehdi Ghatee, Mohammad Amin Seifi, Javad Fazli + 9 more
'Sajed Tavakoli' 'Zahra Rafei' 'Shervin Ghaffari' 'Abolfazl Nikahd' 'Mahdi Razi Gandomani' 'Alireza Orouji' 'Ramtin Mahmoudi Kashani' 'Sarina Heshmati' 'Negin Sadat Mousavi'] Imbalanced data poses a significant obstacle in machine learning, as an unequal distribution of class labels often results in skewed predictions…
Christian Thiele, Gerrit Hirschfeld
"Optimal cutpoints" for binary classification tasks are often established by testing which cutpoint yields the best discrimination, for example the Youden index, in a specific sample. This results in "optimal" cutpoints that are highly variable and systematically overestimate the out-of-sample performance. To address…
С.Г. Литвинова, Mervyn J. Silvapulle
We show that the full-sample bootstrap is asymptotically valid for constructing confidence intervals for high-quantiles, tail probabilities, and other tail parameters of a univariate distribution. This resolves the doubts that have been raised about the validity of such bootstrap methods. In our extensive simulation…
Kees Jan van Garderen, Noud van Giersbergen
Mediation analysis is a form of causal inference that investigates indirect effects and causal mechanisms. Confidence intervals for indirect effects play a central role in conducting inference. The problem is non-standard leading to coverage rates that deviate considerably from their nominal level. The default…
Sandra Benítez-Peña, Rafael Blanquero, Emilio Carrizosa, Pepa Ramírez‐Cobo
'Pepa Ramírez‐Cobo'] Support vector machines (SVMs) are widely used and constitute one of the best examined and used machine learning models for two-class classification. Classification in SVM is based on a score procedure, yielding a deterministic classification rule, which can be transformed into a probabilistic rule…
Zhibing He, Yichen Qin, Ben‐Chang Shia, Yang Li
Bootstrap is commonly used as a tool for non-parametric statistical inference to estimate meaningful parameters in Variable Selection Models. However, for massive dataset that has exponential growth rate, the computation of Bootstrap Variable Selection (BootVS) can be a crucial issue. In this paper, we propose the…
Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Observable performance is commonly used to characterize biological systems. In adaptive systems, however, similar performances may arise from distinct organizations, and configurations that appear comparable at a given time may follow different longitudinal trajectories. This limitation motivates a methodological…
Shankhyajyoti De, Arabin Kumar Dey, Deepak Gauda
In this paper, we show an innovative way to construct bootstrap confidence interval of a signal estimated based on a univariate LSTM model. We take three different types of bootstrap methods for dependent set up. We prescribe some useful suggestions to select the optimal block length while performing the bootstrapping…