11 papers · ranked by Valyu relevance
George Stoica, Mihaela Breaban, Vlad Ștefan Barbu
> Abstract. Using additional training data is known to improve the results, especially for medical image 3D segmentation where there is a lack of training material and the model needs to generalize well from few available data. However, the new data could have been acquired using other instruments and preprocessed such…
Étienne Grandjean, Louis Jachiet
| 1 | Introduction | 2 | |…
Victor Y. Pan, Liang Zhao
- If we append sufficiently many standard Gaussian random rows or columns to any matrix A, such that ||A|| = 1, then the augmented matrix has full rank with probability 1 and is well-conditioned with a probability close to 1, even if the matrix A is rank deficient or ill-conditioned. - We specify and prove these…
Lydia R Lucchesi, Petra Kuhnert, Jenny Davis, Lexing Xie
Data preprocessing is a crucial stage in the data analysis pipeline, with both technical and social aspects to consider. Yet, the attention it receives is often lacking in research practice and dissemination. We present the Smallset Timeline, a visualisation to help reflect on and communicate data preprocessing…
Jia Wei, Xingjun Zhang, Witold Pedrycz, Longxiang Wang + 1 more
CPU, GPU and CSD Authors: ['Jia Wei' 'Xingjun Zhang' 'Witold Pedrycz' 'Longxiang Wang' 'Jie Zhao'] Abstract—As accelerator computation speeds increase and the number of single compute node accelerators increases, data reading and preprocessing gradually become a bottleneck for deep learning. Most existing data…
Francesco Taurone, Daniel E. Lucani, Marcell Fehér, Qi Zhang
> Abstract. Data compression algorithms typically rely on identifying repeated sequences of symbols from the original data to provide a compact representation of the same information, while maintaining the ability to recover the original data from the compressed sequence. Using data transformations prior to the…
Ovi Paul
—NumtaDB is by far the largest data-set collection for handwritten digits in Bengali. This is a diverse dataset containing more than 85000 images. But this diversity also makes this dataset very difficult to work with. The goal of this paper is to find the benchmark for pre-processed images which gives good accuracy on…
Alejandro Murillo-González, José David Ortega Pabón, Juan Guillermo Paniagua, Olga Lucía Quintero Montoya
'Juan Guillermo Paniagua' 'Olga Lucía Quintero Montoya'] An image preprocessing methodology based on Fourier analysis together with the Laguerre-Gauss Spatial Filter is proposed. This is an alternative to obtain features from aerial images that reduces the feature space significantly, preserving enough information for…
Guzzetta, Gianluca
In this paper, we present a comprehensive study and analysis of the Chan Vese algorithm for image segmentation, using a discretized scheme obtained from the empirical study of Chan-Vese model's functional energy and its partial differential equation based on its level set function. The proof to such result is shown…
Huiyu Zhao, Jiahao Yan, Catherine Dawson, Haitao Yang + 1 more
We present AngstromPro, a versatile, modular and open-source software built on Python for managing, visualizing and analyzing large datasets acquired via Scanning Tunneling Microscopes (STM). Its robust architecture features a top-level module that manages a Global Variables List and a sub-modules List. Each…
Authors not listed
—Many scientific and engineering problems involving multi-physics span a wide range of scales. Understanding the interactions across these scales is essential for fully comprehending such complex problems. However, visualizing multivariate, multiscale data within an integrated view where correlations across space…