19 papers · ranked by Valyu relevance
Rahi Jain, Wei Xu
Feature selection is important in high dimensional data analysis. The wrapper approach is one of the ways to perform feature selection, but it is computationally intensive as it builds and evaluates models of multiple subsets of features. The existing wrapper approaches primarily focus on shortening the path to find an…
Francisco Aragón-Royón, Alfonso Jiménez-Vílchez, Antonio Araúzo-Azofra, José M. Benítez
'Antonio Araúzo-Azofra' 'José M. Benítez'] Feature Selection (FS) is a key task in Machine Learning. It consists in selecting a number of relevant variables for the model construction or data analysis. We present the R package, FSinR, which implements a variety of widely known filter and wrapper methods, as well as…
Reiner Jedermann, Walter Lang, Leopoldo Angrisani, Domenico Accardo
Analog sensors often require complex mathematical models for data analysis. Digital twins (DTs) provide platforms to display sensor data in real time but still lack generic solutions regarding how mathematical models and algorithms can be integrated. Based on previous tests for monitoring and predicting banana fruit…
Yanxiong Peng, Wenyuan Li, Ying Liu
Microarrays allow researchers to monitor the gene expression patterns for tens of thousands of genes across a wide range of cellular responses, phenotype and conditions. Selecting a small subset of discriminate genes from thousands of genes is important for accurate classification of diseases and phenotypes. Many…
Fujun Wang, Xing Wang, Seyedali Mirjalili
Feature selection is an important task in big data analysis and information retrieval processing. It reduces the number of features by removing noise, extraneous data. In this paper, one feature subset selection algorithm based on damping oscillation theory and support vector machine classifier is proposed. This…
Dibri Nsofor, Ben Greenman
Gradually-typed languages feature a dynamic type that supports implicit coercions, greatly weakening the type system but making types easier to adopt. Understanding how developers use this dynamic type is a critical question for the design of useful and usable type systems. This paper reports on an in-progress corpus…
Yukiyoshi Iwata, Shun Hasegawa, Kento Kawaharazuka, Kei Okada + 1 more
'Masayuki Inaba'] Flexible object manipulation of paper and cloth is a major research challenge in robot manipulation. Although there have been efforts to develop hardware that enables specific actions and to realize a single action of paper folding using sim-to-real and learning, there have been few proposals for…
Grace Yee Lin Ng, Shing Chiang Tan, Chia Sui Ong, Guanghui Liu
Cell type identification is one of the fundamental tasks in single-cell RNA sequencing (scRNA-seq) studies. It is a key step to facilitate downstream interpretations such as differential expression, trajectory inference, etc. scRNA-seq data contains technical variations that could affect the interpretation of the cell…
Yongtao Shi, Yuefeng Zheng, Xiaotong Bai, Elnaz Pashaei
Recently, hybrid feature selection methods have demonstrated excellent performance on high-dimensional data, but many of these methods tend to yield relatively homogeneous feature subsets. To address this, we propose a novel hybrid feature selection algorithm called the Hybrid Multiple Filter-Wrapper algorithm. This…
K. Tsolakidis, Artu Breuer, S. Bender, Stavroula Margaritaki + 3 more
Advances in fluorescence microscopy have dramatically expanded the range of biological questions that can be addressed, enabling quantitative observations of molecular interactions and cellular dynamics with unprecedented spatial and temporal resolution. However, the growing complexity of imaging data has outpaced our…
Farzane Karami, Olaf Owe, Gerardo Schneider
This paper introduces a run-time mechanism for preventing leakage of secure information in distributed systems. We consider a general concurrency language model, where concurrent objects interact by asynchronous method calls and futures. The aim is to prevent leakage of confidential information to low-level viewers.…
Dimitar Georgiev, Simon Vilms Pedersen, Ruoxiao Xie, Álvaro Fernández-Galiana + 2 more
Raman spectroscopy is a non-destructive and label-free chemical analysis technique, which plays a key role in the analysis and discovery cycle of various branches of science. Nonetheless, progress in Raman spectroscopic analysis is still impeded by the lack of software, methodological and data standardisation, and the…
Abhinav Sharma, Davi Josué Marcon, Johannes Loubser, Karla Valéria Batista Lima + 2 more
The MTBseq pipeline, published in 2018, was designed to address bioinformatics challenges in tuberculosis research using whole-genome sequencing data. It was the first publicly available pipeline on Github to perform full analysis of whole-genome sequencing (WGS) data for Mycobacterium tuberculosis encompassing quality…
Dhruti Parikh, Angli Xue, Hsiao-Chi Liao, Claire Wishart + 7 more
Over the past decade, there has been an explosion in the characterisation and discovery of cell populations using single-cell technologies. Single-cell multi-omics data, particularly those incorporating gene and protein expression, are increasingly commonplace and can lead to more refined characterisation of cell…
Tianwei Cao, Qianqian Xu, Zhiyong Yang, Qingming Huang
—Click-through rate (CTR) prediction, whose goal is to predict the probability of the user to click on an item, has become increasingly significant in the recommender systems. Recently, some deep learning models with the ability to automatically extract the user interest from his/her behaviors have achieved great…
Soreangsey Kiv, Yves Wautelet, Samedi Heng, Manuel Kolp
Tools like Prot´eg´e support the creation and edition of one or more ontologies in a single workspace. They nevertheless require a user to be familiar with this kind of abstractions and their supporting techniques such as a reasoner and SPARQL queries. This paper presents a step-by-step implementation of a…
Authors not listed
The analysis of nonadiabatic molecular dynamics (NAMD) data presents significant challenges due to its high dimensionality and complexity. To address these issues, we introduce ULaMDyn, a Python-based, open-source package designed to automate the unsupervised analysis of large datasets generated by NAMD simulations.…
Authors not listed
High-level quantum mechanical (QM) simulations provide accurate electronic information of chemical systems but scale unfavourably with system size, making calculations of applied systems challenging. Hierarchical quantum mechanics in quantum mechanics embedding (QM/QM) addresses this issue by localising the highly…
Zachary Sierzega, Jeff Wereszczynski, Chris Prior
We introduce the Writhe Application Software Package (WASP) which can be used to characterise the topology of ribbon structures, the underlying mathematical model of DNA, Biopolymers, superfluid vorticies, elastic ropes and magnetic flux ropes. This characterisation is achieved by the general twist-writhe decomposition…