20 papers · ranked by Valyu relevance
Tianwei Cao, Qianqian Xu, Zhiyong Yang, Qingming Huang
—Click-through rate (CTR) prediction, whose goal is to predict the probability of the user to click on an item, has become increasingly significant in the recommender systems. Recently, some deep learning models with the ability to automatically extract the user interest from his/her behaviors have achieved great…
Rahi Jain, Wei Xu
Feature selection is important in high dimensional data analysis. The wrapper approach is one of the ways to perform feature selection, but it is computationally intensive as it builds and evaluates models of multiple subsets of features. The existing wrapper approaches primarily focus on shortening the path to find an…
Reiner Jedermann, Walter Lang, Leopoldo Angrisani, Domenico Accardo
Analog sensors often require complex mathematical models for data analysis. Digital twins (DTs) provide platforms to display sensor data in real time but still lack generic solutions regarding how mathematical models and algorithms can be integrated. Based on previous tests for monitoring and predicting banana fruit…
Hohyun Sim, Hyeonjoong Cho, Ali Shokri, Zhoulai Fu + 1 more
We present Encapsulated Substitution and Agentic Refinement on a Live Scaffold for Safe C-to-Rust Translation, a two-phase pipeline for translating real-world C projects to safe Rust. Existing approaches either produce unsafe output without memory-safety guarantees or translate functions in isolation, failing to detect…
Muhammad Umair Ali, Shaik Javeed Hussain, Amad Zafar, Muhammad Raheel Bhutta + 2 more
This study presents wrapper-based metaheuristic deep learning networks (WBM-DLNets) feature optimization algorithms for brain tumor diagnosis using magnetic resonance imaging. Herein, 16 pretrained deep learning networks are used to compute the features. Eight metaheuristic optimization algorithms, namely, the marine…
Dibri Nsofor, Ben Greenman
Gradually-typed languages feature a dynamic type that supports implicit coercions, greatly weakening the type system but making types easier to adopt. Understanding how developers use this dynamic type is a critical question for the design of useful and usable type systems. This paper reports on an in-progress corpus…
Yukiyoshi Iwata, Shun Hasegawa, Kento Kawaharazuka, Kei Okada + 1 more
'Masayuki Inaba'] Flexible object manipulation of paper and cloth is a major research challenge in robot manipulation. Although there have been efforts to develop hardware that enables specific actions and to realize a single action of paper folding using sim-to-real and learning, there have been few proposals for…
Grace Yee Lin Ng, Shing Chiang Tan, Chia Sui Ong, Guanghui Liu
Cell type identification is one of the fundamental tasks in single-cell RNA sequencing (scRNA-seq) studies. It is a key step to facilitate downstream interpretations such as differential expression, trajectory inference, etc. scRNA-seq data contains technical variations that could affect the interpretation of the cell…
Yongtao Shi, Yuefeng Zheng, Xiaotong Bai, Elnaz Pashaei
Recently, hybrid feature selection methods have demonstrated excellent performance on high-dimensional data, but many of these methods tend to yield relatively homogeneous feature subsets. To address this, we propose a novel hybrid feature selection algorithm called the Hybrid Multiple Filter-Wrapper algorithm. This…
K. Tsolakidis, Artu Breuer, S. Bender, Stavroula Margaritaki + 3 more
Advances in fluorescence microscopy have dramatically expanded the range of biological questions that can be addressed, enabling quantitative observations of molecular interactions and cellular dynamics with unprecedented spatial and temporal resolution. However, the growing complexity of imaging data has outpaced our…
Jayadev Joshi, Fabio Cumbo, Daniel Blankenberg
R is widely used in statistical computing, data analysis, and bioinformatics. A key contributor to its success in bioinformatics and computational biology is the open-source project Bioconductor. As of its latest release (3.20), the Bioconductor community offers 2,289 software packages for biomedical research…
Yingxia Li, Ulrich Mansmann, Shangming Du, Roman Hornung
Background In the last few years, multi-omics data, that is, datasets containing different types of high-dimensional molecular variables for the same samples, have become increasingly available. To date, several comparison studies focused on feature selection methods for omics data, but to our knowledge, none compared…
Juraj Dončević, Krešimir Fertalj, Mario Brčić, Agneza Krajna
This paper deals with the mediator-wrapper architecture. It is an important architectural pattern that enables a more flexible and modular architecture in opposition to monolithic architectures for data source integration systems. This paper identifies certain realistic and concrete scenarios where the mediator-wrapper…
Dimitar Georgiev, Simon Vilms Pedersen, Ruoxiao Xie, Álvaro Fernández-Galiana + 2 more
Raman spectroscopy is a non-destructive and label-free chemical analysis technique, which plays a key role in the analysis and discovery cycle of various branches of science. Nonetheless, progress in Raman spectroscopic analysis is still impeded by the lack of software, methodological and data standardisation, and the…
Michał Zawada, Mateusz Nijak, Jarosław Mac, Jan Szczepaniak + 11 more
'Stanisław Legutko' 'Julia Gościańska-Łowińska' 'Sebastian Szymczyk' 'Michał Kaźmierczak' 'Mikołaj Zwierzyński' 'Jacek Wojciechowski' 'Tomasz Szulc' 'Roman Rogacki' 'José Miguel Molina Martínez' 'Dolores Parras-Burgos' 'Daniel García Fernández-Pacheco'] Baler-wrappers are machines designed to produce high-quality…
Dhruti Parikh, Angli Xue, Hsiao-Chi Liao, Claire Wishart + 7 more
Over the past decade, there has been an explosion in the characterisation and discovery of cell populations using single-cell technologies. Single-cell multi-omics data, particularly those incorporating gene and protein expression, are increasingly commonplace and can lead to more refined characterisation of cell…
Kris Sankaran, Shuzhen Zhang, Chenab, Marina Meilă
Nonlinear dimensionality reduction methods like UMAP and t-SNE can help to organize high-dimensional genomics data into manageable low-dimensional representations, like cell types or differentiation trajectories. Such reductions can be powerful, but inevitably introduce distortion. A growing body of work has…
Soreangsey Kiv, Yves Wautelet, Samedi Heng, Manuel Kolp
Tools like Prot´eg´e support the creation and edition of one or more ontologies in a single workspace. They nevertheless require a user to be familiar with this kind of abstractions and their supporting techniques such as a reasoner and SPARQL queries. This paper presents a step-by-step implementation of a…
Authors not listed
The analysis of nonadiabatic molecular dynamics (NAMD) data presents significant challenges due to its high dimensionality and complexity. To address these issues, we introduce ULaMDyn, a Python-based, open-source package designed to automate the unsupervised analysis of large datasets generated by NAMD simulations.…
Authors not listed
High-level quantum mechanical (QM) simulations provide accurate electronic information of chemical systems but scale unfavourably with system size, making calculations of applied systems challenging. Hierarchical quantum mechanics in quantum mechanics embedding (QM/QM) addresses this issue by localising the highly…