16 papers · ranked by Valyu relevance
Javier Geijo-Fernández, Alexander Pfundner, Carlos A Garcia-Perez
Microbial community profiling relies on comprehensive reference databases, yet full-length 16S rRNA amplicons remain sparse for many bacterial taxa. We present SGenerator, a neural network-based data augmentation method that generates biologically informative, full-length (1500 bp) 16S rRNA sequences for…
Katariina Perkonoja, Kari Auranen, Joni Virta
The rapid growth in data availability has facilitated research and development, yet not all industries have benefited equally due to legal and privacy constraints. The healthcare sector faces significant challenges in utilizing patient data because of concerns about data security and confidentiality. To address this…
Amy Schwartz, Pierre-Antoine Gourraud, Anil Kumar Vadathya, Ayush Bhattacharya + 22 more
'Ayush Bhattacharya' 'Samer El Kababji' 'Nicholas Mitsakakis' 'Elizabeth Jonker' 'Ana-Alicia Beltran-Bless' 'Gregory Pond' 'Lisa Vandermeer' 'Dhenuka Radhakrishnan' 'Lucy Mosquera' 'Alexander Paterson' 'Lois Shepherd' 'Bingshu Chen' 'William Barlow' 'Julie Gralow' 'Marie-France Savard' 'Christian Fesl' 'Dominik…
Hanan Shteingart, Yonatan Loewenstein, Giovanni Ponti
There is a long history of experiments in which participants are instructed to generate a long sequence of binary random numbers. The scope of this line of research has shifted over the years from identifying the basic psychological principles and/or the heuristics that lead to deviations from randomness, to one of…
Gabrielle Josling, Ibrahima Diouf, Sankalp Khanna
Clinical and health research increasingly depends on rich, structured datasets such as electronic health records (EHRs), disease registries, and longitudinal cohort studies. These tabular datasets provide the foundation for epidemiological research, assessment of healthcare outcomes, and data-driven policy development.…
Debapriya Hazra, Mi-Ryung Kim, Yung-Cheol Byun, Luca Agnelli
Nucleic acids are the basic units of deoxyribonucleic acid (DNA) sequencing. Every organism demonstrates different DNA sequences with specific nucleotides. It reveals the genetic information carried by a particular DNA segment. Nucleic acid sequencing expresses the evolutionary changes among organisms and…
Fengwei Jia, Hongli Zhu, Fengyuan Jia, Xinyue Ren + 3 more
'Hongming Tan' 'Wai Kin Victor Chan'] Recently, generative models have been gradually emerging into the extended dataset field, showcasing their advantages. However, when it comes to generating tabular data, these models often fail to satisfy the constraints of numerical columns, which cannot generate high-quality…
Sohom Ghosh, Shefali Yadav, Xin Wang, Bibhash Chakrabarty + 1 more
'Serdar Kadıoğlu'] Sequential pattern mining remains a challenging task due to the large number of redundant candidate patterns and the exponential search space. In addition, further analysis is still required to map extracted patterns to different outcomes. In this paper, we introduce a pattern mining framework that…
Alexey Zaripov, Roman Kulshin, Anatoly Sidorov, Miguel Angel Guevara Lopez + 2 more
'Miguel Angel Guevara Lopez' 'Luís Gonzaga Mendes Magalhães' 'Edel Bartolo Garcia Reyes'] This work is dedicated to the development of a system for generating artificial data for training neural networks used within a conveyor-based technology framework. It presents an overview of the application areas of computer…
Alfred Ultsch, Jörn Lötsch
Small sample sizes in biomedical research often led to poor reproducibility and challenges in translating findings into clinical applications. This problem stems from limited study resources, rare diseases, ethical considerations in animal studies, costly expert diagnosis, and others. As a contribution to the problem…
Osvaldo Navarro, René Cumplido, Luis Villaseñor-Pineda, Claudia Feregrino-Uribe + 2 more
'Claudia Feregrino-Uribe' 'Jesús Ariel Carrasco-Ochoa' 'Francesco Pappalardo'] Sequential Pattern Mining is a widely addressed problem in data mining, with applications such as analyzing Web usage, examining purchase behavior, and text mining, among others. Nevertheless, with the dramatic increase in data volume, the…
Laura Savaré, Francesca Ieva, Giovanni Corrao, Antonio Lora
Background Care pathways are increasingly being used to enhance the quality of care and optimize the use of resources for health care. Nevertheless, recommendations regarding the sequence of care are mostly based on consensus-based decisions as there is a lack of evidence on effective treatment sequences. In a…
Ovidiu Popa, Ellen Oldenburg, Oliver Ebenhöh
Today massive amounts of sequenced metagenomic and metatranscriptomic data from different ecological niches and environmental locations are available. Scientific progress depends critically on methods that allow extracting useful information from the various types of sequence data. Here, we will first discuss types of…
Rafał Deja, Wojciech Froelich, GraŻyna Deja
Background In spite of numerous research efforts on supporting the therapy of diabetes mellitus, the subject still involves challenges and creates active interest among researchers. In this paper, a decision support tool is presented for setting insulin therapy in new-onset type 1 diabetes. Methods The concept of…
Tian Gan, Jindong Chen, Hao Wang, Conghui Shang + 5 more
Objective To evaluate the impact of sequential (first- to third-generation) epidermal growth factor receptor tyrosine kinase inhibitor (EGFR-TKI) treatment on top-corrected QT interval (top-QTc) in non-small cell lung cancer (NSCLC) patients. Methods We retrospectively reviewed the medical records of NSCLC patients…
Xiya Ma, Shaoxing Yang, Kun Zhang, Jing Xu + 5 more
To investigate the difference in survival of patients with advanced ALK-positive NSCLC after crizotinib progression under different treatment patterns, we retrospectively analyzed the clinical data of 128 patients who received crizotinib as initial ALK-TKI and demonstrated the clinical outcomes of different sequential…