3 papers · ranked by Valyu relevance
Seonghwan Seo, Hyeongwoo Kim, Seokhyun Moon, Woo Youn Kim
Protein language models (PLMs) trained on evolutionary sequences learn representations that encode protein structure, enabling direct structure prediction without multiple-sequence alignments (MSAs). Here we present the Atlas model family, an open and trainable system spanning protein language modeling, monomer…
Fabio Cumbo, Kabir Dhillon, M. Hassan Najafi, Sercan Aygun + 1 more
The exponential growth of genomic databases necessitates alignment-free methods for comparing genomes. While MinHash-based tools have revolutionized this field by efficiently estimating the Average Nucleotide Identity based on k-mer sets, they inherently discard structural genomic information. We introduce HyperSketch…
Anthony Christidis, Andrew Ghazi, Smriti Chawla, Nitesh Turaga + 2 more
Although cell type annotation has become an integral part of single-cell analysis workflows, the assessment of computational annotations remains challenging. Many annotation tools transfer labels from an annotated reference dataset to a new query dataset of interest, but blindly transferring labels from one dataset to…