6 papers · ranked by Valyu relevance
Fabio Cumbo, Kabir Dhillon, M. Hassan Najafi, Sercan Aygun + 1 more
The exponential growth of genomic databases necessitates alignment-free methods for comparing genomes. While MinHash-based tools have revolutionized this field by efficiently estimating the Average Nucleotide Identity based on k-mer sets, they inherently discard structural genomic information. We introduce HyperSketch…
Eliezer Masliah
How transient neural representations become integrated and stable enough to function as internal neural models remains incompletely understood. Grounded in efficient coding, Bayesian and predictive frameworks, recurrent and attractor dynamics, neural state-space models, and systems neuroscience, the Principle of…
Hao Ding, Nannan Wu, Tianyi Qiu
DNA foundation models such as Evo2 7B adopt hybrid Hyena/attention architectures (Striped-Hyena2) whose single-stream autoregressive decoding is bounded by weight bandwidth at ∼45 tok/s. Speculative decoding on such hybrids faces a systems problem that prior SSM work solves only partially: after a draft is verified…
Xiaoyue Hu, Yuhao Ma, Ruixing Ming, Heping Zhang + 1 more
Identifying essential biomarkers remains a core challenge in elucidating the pathogenic mechanisms and achieving precise diagnosis of complex diseases. Deep neural networks offer immense predictive power, yet their lack of interpretability severely limits downstream biological insight. Here, we introduce DeepVaris, an…
Namasi G Sankar, Georgios Miliotis, Simon Caton
Genome assembly is important in infectious disease surveillance, antimicrobial resistance monitoring, and cancer genomics. The task of reconstructing full genomic sequences from fragmented reads, can be framed as a large scale combinatorial optimisation problem. Recent advances in quantum computing have introduced new…
Xuebin Feng, Emma R. Master
Sequence similarity networks (SSNs) are graphical representations of sequence relationship frequently used for exploring protein sequence space. Conventional SSN workflows typically use BLAST to calculate sequence similarities and rely on external visualization tools to generate the final networks. Consequently, raw…