15 papers · ranked by Valyu relevance
Melissa Lober, Younes Bouhadjar, Markus Diesmann, Tom Tetzlaff
Processing sequential inputs is a fundamental brain function, underlying tasks such as sensory perception, language, and motor control. A challenge in sequence processing is to represent not only the order of events, but also their precise timing. While existing computational models can learn sequential structure, many…
Yen-Tung Yeh, Chung-Jui Chan, Yun-Ning, Hung + 1 more
Automatic music mixing, the task of automatically combining individual audio tracks into a cohesive mixture, is typically addressed by parallelized architectures that process all input tracks in a single pass. In this work, inspired by how human mix engineers process stems one at a time, we propose a paradigm shift and…
Jon Sporring, David Stansby
This report addresses larger-than-memory image analysis for petascale datasets such as 1.4 PB electron-microscopy volumes [[4]] and 150 TB human-organ atlases [[6]]. We argue that performance is fundamentally I/O-bound rather than compute-bound. We show that structuring analysis as streaming passes over data is…
Daniel Waxman, Fernando Llorente, Petar M. Djurić
The proliferation of capable and efficient machine learning (ML) models marks one of the strongest methodological shifts in signal processing (SP) in its nearly 100-year history. ML models support the development of SP systems that represent complex, nonlinear relationships with high predictive accuracy. Adapting these…
Rinku Sebastian, Simon O Keefe, Martin A Trefzer
This paper evaluates Reservoir Computing (RC) as an autonomous, 'feature-free' framework for audio processing, designed to eliminate traditional, handcrafted feature extraction stages. We investigate whether the high-dimensional temporal dynamics inherent in a reservoir can function as a robust end-to-end processor for…
Ziyan Guo, Wenji Fang, Wenkai Li, Yuchao Wu + 2 more
Accurate timing prediction at the register-transfer level (RTL) is a longstanding challenge in design automation. Existing graph-based methods struggle with limited receptive fields, high complexity, and a lack of signal directionality. We present RTL-Sequencer, a novel sequence-based paradigm that enables scalable RTL…
Wu Yonggang
| 1. Introduction4 | | |----------------------------------------------------------------------|--| | 2. Module 1: Visual processing system (VPS) module4 | | | 2.1 Complementary plasticity hypothesis (CPH)5 | | | 2.2 Training and tuning6 | | | 2.3 Specificity and invariance7 | | | 2.4 Invariance through excitatory…
Alan Zhao, Cyril Y. He, Wei Xu
Deployers of online LLM services usually seek to maximize cluster-wide performance given a fixed number of GPUs. Tensor parallelism (TP) is necessary to fit modern models but scales sub-linearly as the TP degree t grows, due to cross-GPU communication and non-scalable runtime work, as predicted by Amdahl's Law.…
Martinuzzi, Francesco
Recurrent neural networks (RNNs) are a cornerstone of sequence modeling across various scientific and industrial applications. Owing to their versatility, numerous RNN variants have been proposed over the past decade, aiming to improve the modeling of long-term dependencies and to address challenges such as vanishing…
Victor Norgren
Conventional transformer inference engines are request-driven, paying an O(n) prefill cost on every query. In streaming workloads, where data arrives continuously and queries probe an ever-growing context, this cost is prohibitive. We introduce a data-driven computational model centred on stateful sessions: a…
Tairan Xu, Leyang Xue, Zhan Lu, Jinfu Deng + 6 more
Batch inference has become a central mode of AI computation, yet existing inference engines still rely on execution models designed for interactive serving. When scaled to millions of sequences, batch workloads reveal two fundamental requirements: the ability to handle extreme inter- and intra-sequence load variation…
David Campos, Bin Yang, Tung Kieu, Lei Chen + 2 more
The ongoing digitization has led to a proliferation of time-series data streams that monitor a variety of processes, from which valuable insights may be obtained. Further, the emergence of successful foundational language models begs the question of whether it is possible to achieve time-series models with the…
Zhi Liu, Guangzhi Wang
Humans solve complex problems under limited cognitive resources through temporalized sequential reasoning [[31]]. Language relies on problem space search for deep semantic reasoning [[24]]. While early large language models (LLMs) could generate fluent text, they lacked robust semantic reasoning capabilities. Prompting…
Jintao Li, Weichang Li, Kai Tong, Xaingyu Guo
Distributed acoustic sensing (DAS) systems generate continuous, ultra-high-channel-count data streams at rates that exceed the capabilities of conventional batch-oriented analysis frameworks. As a result, essential tasks such as interactive exploration of long-duration recordings, scalable event annotation, and…
Shilpika Shilpika, Bethany Lusch, Eric Pershey, Carlo Graziani + 2 more
Modern supercomputers housed in High Performance Computing (HPC) environments generate massive volumes of log data daily, revealing intricate information and performance metrics about these complex systems. The sheer size and heterogeneous nature of HPC logs, especially text data, pose significant challenges for…