20 papers · ranked by Valyu relevance
Lulu Ding, Kun Wang, Hongmei Zhang, Shaohui Xie + 5 more
DNA storage offers exceptional information density and archival longevity, but is constrained by the complex, heterogeneous errors inherent to synthesis, storage, and sequencing. Conventional error-correction schemes often rely on excessive logical redundancy to mitigate these biochemical imperfections, thereby…
Wei-Cheng Tseng, David Harwath
—Recent advancements in neural audio codecs have not only enabled superior audio compression but also enhanced speech synthesis techniques. Researchers are now exploring their potential as universal acoustic feature extractors for a broader range of speech processing tasks. Building on this trend, we introduce…
Florian Grötschla, Arunasish Sen, Alessandro Lombardi, Guillermo Cámbara + 1 more
—We present VCNAC, a variable channel neural audio codec. Our approach features a single encoder and decoder parametrization that enables native inference for different channel setups, from mono speech to cinematic 5.1 channel surround audio. Channel compatibility objectives ensure that multichannel content maintains…
Shi, Jiatong, Wang Haoran, Chen + 7 more
—Neural speech codecs have achieved strong performance in low-bitrate compression, but residual vector quantization (RVQ) often suffers from unstable training and ineffective decomposition, limiting reconstruction quality and efficiency. We propose PURE Codec (Progressive Unfolding of Residual Entropy), a novel…
Authors not listed
Modeling of chemical reactions is essential for understanding kinetic mechanisms and predicting possible outcomes of reacting systems. Quantum mechanical calculations are accurate but often prohibitively expensive. Deep learning has emerged as a faster alternative, but progress is slowed by a fragmented software…
Ramy Khabbaz, Jérémy Mateos, Marc Antonini, Serge Kas Hanna
The biochemical processes underlying DNA data storage, including synthesis, amplification, and sequencing, are inherently noisy. Consequently, base-level insertion, deletion, and substitution (IDS) errors, as well as sequence-level dropouts, occur and pose major challenges for reliable data retrieval. Here we introduce…
Wan, Zixiang, Zhang, Guochang + 4 more
Neural Audio Codecs (NACs) have gained growing attention in recent years as technologies for audio compression and audio representation in speech language models. While mainstream NACs typically require G-level computation and M-level parameters, the performance of lightweight and streaming NACs remains underexplored.…
Wojcicki, Kamil, Isik, Yusuf Ziya + 18 more
While recent neural audio codecs deliver superior speech quality at ultralow bitrates over traditional methods, their practical adoption is hindered by obstacles related to low-resource operation and robustness to acoustic distortions. Edge deployment scenarios demand codecs that operate under stringent compute…
Ricardo Moreno, Jorge Ortigoso-Narro, Daniel de la Prida, Luis A. Azpicueta-Ruiz + 5 more
Highlights What are the main findings?1. The research and development of a scalable 1024-MEMS microphone array using a distributed network of 16 BeagleBone Black modules and TDM architecture. 2. Microsecond synchronization across the distributed nodes is achieved using PTP and PHC discipline with a priority-scheduled…
Qiyu Zha, Jiangling Guo, Hocine Cherifi
Video compression is central to large-scale video delivery, where better rate-distortion efficiency directly reduces bandwidth and storage cost. A practical way to improve efficiency is to encode a low-resolution video stream with a standard codec and restore high-resolution details with a learned super-resolution…
Jun Xu, Zhengxue Cheng, Fengxi Zhang, Yuhan Liu + 2 more
Learning-based speech compression has achieved promising low-bitrate performance, but many neural speech codecs still describe quantized latents with preset-rate discrete symbols or apply entropy coding only after symbol generation. Such designs decouple representation learning from probability modeling, limiting their…
Kalyani Rohidas Vaidya, Dhanalakshmi Kannur Munirathnam, Patrick Seeling, Raffaele Bruno + 1 more
As Large Language Models (LLMs) increasingly operate in agentic environments communicating via the Model Context Protocol (MCP), the structured JSON-RPC message traffic is growing rapidly. Yet the current MCP specification does not include any compression mechanism, leaving potential bandwidth optimizations…
Xiao-Hang Jiang, Yang Ai, Hui-Peng Du, Zhen-Hua Ling + 1 more
High-quality speech coding at low bitrates is crucial for bandwidth-constrained applications, yet remains challenging due to the severe loss of quality-critical information in highly compressed representations. To overcome this challenge, we propose CFMDCTCodec, a low-bitrate neural speech codec that operates entirely…
Linh T.P. Le, Omkar Hegde, Wei-Huan Wu, Ayesha Ejaz + 3 more
High-throughput microfluidics has transformed biomedical research by enabling precise and parallel sample handling, but most devices are single-use due to channel occlusion and contamination from experiments. Alongside low fabrication yield and reduced experimental success associated with dense microfeatures, this…
Alin-Adrian Alecu, Mohammad Ali Tahouri, Adrian Munteanu, Bujor Păvăloiu + 1 more
Near-lossless coding schemes traditionally rely on uniform quantization to control the maximum absolute error ( $L_{\infty}$ norm) of residual signals, often assuming a parametric model for the source distribution. This paper introduces a novel design framework for non-uniform, entropy-aware $L_{\infty}$-oriented…
Sina A. Schwarze, Ulman Lindenberger, Silvia A. Bunge, Yana Fandakova
Cognitive training often aims to improve cognitive skills, but outcomes have been variable in terms of their success. One factor that has been found to predict training outcomes is the degree of modularity of functional brain networks, defined as the extent to which brain regions are more strongly connected to regions…
Duong, Thien T., Springer, Jan P.
Perceptual quality of audio is the combination of aural accuracy and listener-perceived sound fidelity. It is how humans respond to the accuracy, intelligibility, and fidelity of aural media. Today this fidelity is also heavily influenced by the use of audio compression codecs for storing aural media in digital form.…
Ahmed A. Harby, Farhana Zulkernine, Hanady M. Abdulsalam
The rapid growth of multimedia content has increased the demand for effective methods to reduce storage requirements while maintaining quality and enabling fast data transmission. Existing standards and generative model approaches often involve high computational cost, require extensive parameter tuning, and produce…
Zhengda Wu, Jie Zhang, Agnieszka Konys
Consumer Electronics (CE) currently face dual pressures: the growing demand for personalization and multi-scenario usage, and the need to enhance product sustainability. Traditional single-function designs struggle to adapt to these changes, leading to resource underutilization and the generation of significant…
Ya Liu, Rui Zhang, Yong Zhang, Yuwei Chen + 1 more
Large field-of-view (FOV) infrared imaging, widely utilized in applications including target detection and remote sensing, generates massive datasets that pose significant challenges for transmission and storage. To address this issue, we propose an efficient lossless compression method for large FOV infrared video.…