14 papers · ranked by Valyu relevance
Wang, Zhichao, Ma, Dongyang + 14 more
The "end-to-end" label for LLMs is a misnomer. In practice, they depend on a nondifferentiable decoding process that requires laborious, hand-tuning of hyperparameters like temperature and top-p. This paper introduces AutoDeco, a novel architecture that enables truly "end-to-end" generation by learning to control its…
Bajpai, Divya Jyoti, Hanawal, Manjesh Kumar
Vision-language Models (VLMs) have made significant strides in visual understanding and query response generation, but often face challenges of high computational cost and inference latency due to autoregressive decoding. In this work, we introduce an imitation-learningbased Self-Speculative Decoding (SSD) framework…
Jun Zhang, Jue Wang, Huan Li, Lidan Shou + 3 more
'Sharad Mehrotra'] We present a novel inference scheme, selfspeculative decoding, for accelerating Large Language Models (LLMs) without the need for an auxiliary model. This approach is characterized by a two-stage process: drafting and verification. The drafting stage generates draft tokens at a slightly lower quality…
Eyal Cohen, Bhiksha Raj, Joseph Keshet
—Self-supervised automatic speech recognition (SSL-ASR) is an ASR approach that uses speech encoders pretrained on large amounts of unlabeled audio (e.g., wav2vec2.0 or Hu-BERT) and then fine-tunes them with limited labeled data to perform transcription. Decoding is usually performed with a CTC decoder, whose…
Nikolaus Kriegeskorte, Pamela K. Douglas
Encoding and decoding models are widely used in systems, cognitive, and computational neuroscience to make sense of brain-activity data. However, the interpretation of their results requires care. Decoding models can help reveal whether particular information is present in a brain region in a format the decoder can…
Alon Helvits, Eliya Nachmani
Error-correcting codes enable reliable communication, yet practical soft decoding remains challenging across code families and block lengths. We propose SB-ECC, a score-based decoder that casts decoding as continuous-time denoising. A neural denoiser defines a probability-flow ordinary differential equation (ODE) that…
Yukun Cheng, Wei Chen, Lun Li, Bo Ai
—Deep learning based decoding networks have shown significant improvement in decoding LDPC codes, but the neural decoders are limited by rate-matching operations such as puncturing or extending, thus needing to train multiple decoders with different code rates for a variety of channel conditions. In this…
Raphaël Le Bidan, Ahmad Ismail, Elsa Dupraz, Charbel Abdel Nour
Syndrome-based neural decoding (SBND) has emerged as a promising deep learning approach for soft-decision decoding of high-rate, short-length codes. However, this approach still has substantial room for improvement. In this paper, we show how to leverage code automorphisms to enhance the ability of existing SBND models…
Hung T. Nguyen, Steven Bottone, Kwang Taik Kim, Mung Chiang + 1 more
'H. Vincent Poor'] Abstract—Error correcting codes are a fundamental component in modern day communication systems, demanding extremely high throughput, ultra-reliability and low latency. Recent approaches using machine learning (ML) models as the decoders offer both improved performance and great adaptability to…
Biao Zhang, Deyi Xiong, Jinsong Su
With parallelizable attention networks, the neural Transformer is very fast to train. However, due to the auto-regressive architecture and self-attention in the decoder, the decoding procedure becomes slow. To alleviate this issue, we propose an average attention network as an alternative to the self-attention network…
Łukasz Kaiser, Samy Bengio
Recurrent models for sequences have been recently successful at many tasks, especially for language modeling and machine translation. Nevertheless, it remains challenging to extract good representations from these models. For instance, even though language has a clear hierarchical structure going from characters…
Johannes Zenn, Jonas Geiping
Many decoding methods for large language models can be understood as shifting probability mass toward outputs that are more likely under the model, either locally at the token level or globally at the sequence level. Therefore, their success depends on a fundamental question: when does sequence probability, that is…
Vickram N. Premakumar, Michael Vaiana, Florin Pop, Judd Rosenblatt + 3 more
'Diogo Schwerz de Lucena' 'Kirsten Ziman' 'Michael S. A. Graziano'] Self-models have been a topic of great interest for decades in studies of human cognition and more recently in machine learning. Yet what benefits do self-models confer? Here we show that when artificial networks learn to predict their internal states…
Yixin Gao, Runsen Feng, Zongyu Guo, Zhibo Chen
—Despite a short history, neural image codecs have been shown to surpass classical image codecs in terms of ratedistortion performance. However, most of them suffer from significantly longer decoding times, which hinders the practical applications of neural image codecs. This issue is especially pronounced when…