17 papers · ranked by Valyu relevance
Yan Liu, Erik D. Reichle
Although different learning systems are coordinated to afford complex behavior, little is known about how this occurs. This article describes a theoretical framework that specifies how complex behaviors that might be thought to require error-driven learning might instead be acquired through simple reinforcement. This…
Cheng Chen, Yong Wang, Lizi Liao, Yueguo Chen + 1 more
> Abstract. Given a limited labeling budget, active learning (al) aims to sample the most informative instances from an unlabeled pool to acquire labels for subsequent model training. To achieve this, al typically measures the informativeness of unlabeled instances based on uncertainty and diversity. However, it does…
Krishnan Raghavan, Vignesh Narayanan, Jagannathan Saraangapani
— Learning to control complex systems using nontraditional feedback, e.g., in the form of snapshot images, is an important task encountered in diverse domains such as robotics, neuroscience, and biology (cellular systems). In this paper, we present a two neural-network (NN)-based feedback control framework to design…
Harshil Vejendla
Transformer models often exhibit brittle extrapolation, failing on inputs that are longer or structurally more complex than those seen during training. We introduce Counter-Example-Driven Curricula (CEDC), an automated framework that improves model robustness by iteratively focusing on its own failures. At each step…
Christopher Meek
Understanding prediction errors and determining how to fix them is critical to building effective predictive systems. In this paper, we delineate four types of prediction errors (mislabeling, representation, learner and boundary errors) and demonstrate that these four types characterize all prediction errors. In…
Akash Samanta, Sheldon Williamson
Learning systems deployed in nonstationary and safety-critical environments often suffer from instability, slow convergence, or brittle adaptation when learning dynamics evolve over time. While modern optimization, reinforcement learning, and meta-learning methods adapt to gradient statistics, they largely ignore the…
Samuel J. Gershman
Where do objective functions come from? How do we select what goals to pursue? Human intelligence is adept at synthesizing new objective functions on the fly. How does this work, and can we endow artificial systems with the same ability? This paper proposes an approach to answering these questions, starting with the…
Samuel Blad, Martin Längkvist, Amy Loutfi
Measuring learning progress is essential for curiosity-driven exploration in reinforcement learning, but widely used signals such as prediction error often fail to distinguish meaningful, learnable patterns from random noise. This paper proposes Gradient-Momentum Coupling (GMC), a signal derived from optimization…
Yariv Yanay
Quantum error correction is one of the fundamental building blocks of digital quantum computation. The Quantum Lego formalism has introduced a systematic way of constructing new stabilizer codes out of basic lego-like building blocks, which in previous work we have used to generate improved error correcting codes via…
Eric Pulick, Vladimir Meñkov, Yonatan Mintz, Paul B. Kantor + 1 more
'Vicki M. Bier'] Reliable real-world deployment of reinforcement learning (RL) methods requires a nuanced understanding of their strengths and weaknesses and how they compare to those of humans. Human-machine systems are becoming more prevalent and the design of these systems relies on a task-oriented understanding of…
Maier, Antoine, Maier, Aude + 2 more
—A common but rarely examined assumption in machine learning is that training yields models that actually satisfy their specified objective function. We call this the Objective Satisfaction Assumption (OSA). Although deviations from OSA are acknowledged, their implications are overlooked. We argue, in a…
Inês Lourenço, Rebecka Winqvist, Cristian R. Rojas, Bo Wahlberg
— A classical learning setting typically concerns an agent/student who collects data, or observations, from a system in order to estimate a certain property of interest. Correctional learning is a type of cooperative teacher-student framework where a teacher, who has partial knowledge about the system, has the ability…
Kumarjit Pathak, Jitin Kapila
In statistical modelling the biggest threat is concept drift which makes the model gradually showing deteriorating performance over time. There are state of the art methodologies to detect the impact of concept drift, however general strategy considered to overcome the issue in performance is to rebuild or re-calibrate…
Anestis Fachantidis, Matthew E. Taylor, Ioannis Vlahavas
—In this article we study the transfer learning model of action advice under a budget. We focus on reinforcement learning teachers providing action advice to heterogeneous students playing the game of Pac-Man under a limited advice budget. First, we examine several critical factors affecting advice quality in this…
Xiao Li, Hanchen Xu, Jinming Zhang, Hua Hua Chang
In this paper, we formulate the adaptive learning problem—the problem of how to find an individualized learning plan (called policy) that chooses the most appropriate learning materials based on learner's latent traits—faced in adaptive learning systems as a Markov decision process (MDP). We assume latent traits to be…
George Leu, Jiangjun Tang
Machine education is an emerging research field that focuses on the problem which is inverse to machine learning. To date, the literature on educating machines is still in its infancy. A fairly low number of methodology and method papers are scattered throughout various formal and informal publication avenues, mainly…
Yun‐Shiuan Chuang, Xuezhou Zhang, Yuzhe Ma, Mark K. Ho + 2 more
'Joseph L. Austerweil' 'Junwei Zhu'] Successful teaching requires an assumption of how the learner learns - how the learner uses experiences from the world to update their internal states. We investigate what expectations people have about a learner when they teach them in an online manner using rewards and punishment.…