13 papers · ranked by Valyu relevance
Pranav Mahajan, Ben Seymour
The seminal reward prediction error theory of dopamine function faces several key challenges. Most notable is the difficulty learning multiple rewards simultaneously, inefficient on-policy learning, and accounting for heterogeneous striatal responses in the tail of the striatum. We propose a normative framework, based…
James D. Boyko, Kyle J. Gontjes, Evan S. Snitkin, Stephen A. Smith
Ancestral state reconstruction (ASR) is a foundational tool in comparative biology, offering insights into the evolutionary history of lineages. With each new evolutionary model, our ability to estimate ancestral states has improved alongside the increased biological realism of these models. However, the field has…
Yusi Chen, Angela Radulescu, Herbert Zheng Wu
Understanding the intentions and beliefs of others, a phenomenon known as “theory of mind”, is a crucial element in social behavior. These beliefs and perceptions are inherently subjective and latent, making them often unobservable for investigation. Social interactions further complicate the matter, as multiple agents…
Jay A. Hennig, Sandra A. Romero Pinto, Takahiro Yamaguchi, Scott W. Linderman + 2 more
To behave adaptively, animals must learn to predict future reward, or value. To do this, animals are thought to learn reward predictions using reinforcement learning. However, in contrast to classical models, animals must learn to estimate value using only incomplete state information. Previous work suggests that…
Pranav Mahajan, Ben Seymour
The seminal reward prediction error account of dopamine has been highly successful, but faces several key challenges. Most notable are the difficulty of learning multiple rewards simultaneously, inefficient on-policy learning, and accounting for the heterogeneous striatal responses observed across and within striatal…
Sarah Schwöbel, Dimitrije Markovic, Michael N. Smolka, Stefan Kiebel
In cognitive neuroscience and psychology, reaction times are an important behavioral measure. However, in instrumental learning and goal-directed decision making experiments, findings often rely only on choice probabilities from a value-based model, instead of reaction times. Recent advancements have shown that it is…
Armin Bazarjani, Payam Piray
Cognitive maps enable flexible behavior by providing reusable internal representations of task structure. The successor representation, a predictive map that encodes expected future state occupancy, has been proposed as one way such maps might be computed in the brain, but its policy dependence severely limits flexible…
Justin Chow, Yunran Yang, Brokoslaw Laschowski
Inverse reinforcement learning can recover reward functions from observed behavior, but interpreting those rewards remains a fundamental challenge for understanding intelligent behavior and decision-making. To address this challenge, we introduce a novel framework for reward interpretation that combines reward-function…
Jonathan Ferrer-Mestres, Thomas G. Dietterich, Olivier Buffet, Iadine Chadès
In conservation of biodiversity, natural resource management and behavioural ecology, stochastic dynamic programming, and its mathematical framework, Markov decision processes (MDPs), are used to inform sequential decision-making under uncertainty. Models and solutions of Markov decision problems should be…
Bowen Zheng, Scott L. Brincat, Jacob A. Donoghue, Earl K. Miller + 1 more
Under a range of behavioral and physiological conditions, spike times and local field potential (LFP) oscillations exhibit phase coupling within specific frequency bands. Classical measures such as spike–field coherence (SFC) and the phase-locking value (PLV) quantify this coupling but estimate the LFP spectrum…
Takayuki Tsurumi, Ayaka Kato, Arvind Kumar, Kenji Morita
How external/internal ‘state’ is represented in the brain is crucial, since appropriate representation enables goal-directed behavior. Recent studies suggest that state representation and state value can be simultaneously learnt through reinforcement learning (RL) using reward-prediction-error in recurrent-…
Varun Canamedi
The structural wiring of the brain is expected to produce a repertoire of functional networks, across time, context, individuals and vice versa. Therefore, a method to infer the joint distribution of structural and functional connectomes would be of immense value. However, existing approaches only provide deterministic…
Callan M. Gillespie, Nicholas J. Haas, Tara F. Nagle, Robb W. Colbrunn
To quantify the contribution of specific ligaments to overall joint movement, the principle of superposition has been used for nearly 30 years. This principle relies on using a robotic test system to move a biological joint to the same position before and after transecting a ligament. The difference in joint forces…