COOKIES: By using this website you agree that we can place Google Analytics Cookies on your device for performance monitoring. |
University of Cambridge > Talks.cam > Microsoft Research Cambridge, public talks > Distributed distributional codes for learning successor features in partially observable environments
Distributed distributional codes for learning successor features in partially observable environmentsAdd to your list(s) Download to your calendar using vCal
If you have a question about this talk, please contact Microsoft Research Cambridge Talks Admins. Please note, this event may be recorded. Microsoft will own the copyright of any recording and reserves the right to distribute it as required. Animals need to devise strategies to maximise returns while interacting with their environment based on incoming noisy sensory observations. Task-relevant states, such as the agent’s location within an environment or the presence of a predator, are often not directly observable but must be inferred using available sensory information. Successor representations (SR) have been proposed as a middle-ground between model-based and model-free reinforcement learning strategies, allowing for fast value computation and rapid adaptation to changes in the reward function or goal locations. Indeed, recent studies suggest that features of neural responses are consistent with the SR framework. However, it is not clear how such representations might be learned and computed in partially observed, noisy environments. Here, we introduce a neurally plausible model using distributional successor features, which builds on the distributed distributional code for the representation and computation of uncertainty, and which allows for efficient value function computation in partially observed environments via the successor representation. We show that distributional successor features can support reinforcement learning in noisy environments in which direct learning of successful policies is infeasible. This talk is part of the Microsoft Research Cambridge, public talks series. This talk is included in these lists:
Note that ex-directory lists are not shown. |
Other listsWildlife and Environment Visiting African Fellows' Research Showcase Qatar Carbonates and Carbon Storage Research Centre: Status update after three years of fundamental researchOther talksCANCELLED: Curator’s introduction to Virtue, Vice & the Senses: Prints 1540 – 1650 Majorana Fermions in Condensed Matter 1 Predicting Long Runout Landslides Uncovering the role of centrioles and cilia in signal transduction and metabolism POSTPONED: Neuro-oncology Seminar April 2020 Distinct mucosal or systemic responses to non-pathogenic intestinal microbes build the baseline B cell repertoire and its functional responsiveness. |