How To Stay Curious While Avoiding Noisy Tvs Using Aleatoric Uncertainty Estimation
2021 Β· Augustine N. Mavor-Parker, Kimberly A. Young, Caswell Barry, et al.
Abstract
Exploration in environments with sparse rewards is difficult for artificial agents. Curiosity driven learning -- using feed-forward prediction errors as intrinsic rewards -- has achieved some success in these scenarios, but fails when faced with action-dependent noise sources. We present aleatoric mapping agents (AMAs), a neuroscience inspired solution modeled on the cholinergic system of the mammalian brain. AMAs aim to explicitly ascertain which dynamics of the environment are unpredictable, regardless of whether those dynamics are induced by the actions of the agent. This is achieved by generating separate forward predictions for the mean and variance of future states and reducing intrinsic rewards for those transitions with high aleatoric variance. We show AMAs are able to effectively circumvent action-dependent stochastic traps that immobilise conventional curiosity driven agents. The code for all experiments presented in this paper is open sourced: http://github.com/self-supervis
Authors
(none)
Tags
Stats
Related papers
- Beyond Noisy-tvs: Noise-robust Exploration Via Learning Progress Monitoring (2025)0.00
- Wonder Wins Ways: Curiosity-driven Exploration Through Multi-agent Contextual Calibration (2025)0.00
- Curiosity-driven Exploration Via Latent Bayesian Surprise (2021)0.00
- Intrinsic Rewards For Exploration Without Harm From Observational Noise: A Simulation Study Based On The Free Energy Principle (2024)0.00
- Curiosity-driven Multi-agent Exploration With Mixed Objectives (2022)0.00
- Episodic Curiosity Through Reachability (2018)0.00
- Adaptive Symmetric Reward Noising For Reinforcement Learning (2019)0.00
- Is Curiosity All You Need? On The Utility Of Emergent Behaviours From Curious Exploration (2021)0.00