Skip to main content

Michal Valko : Paper

Reinforcement learning using density estimation with online clustering for exploration

Alaa Saade, Steven Kapturowski, Daniele Calandriello, Charles Blundell, Michal Valko, Pablo Sprechmann, Bilal Piot

2023 · International Patent App. PCT/EP2023/076893 (2023)

Abstract

A method for exploration in reinforcement learning that combines online clustering with density estimation to compute novelty-based intrinsic rewards, enabling agents to efficiently track state visitation counts across thousands of episodes in sparse-reward environments.