Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Learning when to observe: A frugal reinforcement learning framework for a high-cost world

Jul 24, 2023

Colin Bellinger, Mark Crowley, Isaac Tamblyn

Figure 1 for Learning when to observe: A frugal reinforcement learning framework for a high-cost world

Figure 2 for Learning when to observe: A frugal reinforcement learning framework for a high-cost world

Figure 3 for Learning when to observe: A frugal reinforcement learning framework for a high-cost world

Figure 4 for Learning when to observe: A frugal reinforcement learning framework for a high-cost world

Share this with someone who'll enjoy it:

Abstract:Reinforcement learning (RL) has been shown to learn sophisticated control policies for complex tasks including games, robotics, heating and cooling systems and text generation. The action-perception cycle in RL, however, generally assumes that a measurement of the state of the environment is available at each time step without a cost. In applications such as materials design, deep-sea and planetary robot exploration and medicine, however, there can be a high cost associated with measuring, or even approximating, the state of the environment. In this paper, we survey the recently growing literature that adopts the perspective that an RL agent might not need, or even want, a costly measurement at each time step. Within this context, we propose the Deep Dynamic Multi-Step Observationless Agent (DMSOA), contrast it with the literature and empirically evaluate it on OpenAI gym and Atari Pong environments. Our results, show that DMSOA learns a better policy with fewer decision steps and measurements than the considered alternative from the literature. The corresponding code is available at: \url{https://github.com/cbellinger27/Learning-when-to-observe-in-RL

* Accepted for presentation at ECML-PKDD 2023 workshop track: Simplification, Compression, Efficiency and Frugality for Artificial Intelligence (SCEFA)

View paper on

Share this with someone who'll enjoy it:

Title:Learning when to observe: A frugal reinforcement learning framework for a high-cost world

Paper and Code