Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:LLMs for Generalizable Language-Conditioned Policy Learning under Minimal Data Requirements

Dec 09, 2024

Thomas Pouplin, Katarzyna Kobalczyk, Hao Sun, Mihaela van der Schaar

Figure 1 for LLMs for Generalizable Language-Conditioned Policy Learning under Minimal Data Requirements

Figure 2 for LLMs for Generalizable Language-Conditioned Policy Learning under Minimal Data Requirements

Figure 3 for LLMs for Generalizable Language-Conditioned Policy Learning under Minimal Data Requirements

Figure 4 for LLMs for Generalizable Language-Conditioned Policy Learning under Minimal Data Requirements

Share this with someone who'll enjoy it:

Abstract:To develop autonomous agents capable of executing complex, multi-step decision-making tasks as specified by humans in natural language, existing reinforcement learning approaches typically require expensive labeled datasets or access to real-time experimentation. Moreover, conventional methods often face difficulties in generalizing to unseen goals and states, thereby limiting their practical applicability. This paper presents TEDUO, a novel training pipeline for offline language-conditioned policy learning. TEDUO operates on easy-to-obtain, unlabeled datasets and is suited for the so-called in-the-wild evaluation, wherein the agent encounters previously unseen goals and states. To address the challenges posed by such data and evaluation settings, our method leverages the prior knowledge and instruction-following capabilities of large language models (LLMs) to enhance the fidelity of pre-collected offline data and enable flexible generalization to new goals and states. Empirical results demonstrate that the dual role of LLMs in our framework-as data enhancers and generalizers-facilitates both effective and data-efficient learning of generalizable language-conditioned policies.

View paper on

Share this with someone who'll enjoy it:

Title:LLMs for Generalizable Language-Conditioned Policy Learning under Minimal Data Requirements

Paper and Code