Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Mary Cassin

Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning

Dec 07, 2021

DeepMind Interactive Agents Team, Josh Abramson, Arun Ahuja, Arthur Brussee, Federico Carnevale, Mary Cassin, Felix Fischer, Petko Georgiev, Alex Goldin, Tim Harley(+14 more)

Figure 1 for Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning

Figure 2 for Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning

Figure 3 for Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning

Figure 4 for Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning

Via

Access Paper or Ask Questions

Alchemy: A structured task distribution for meta-reinforcement learning

Feb 04, 2021

Jane X. Wang, Michael King, Nicolas Porcel, Zeb Kurth-Nelson, Tina Zhu, Charlie Deck, Peter Choy, Mary Cassin, Malcolm Reynolds, Francis Song(+7 more)

Figure 1 for Alchemy: A structured task distribution for meta-reinforcement learning

Figure 2 for Alchemy: A structured task distribution for meta-reinforcement learning

Figure 3 for Alchemy: A structured task distribution for meta-reinforcement learning

Figure 4 for Alchemy: A structured task distribution for meta-reinforcement learning

Abstract:There has been rapidly growing interest in meta-learning as a method for increasing the flexibility and sample efficiency of reinforcement learning. One problem in this area of research, however, has been a scarcity of adequate benchmark tasks. In general, the structure underlying past benchmarks has either been too simple to be inherently interesting, or too ill-defined to support principled analysis. In the present work, we introduce a new benchmark for meta-RL research, which combines structural richness with structural transparency. Alchemy is a 3D video game, implemented in Unity, which involves a latent causal structure that is resampled procedurally from episode to episode, affording structure learning, online inference, hypothesis testing and action sequencing based on abstract domain knowledge. We evaluate a pair of powerful RL agents on Alchemy and present an in-depth analysis of one of these agents. Results clearly indicate a frank and specific failure of meta-learning, providing validation for Alchemy as a challenging benchmark for meta-RL. Concurrent with this report, we are releasing Alchemy as public resource, together with a suite of analysis tools and sample agent trajectories.

* 16 pages, 9 figures

Via

Access Paper or Ask Questions

Imitating Interactive Intelligence

Jan 21, 2021

Josh Abramson, Arun Ahuja, Iain Barr, Arthur Brussee, Federico Carnevale, Mary Cassin, Rachita Chhaparia, Stephen Clark, Bogdan Damoc, Andrew Dudzik(+19 more)

Figure 1 for Imitating Interactive Intelligence

Figure 2 for Imitating Interactive Intelligence

Figure 3 for Imitating Interactive Intelligence

Figure 4 for Imitating Interactive Intelligence

Abstract:A common vision from science fiction is that robots will one day inhabit our physical spaces, sense the world as we do, assist our physical labours, and communicate with us through natural language. Here we study how to design artificial agents that can interact naturally with humans using the simplification of a virtual environment. This setting nevertheless integrates a number of the central challenges of artificial intelligence (AI) research: complex visual perception and goal-directed physical control, grounded language comprehension and production, and multi-agent social interaction. To build agents that can robustly interact with humans, we would ideally train them while they interact with humans. However, this is presently impractical. Therefore, we approximate the role of the human with another learned agent, and use ideas from inverse reinforcement learning to reduce the disparities between human-human and agent-agent interactive behaviour. Rigorously evaluating our agents poses a great challenge, so we develop a variety of behavioural tests, including evaluation by humans who watch videos of agents or interact directly with them. These evaluations convincingly demonstrate that interactive training and auxiliary losses improve agent behaviour beyond what is achieved by supervised learning of actions alone. Further, we demonstrate that agent capabilities generalise beyond literal experiences in the dataset. Finally, we train evaluation models whose ratings of agents agree well with human judgement, thus permitting the evaluation of new agent models without additional effort. Taken together, our results in this virtual environment provide evidence that large-scale human behavioural imitation is a promising tool to create intelligent, interactive agents, and the challenge of reliably evaluating such agents is possible to surmount.

Via

Access Paper or Ask Questions