Computer Science > Machine Learning

arXiv:2208.03374 (cs)

[Submitted on 5 Aug 2022]

Title:Learning to Generalize with Object-centric Agents in the Open World Survival Game Crafter

Authors:Aleksandar Stanić, Yujin Tang, David Ha, Jürgen Schmidhuber

View PDF

Abstract:Reinforcement learning agents must generalize beyond their training experience. Prior work has focused mostly on identical training and evaluation environments. Starting from the recently introduced Crafter benchmark, a 2D open world survival game, we introduce a new set of environments suitable for evaluating some agent's ability to generalize on previously unseen (numbers of) objects and to adapt quickly (meta-learning). In Crafter, the agents are evaluated by the number of unlocked achievements (such as collecting resources) when trained for 1M steps. We show that current agents struggle to generalize, and introduce novel object-centric agents that improve over strong baselines. We also provide critical insights of general interest for future work on Crafter through several experiments. We show that careful hyper-parameter tuning improves the PPO baseline agent by a large margin and that even feedforward agents can unlock almost all achievements by relying on the inventory display. We achieve new state-of-the-art performance on the original Crafter environment. Additionally, when trained beyond 1M steps, our tuned agents can unlock almost all achievements. We show that the recurrent PPO agents improve over feedforward ones, even with the inventory information removed. We introduce CrafterOOD, a set of 15 new environments that evaluate OOD generalization. On CrafterOOD, we show that the current agents fail to generalize, whereas our novel object-centric agents achieve state-of-the-art OOD generalization while also being interpretable. Our code is public.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
ACM classes:	I.2.6
Cite as:	arXiv:2208.03374 [cs.LG]
	(or arXiv:2208.03374v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2208.03374

Submission history

From: Aleksandar Stanic [view email]
[v1] Fri, 5 Aug 2022 20:05:46 UTC (960 KB)

Computer Science > Machine Learning

Title:Learning to Generalize with Object-centric Agents in the Open World Survival Game Crafter

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Learning to Generalize with Object-centric Agents in the Open World Survival Game Crafter

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators