Tuning for the Unknown: Revisiting Evaluation Strategies for Lifelong RL

Mesbahi, Golnaz; Mastikhina, Olya; Panahi, Parham Mohammad; White, Martha; White, Adam

Computer Science > Machine Learning

arXiv:2404.02113 (cs)

[Submitted on 2 Apr 2024 (v1), last revised 30 Apr 2024 (this version, v2)]

Title:Tuning for the Unknown: Revisiting Evaluation Strategies for Lifelong RL

Authors:Golnaz Mesbahi, Olya Mastikhina, Parham Mohammad Panahi, Martha White, Adam White

View PDF

Abstract:In continual or lifelong reinforcement learning access to the environment should be limited. If we aspire to design algorithms that can run for long-periods of time, continually adapting to new, unexpected situations then we must be willing to deploy our agents without tuning their hyperparameters over the agent's entire lifetime. The standard practice in deep RL -- and even continual RL -- is to assume unfettered access to deployment environment for the full lifetime of the agent. This paper explores the notion that progress in lifelong RL research has been held back by inappropriate empirical methodologies. In this paper we propose a new approach for tuning and evaluating lifelong RL agents where only one percent of the experiment data can be used for hyperparameter tuning. We then conduct an empirical study of DQN and Soft Actor Critic across a variety of continuing and non-stationary domains. We find both methods generally perform poorly when restricted to one-percent tuning, whereas several algorithmic mitigations designed to maintain network plasticity perform surprising well. In addition, we find that properties designed to measure the network's ability to learn continually indeed correlate with performance under one-percent tuning.

Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2404.02113 [cs.LG]
	(or arXiv:2404.02113v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2404.02113

Submission history

From: Golnaz Mesbahi [view email]
[v1] Tue, 2 Apr 2024 17:13:22 UTC (16,853 KB)
[v2] Tue, 30 Apr 2024 16:41:05 UTC (16,841 KB)

Computer Science > Machine Learning

Title:Tuning for the Unknown: Revisiting Evaluation Strategies for Lifelong RL

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Tuning for the Unknown: Revisiting Evaluation Strategies for Lifelong RL

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators