Improving Computational Efficiency in Visual Reinforcement Learning via   Stored Embeddings

Lili Chen; Kimin Lee; Aravind Srinivas; Pieter Abbeel

arXiv:2103.02886·cs.LG·October 29, 2021·6 cites

Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings

Lili Chen, Kimin Lee, Aravind Srinivas, Pieter Abbeel

PDF

Open Access 1 Repo 1 Video

TL;DR

SEER introduces a simple modification to off-policy deep reinforcement learning by freezing CNN layers early and storing low-dimensional embeddings, significantly reducing computation and memory without performance loss.

Contribution

The paper proposes Stored Embeddings for Efficient Reinforcement Learning (SEER), a method that reduces memory and computational demands in visual RL by freezing CNN layers and storing low-dimensional embeddings.

Findings

01

SEER significantly reduces memory usage and computation.

02

SEER maintains RL performance across various environments.

03

Freezing CNN layers accelerates training without accuracy loss.

Abstract

Recent advances in off-policy deep reinforcement learning (RL) have led to impressive success in complex tasks from visual observations. Experience replay improves sample-efficiency by reusing experiences from the past, and convolutional neural networks (CNNs) process high-dimensional inputs effectively. However, such techniques demand high memory and computational bandwidth. In this paper, we present Stored Embeddings for Efficient Reinforcement Learning (SEER), a simple modification of existing off-policy RL methods, to address these computational and memory requirements. To reduce the computational overhead of gradient updates in CNNs, we freeze the lower layers of CNN encoders early in training due to early convergence of their parameters. Additionally, we reduce memory requirements by storing the low-dimensional latent vectors for experience replay instead of high-dimensional…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

lili-chen/SEER
pytorchOfficial

Videos

Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings· slideslive

Taxonomy

TopicsReinforcement Learning in Robotics · Neural dynamics and brain function · Advanced Neural Network Applications

Methods*Communicated@Fast*How Do I Communicate to Expedia? · Grouped Convolution · Dense Connections · 1x1 Convolution · Batch Normalization · Sigmoid Activation · Squeeze-and-Excitation Block · Average Pooling · Global Average Pooling · Convolution