Continual Reinforcement Learning With Multi-Timescale Replay. (arXiv:2004.07530v1 [cs.LG])

[Submitted on 16 Apr 2020]

Abstract: In this paper, we propose a multi-timescale replay (MTR) buffer for improving
continual learning in RL agents faced with environments that are changing
continuously over time at timescales that are unknown to the agent. The basic
MTR buffer comprises a cascade of sub-buffers that accumulate experiences at
different timescales, enabling the agent to improve the trade-off between
adaptation to new data and retention of old knowledge. We also combine the MTR
framework with invariant risk minimization, with the idea of encouraging the
agent to learn a policy that is robust across the various environments it
encounters over time. The MTR methods are evaluated in three different
continual learning settings on two continuous control tasks and, in many cases,
show improvement over the baselines.

Submission history

From: Christos Kaplanis [view email]
[v1]
Thu, 16 Apr 2020 08:47:40 UTC (4,451 KB)

Source: http://arxiv.org/abs/2004.07530

Plato Data Intelligence.
Vertical Search & Ai.

Continual Reinforcement Learning with Multi-Timescale Replay. (arXiv:2004.07530v1 [cs.LG])

Submission history

‘Liquid Vesting’ Is Oxymoronic Blockchain Feature That Lets Early Investors Sell Without Waiting

8 Reasons Why LayerZero’s Upcoming Airdrop Could Be the Most Complex Ever – Unchained

Latest Intelligence

Binance Releases 18th Proof-of-Researve; Crypto Enthusiasts Seize Profit Opportunity On Milei Moneda ($MEDA) Presale

State of Wisconsin Buys Nearly $100M Worth of BlackRock Spot Bitcoin ETF

Coinbase outage hampers Bitcoin trading amid price swings

PEPE hits record high as Gamestop meme trading returns

Liminal Custody Cements Digital Asset Custody Dominance with ADGM FSP License

Deutsche Bank joins Singapore’s Project Guardian to advance asset tokenization

Chat with us

Plato Data Intelligence.Vertical Search & Ai.

Continual Reinforcement Learning with Multi-Timescale Replay. (arXiv:2004.07530v1 [cs.LG])

Submission history

Latest Intelligence

Chat with us

Plato Data Intelligence.
Vertical Search & Ai.