Real-Time Recurrent Reinforcement Learning

Lemmel, Julian; Grosu, Radu

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 2311

Computer Science > Machine Learning

Title: Real-Time Recurrent Reinforcement Learning

Authors: Julian Lemmel, Radu Grosu

(Submitted on 8 Nov 2023 (this version), latest version 28 Mar 2024 (v2))

Abstract: Recent advances in reinforcement learning, for partially-observable Markov decision processes (POMDPs), rely on the biologically implausible backpropagation through time algorithm (BPTT) to perform gradient-descent optimisation. In this paper we propose a novel reinforcement learning algorithm that makes use of random feedback local online learning (RFLO), a biologically plausible approximation of realtime recurrent learning (RTRL) to compute the gradients of the parameters of a recurrent neural network in an online manner. By combining it with TD($\lambda$), a variant of temporaldifference reinforcement learning with eligibility traces, we create a biologically plausible, recurrent actor-critic algorithm, capable of solving discrete and continuous control tasks in POMDPs. We compare BPTT, RTRL and RFLO as well as different network architectures, and find that RFLO can perform just as well as RTRL while exceeding even BPTT in terms of complexity. The proposed method, called real-time recurrent reinforcement learning (RTRRL), serves as a model of learning in biological neural networks mimicking reward pathways in the mammalian brain.

Comments:	12 pages, 8 figures, includes Appendix
Subjects:	Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE); Systems and Control (eess.SY)
Cite as:	arXiv:2311.04830 [cs.LG]
	(or arXiv:2311.04830v1 [cs.LG] for this version)

Submission history

From: Julian Lemmel [view email]
[v1] Wed, 8 Nov 2023 16:56:16 GMT (3934kb,D)
[v2] Thu, 28 Mar 2024 10:30:57 GMT (4065kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2311.04830v1

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Real-Time Recurrent Reinforcement Learning

Submission history