Attention Loss Adjusted Prioritized Experience Replay

Chen, Zhuoying; Li, Hui**; Wang, Rizhong

Computer Science > Machine Learning

arXiv:2309.06684 (cs)

[Submitted on 13 Sep 2023 (v1), last revised 9 Oct 2023 (this version, v2)]

Title:Attention Loss Adjusted Prioritized Experience Replay

Authors:Zhuoying Chen, Hui** Li, Rizhong Wang

View PDF

Abstract:Prioritized Experience Replay (PER) is a technical means of deep reinforcement learning by selecting experience samples with more knowledge quantity to improve the training rate of neural network. However, the non-uniform sampling used in PER inevitably shifts the state-action space distribution and brings the estimation error of Q-value function. In this paper, an Attention Loss Adjusted Prioritized (ALAP) Experience Replay algorithm is proposed, which integrates the improved Self-Attention network with Double-Sampling mechanism to fit the hyperparameter that can regulate the importance sampling weights to eliminate the estimation error caused by PER. In order to verify the effectiveness and generality of the algorithm, the ALAP is tested with value-function based, policy-gradient based and multi-agent reinforcement learning algorithms in OPENAI gym, and comparison studies verify the advantage and efficiency of the proposed training framework.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2309.06684 [cs.LG]
	(or arXiv:2309.06684v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2309.06684

Submission history

From: Zhuoying Chen [view email]
[v1] Wed, 13 Sep 2023 02:49:32 UTC (876 KB)
[v2] Mon, 9 Oct 2023 03:12:03 UTC (768 KB)

Computer Science > Machine Learning

Title:Attention Loss Adjusted Prioritized Experience Replay

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Attention Loss Adjusted Prioritized Experience Replay

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators