Methods › Reinforcement Learning › Policy Gradient Methods › ACER
ACER
Introduced by Ziyu Wang et al. in Sample Efficient Actor-Critic with Experience Replay
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
ACER, or Actor Critic with Experience Replay, is an actor-critic deep reinforcement learning agent with experience replay. It can be seen as an off-policy extension of A3C, where the off-policy estimator is made feasible by:
- Using Retrace Q-value estimation.
- Using truncated importance sampling with bias correction.
- Using a trust region policy optimization method.
- Using a stochastic dueling network architecture.
Papers archive 2025-07-28
12 shown of 12, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Dynamics of Resource Allocation in O-RANs: An In-depth Exploration of On-Policy and Off-Policy Deep Reinforcement Learning for Real-Time Applications 17 Nov 2024 · 0 repositories · arXiv:2412.01839
-
Joint Physical-Digital Facial Attack Detection Via Simulating Spoofing Clues 12 Apr 2024 · 3 repositories · arXiv:2404.08450
-
Distributional Estimation of Data Uncertainty for Surveillance Face Anti-spoofing 18 Sep 2023 · 0 repositories · arXiv:2309.09485
-
PDVN: A Patch-based Dual-view Network for Face Liveness Detection using Light Field Focal Stack 17 Jan 2023 · 0 repositories
-
Asynchronous Curriculum Experience Replay: A Deep Reinforcement Learning Approach for UAV Autonomous Motion Control in Unknown Dynamic Environments 4 Jul 2022 · 0 repositories · arXiv:2207.01251
-
Is Word Error Rate a good evaluation metric for Speech Recognition in Indic Languages? 30 Mar 2022 · 0 repositories · arXiv:2203.16601
-
Learning Reward Machines: A Study in Partially Observable Reinforcement Learning 17 Dec 2021 · 0 repositories · arXiv:2112.09477
-
A-DeepPixBis: Attentional Angular Margin for Face Anti-Spoofing 1 Mar 2021 · 0 repositories
-
Learning Reward Machines for Partially Observable Reinforcement Learning 1 Dec 2019 · 1 repository
-
Sample Efficient Deep Reinforcement Learning for Dialogue Systems with Large Action Spaces 11 Feb 2018 · 0 repositories · arXiv:1802.03753
-
Pretraining Deep Actor-Critic Reinforcement Learning Algorithms With Expert Demonstrations 31 Jan 2018 · 0 repositories · arXiv:1801.10459
-
Sample Efficient Actor-Critic with Experience Replay 3 Nov 2016 · 7 repositories · arXiv:1611.01224Syntology ran 0 of 1 samples · 1 unverified · 1 pointer-only (licence)
Tasks archive 2025-07-28
20 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections