Methods › Reinforcement Learning › Policy Gradient Methods › PPO › Papers, page 10
Proximal Policy Optimization
PPO
Papers archive 2025-07-28
archive papers tagged: 949 · with a code link: 397 · where Syntology ran a sample: 139 (114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (139 of 949 tagged: 114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument)
Page 10 of 10: papers 901 to 949 of 949, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Modified Actor-Critics 2 Jul 2019 · 0 repositories · arXiv:1907.01298
-
Learning Data Augmentation Strategies for Object Detection 26 Jun 2019 · 6 repositories · arXiv:1906.11172Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Neural Proximal/Trust Region Policy Optimization Attains Globally Optimal Policy 25 Jun 2019 · 0 repositories · arXiv:1906.10306
-
Proximal Distilled Evolutionary Reinforcement Learning 24 Jun 2019 · 1 repository · arXiv:1906.09807Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Multimodal End-to-End Autonomous Driving 7 Jun 2019 · 0 repositories · arXiv:1906.03199
-
RL-Based Method for Benchmarking the Adversarial Resilience and Robustness of Deep Reinforcement Learning Policies 3 Jun 2019 · 0 repositories · arXiv:1906.01110
-
Policy Search by Target Distribution Learning for Continuous Control 27 May 2019 · 0 repositories · arXiv:1905.11041
-
Combine PPO with NES to Improve Exploration 23 May 2019 · 0 repositories · arXiv:1905.09492
-
Deep Q-Learning with Q-Matrix Transfer Learning for Novel Fire Evacuation Environment 23 May 2019 · 0 repositories · arXiv:1905.09673
-
Multimodal 3D Object Detection from Simulated Pretraining 19 May 2019 · 2 repositories · arXiv:1905.07754
-
Dimension-Wise Importance Sampling Weight Clipping for Sample-Efficient Reinforcement Learning 7 May 2019 · 1 repository · arXiv:1905.02363Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 1 violated, 12 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 18 pointer-only (licence)
-
Autonomous Air Traffic Controller: A Deep Multi-Agent Reinforcement Learning Approach 2 May 2019 · 0 repositories · arXiv:1905.01303
-
SUPERVISED POLICY UPDATE 1 May 2019 · 1 repository
-
Towards Combining On-Off-Policy Methods for Real-World Applications 24 Apr 2019 · 0 repositories · arXiv:1904.10642
-
Talk Proposal: Towards the Realistic Evaluation of Evasion Attacks using CARLA 18 Apr 2019 · 3 repositories · arXiv:1904.12622
-
Rogue-Gym: A New Challenge for Generalization in Reinforcement Learning 17 Apr 2019 · 2 repositories · arXiv:1904.08129Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ShapeMask: Learning to Segment Novel Objects by Refining Shape Priors 5 Apr 2019 · 1 repository · arXiv:1904.03239
-
Truly Proximal Policy Optimization 19 Mar 2019 · 1 repository · arXiv:1903.07940
-
Sample-Efficient Model-Free Reinforcement Learning with Off-Policy Critics 11 Mar 2019 · 1 repository · arXiv:1903.04193
-
Visual-based Autonomous Driving Deployment from a Stochastic and Uncertainty-aware Perspective 3 Mar 2019 · 1 repository · arXiv:1903.00821
-
Trust Region-Guided Proximal Policy Optimization 29 Jan 2019 · 2 repositories · arXiv:1901.10314
-
Distillation Strategies for Proximal Policy Optimization 23 Jan 2019 · 0 repositories · arXiv:1901.08128
-
On-Policy Trust Region Policy Optimisation with Replay Buffers 18 Jan 2019 · 2 repositories · arXiv:1901.06212
-
Exploring applications of deep reinforcement learning for real-world autonomous driving systems 6 Jan 2019 · 0 repositories · arXiv:1901.01536
-
A Logarithmic Barrier Method For Proximal Policy Optimization 16 Dec 2018 · 0 repositories · arXiv:1812.06502
-
SADA: Semantic Adversarial Diagnostic Attacks for Autonomous Applications 5 Dec 2018 · 1 repository · arXiv:1812.02132
-
Policy Optimization with Model-based Explorations 18 Nov 2018 · 0 repositories · arXiv:1811.07350
-
On the Complexity of Exploration in Goal-Driven Navigation 16 Nov 2018 · 0 repositories · arXiv:1811.06889
-
TrolleyMod v1.0: An Open-Source Simulation and Data-Collection Platform for Ethical Decision Making in Autonomous Vehicles 14 Nov 2018 · 1 repository · arXiv:1811.05594
-
Equivalent Constraints for Two-View Geometry: Pose Solution/Pure Rotation Identification and 3D Reconstruction 13 Oct 2018 · 0 repositories · arXiv:1810.05863
-
NSGA-Net: Neural Architecture Search using Multi-Objective Genetic Algorithm 8 Oct 2018 · 2 repositories · arXiv:1810.03522
-
PPO-CMA: Proximal Policy Optimization with Covariance Matrix Adaptation 5 Oct 2018 · 1 repository · arXiv:1810.02541
-
Reinforcement Learning with Perturbed Rewards 2 Oct 2018 · 1 repository · arXiv:1810.01032Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Real-time Dynamic Object Detection for Autonomous Driving using Prior 3D-Maps 28 Sep 2018 · 0 repositories · arXiv:1809.11036
-
Learning End-to-end Autonomous Driving using Guided Auxiliary Supervision 30 Aug 2018 · 1 repository · arXiv:1808.10393
-
Adversarial Deep Reinforcement Learning in Portfolio Management 29 Aug 2018 · 5 repositories · arXiv:1808.09940
-
Proximal Policy Optimization and its Dynamic Version for Sequence Generation 24 Aug 2018 · 0 repositories · arXiv:1808.07982
-
Policy Optimization With Penalized Point Probability Distance: An Alternative To Proximal Policy Optimization 2 Jul 2018 · 2 repositories · arXiv:1807.00442
-
Conditional Affordance Learning for Driving in Urban Environments 18 Jun 2018 · 1 repository · arXiv:1806.06498Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
Semantic Road Layout Understanding by Generative Adversarial Inpainting 29 May 2018 · 0 repositories · arXiv:1805.11746
-
Supervised Policy Update for Deep Reinforcement Learning 29 May 2018 · 1 repository · arXiv:1805.11706
-
An Adaptive Clipping Approach for Proximal Policy Optimization 17 Apr 2018 · 0 repositories · arXiv:1804.06461
-
Variational Inference for Policy Gradient 21 Feb 2018 · 0 repositories · arXiv:1802.07833
-
An Empirical Analysis of Proximal Policy Optimization with Kronecker-factored Natural Gradients 17 Jan 2018 · 0 repositories · arXiv:1801.05566
-
Exploring Deep Recurrent Models with Reinforcement Learning for Molecule Design 1 Jan 2018 · 0 repositories
-
CARLA: An Open Urban Driving Simulator 10 Nov 2017 · 1 repository · arXiv:1711.03938
-
AMBER: Adaptive Multi-Batch Experience Replay for Continuous Action Control 12 Oct 2017 · 0 repositories · arXiv:1710.04423
-
Learning Transferable Architectures for Scalable Image Recognition 21 Jul 2017 · 17 repositories · arXiv:1707.07012Syntology 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Proximal Policy Optimization Algorithms 20 Jul 2017 · 188 repositories · arXiv:1707.06347Syntology 99 ran (of which 53 constructed an object rather than computing a result; 71 with no instrument failure: 7 honoured, 2 violated, 62 with no contract checked; 28 where Syntology's instrument failed) · 77 unverified (of 176 harvested samples) · 94 pointer-only (licence)