Methods › Reinforcement Learning › Policy Gradient Methods › PPO › Papers, page 8
Proximal Policy Optimization
PPO
Papers archive 2025-07-28
archive papers tagged: 949 · with a code link: 397 · where Syntology ran a sample: 139 (114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (139 of 949 tagged: 114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument)
Page 8 of 10: papers 701 to 800 of 949, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Joint Self-Supervised Learning for Vision-based Reinforcement Learning 29 Sep 2021 · 0 repositories
-
P4O: Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization 29 Sep 2021 · 0 repositories
-
A Step Towards Efficient Evaluation of Complex Perception Tasks in Simulation 28 Sep 2021 · 0 repositories · arXiv:2110.02739
-
Fast nonlinear risk assessment for autonomous vehicles using learned conditional probabilistic models of agent futures 21 Sep 2021 · 1 repository · arXiv:2109.09975
-
Stochastic MPC with Multi-modal Predictions for Traffic Intersections 20 Sep 2021 · 0 repositories · arXiv:2109.09792
-
POAR: Efficient Policy Optimization via Online Abstract State Representation Learning 17 Sep 2021 · 1 repository · arXiv:2109.08642
-
OPV2V: An Open Benchmark Dataset and Fusion Pipeline for Perception with Vehicle-to-Vehicle Communication 16 Sep 2021 · 2 repositories · arXiv:2109.07644
-
Border-SegGCN: Improving Semantic Segmentation by Refining the Border Outline using Graph Convolutional Network 11 Sep 2021 · 0 repositories · arXiv:2109.05353
-
NEAT: Neural Attention Fields for End-to-End Autonomous Driving 9 Sep 2021 · 1 repository · arXiv:2109.04456Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
roadscene2vec: A Tool for Extracting and Embedding Road Scene-Graphs 2 Sep 2021 · 1 repository · arXiv:2109.01183
-
Deep Reinforcement Learning at the Edge of the Statistical Precipice 30 Aug 2021 · 3 repositories · arXiv:2108.13264Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 4 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
WAD: A Deep Reinforcement Learning Agent for Urban Autonomous Driving 27 Aug 2021 · 0 repositories · arXiv:2108.12134
-
Efficient Out-of-Distribution Detection Using Latent Space of β-VAE for Cyber-Physical Systems 26 Aug 2021 · 0 repositories · arXiv:2108.11800
-
Graph Laplacian Diffusion Localization of Connected and Automated Vehicles 24 Aug 2021 · 0 repositories · arXiv:2108.10678
-
Settling the Variance of Multi-Agent Policy Gradients 19 Aug 2021 · 1 repository · arXiv:2108.08612
-
End-to-End Urban Driving by Imitating a Reinforcement Learning Coach 18 Aug 2021 · 3 repositories · arXiv:2108.08265Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
KITTI-CARLA: a KITTI-like dataset generated by CARLA Simulator 17 Aug 2021 · 1 repository · arXiv:2109.00892
-
Evaluating the Robustness of Semantic Segmentation for Autonomous Driving against Real-World Adversarial Patch Attacks 13 Aug 2021 · 1 repository · arXiv:2108.06179
-
A general class of surrogate functions for stable and efficient reinforcement learning 12 Aug 2021 · 1 repository · arXiv:2108.05828
-
Capture Uncertainties in Deep Neural Networks for Safe Operation of Autonomous Driving Vehicles 11 Aug 2021 · 0 repositories · arXiv:2108.05118
-
CARLA: A Python Library to Benchmark Algorithmic Recourse and Counterfactual Explanation Algorithms 2 Aug 2021 · 4 repositories · arXiv:2108.00783Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples)
-
Value-Based Reinforcement Learning for Continuous Control Robotic Manipulation in Multi-Task Sparse Reward Settings 28 Jul 2021 · 0 repositories · arXiv:2107.13356
-
MarsExplorer: Exploration of Unknown Terrains via Deep Reinforcement Learning and Procedurally Generated Environments 21 Jul 2021 · 2 repositories · arXiv:2107.09996
-
Is attention to bounding boxes all you need for pedestrian action prediction? 16 Jul 2021 · 0 repositories · arXiv:2107.08031
-
Distributed Online Service Coordination Using Deep Reinforcement Learning 7 Jul 2021 · 1 repository
-
Structure-aware reinforcement learning for node-overload protection in mobile edge computing 29 Jun 2021 · 0 repositories · arXiv:2107.01025
-
Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation 24 Jun 2021 · 1 repository · arXiv:2106.13281Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Safe Local Motion Planning With Self-Supervised Freespace Forecasting 19 Jun 2021 · 1 repository
-
Multi-modal Scene-compliant User Intention Estimation in Navigation 13 Jun 2021 · 0 repositories · arXiv:2106.06920
-
Keyframe-Focused Visual Imitation Learning 11 Jun 2021 · 0 repositories · arXiv:2106.06452
-
Learning by Watching 10 Jun 2021 · 0 repositories · arXiv:2106.05966
-
Don't Get Yourself into Trouble! Risk-aware Decision-Making for Autonomous Vehicles 8 Jun 2021 · 0 repositories · arXiv:2106.04625
-
Safe Deep Q-Network for Autonomous Vehicles at Unsignalized Intersection 8 Jun 2021 · 0 repositories · arXiv:2106.04561
-
Average-Reward Reinforcement Learning with Trust Region Methods 7 Jun 2021 · 0 repositories · arXiv:2106.03442
-
Cross-Trajectory Representation Learning for Zero-Shot Generalization in RL 4 Jun 2021 · 1 repository · arXiv:2106.02193Syntology official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Lifetime policy reuse and the importance of task capacity 3 Jun 2021 · 1 repository · arXiv:2106.01741
-
Clipping Loops for Sample-Efficient Dialogue Policy Optimisation 1 Jun 2021 · 0 repositories
-
Urban Traffic Surveillance (UTS): A fully probabilistic 3D tracking approach based on 2D detections 31 May 2021 · 0 repositories · arXiv:2105.14993
-
Pylot: A Modular Platform for Exploring Latency-Accuracy Tradeoffs in Autonomous Vehicles 30 May 2021 · 1 repository
-
Shaped Policy Search for Evolutionary Strategies using Waypoints 30 May 2021 · 0 repositories · arXiv:2105.14639
-
SBEVNet: End-to-End Deep Stereo Layout Estimation 25 May 2021 · 1 repository · arXiv:2105.11705
-
A parallel-network continuous quantitative trading model with GARCH and PPO 8 May 2021 · 0 repositories · arXiv:2105.03625
-
On the Linear convergence of Natural Policy Gradient Algorithm 4 May 2021 · 0 repositories · arXiv:2105.01424
-
Learning to drive from a world on rails 3 May 2021 · 1 repository · arXiv:2105.00636Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Investigating the Impact of Multi-LiDAR Placement on Object Detection for Autonomous Driving 2 May 2021 · 1 repository · arXiv:2105.00373Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Network Space Search for Pareto-Efficient Spaces 22 Apr 2021 · 0 repositories · arXiv:2104.11014
-
Multi-Modal Fusion Transformer for End-to-End Autonomous Driving 19 Apr 2021 · 2 repositories · arXiv:2104.09224Syntology official (archive's flag): 3 ran · 9 ran (of which 7 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
End-to-End Interactive Prediction and Planning with Optical Flow Distillation for Autonomous Driving 18 Apr 2021 · 0 repositories · arXiv:2104.08862
-
End-to-end Keyword Spotting using Neural Architecture Search and Quantization 14 Apr 2021 · 0 repositories · arXiv:2104.06666
-
Building Mental Models through Preview of Autopilot Behaviors 12 Apr 2021 · 0 repositories · arXiv:2104.05470
-
A Bayesian Approach to Reinforcement Learning of Vision-Based Vehicular Control 8 Apr 2021 · 1 repository · arXiv:2104.03807
-
A Reinforcement Learning Environment For Job-Shop Scheduling 8 Apr 2021 · 4 repositories · arXiv:2104.03760
-
Risk-Aware Lane Selection on Highway with Dynamic Obstacles 8 Apr 2021 · 0 repositories · arXiv:2104.04105
-
Progressive extension of reinforcement learning action dimension for asymmetric assembly tasks 6 Apr 2021 · 0 repositories · arXiv:2104.04078
-
Character Controllers Using Motion VAEs 26 Mar 2021 · 1 repository · arXiv:2103.14274Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Flexible MPC-based Conflict Resolution Using Online Adaptive ADMM 25 Mar 2021 · 0 repositories · arXiv:2103.14118
-
Hierarchical Program-Triggered Reinforcement Learning Agents For Automated Driving 25 Mar 2021 · 0 repositories · arXiv:2103.13861
-
Convex Online Video Frame Subset Selection using Multiple Criteria for Data Efficient Autonomous Driving 24 Mar 2021 · 0 repositories · arXiv:2103.13021
-
Self-Supervised Steering Angle Prediction for Vehicle Control Using Visual Odometry 20 Mar 2021 · 0 repositories · arXiv:2103.11204
-
An Energy-Saving Snake Locomotion Gait Policy Obtained Using Deep Reinforcement Learning 8 Mar 2021 · 0 repositories · arXiv:2103.04511
-
The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games 2 Mar 2021 · 19 repositories · arXiv:2103.01955Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
AutoPreview: A Framework for Autopilot Behavior Understanding 25 Feb 2021 · 0 repositories · arXiv:2102.13034
-
Spatio-Temporal Look-Ahead Trajectory Prediction using Memory Neural Network 24 Feb 2021 · 0 repositories · arXiv:2102.12070
-
On Proximal Policy Optimization's Heavy-tailed Gradients 20 Feb 2021 · 0 repositories · arXiv:2102.10264
-
Combining Events and Frames using Recurrent Asynchronous Multimodal Networks for Monocular Depth Prediction 18 Feb 2021 · 1 repository · arXiv:2102.09320Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Domain Adaptation In Reinforcement Learning Via Latent Unified State Representation 10 Feb 2021 · 1 repository · arXiv:2102.05714Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
A review of motion planning algorithms for intelligent robotics 4 Feb 2021 · 0 repositories · arXiv:2102.02376
-
Fast Concept Mapping: The Emergence of Human Abilities in Artificial Neural Networks when Learning Embodied and Self-Supervised 3 Feb 2021 · 1 repository · arXiv:2102.02153
-
Affordance-based Reinforcement Learning for Urban Driving 15 Jan 2021 · 0 repositories · arXiv:2101.05970
-
Instance-Aware Predictive Navigation in Multi-Agent Environments 14 Jan 2021 · 1 repository · arXiv:2101.05893
-
A Strong On-Policy Competitor To PPO 1 Jan 2021 · 0 repositories
-
Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms 1 Jan 2021 · 0 repositories
-
Deep Coherent Exploration For Continuous Control 1 Jan 2021 · 0 repositories
-
Fast MNAS: Uncertainty-aware Neural Architecture Search with Lifelong Learning 1 Jan 2021 · 0 repositories
-
Inverse reinforcement learning for autonomous navigation via differentiable semantic mapping and planning 1 Jan 2021 · 0 repositories · arXiv:2101.00186
-
Optimizing Information Bottleneck in Reinforcement Learning: A Stein Variational Approach 1 Jan 2021 · 0 repositories
-
PGPS : Coupling Policy Gradient with Population-based Search 1 Jan 2021 · 0 repositories
-
Policy Optimization in Zero-Sum Markov Games: Fictitious Self-Play Provably Attains Nash Equilibria 1 Jan 2021 · 0 repositories
-
TMCOSS: Thresholded Multi-Criteria Online Subset Selection for Data-Efficient Autonomous Driving 1 Jan 2021 · 0 repositories
-
Reconfigurable Intelligent Surface Assisted Mobile Edge Computing with Heterogeneous Learning Tasks 25 Dec 2020 · 0 repositories · arXiv:2012.13533
-
myGym: Modular Toolkit for Visuomotor Robotic Tasks 21 Dec 2020 · 0 repositories · arXiv:2012.11643
-
Multi-Modal Depth Estimation Using Convolutional Neural Networks 17 Dec 2020 · 0 repositories · arXiv:2012.09667
-
CARLA Real Traffic Scenarios -- novel training ground and benchmark for autonomous driving 16 Dec 2020 · 0 repositories · arXiv:2012.11329
-
Increasing Data Efficiency of Driving Agent By World Model 14 Dec 2020 · 1 repository
-
IPM Move Planner: AN EFFICIENT EXPLOITING DEEP REINFORCEMENT LEARNING WITH MONTE CARLO TREE SEARCH 14 Dec 2020 · 0 repositories
-
Mobile Robots Exploration via Deep Reinforcement Learning 14 Dec 2020 · 0 repositories
-
Policy Gradient for items Recommendation on Virtual Taobao 14 Dec 2020 · 0 repositories
-
Policy Gradient RL Algorithms as Directed Acyclic Graphs 14 Dec 2020 · 1 repository · arXiv:2012.07763
-
Train a snake with reinforcement learning algorithms 14 Dec 2020 · 0 repositories
-
Simple Copy-Paste is a Strong Data Augmentation Method for Instance Segmentation 13 Dec 2020 · 5 repositories · arXiv:2012.07177Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Proximal Policy Optimization Smoothed Algorithm 4 Dec 2020 · 0 repositories · arXiv:2012.02439
-
An End-to-end Deep Reinforcement Learning Approach for the Long-term Short-term Planning on the Frenet Space 26 Nov 2020 · 1 repository · arXiv:2011.13098
-
Enhanced Scene Specificity with Sparse Dynamic Value Estimation 25 Nov 2020 · 0 repositories · arXiv:2011.12574
-
Deep reinforcement learning for feedback control in a collective flashing ratchet 20 Nov 2020 · 1 repository · arXiv:2011.10357
-
FinRL: A Deep Reinforcement Learning Library for Automated Stock Trading in Quantitative Finance 19 Nov 2020 · 6 repositories · arXiv:2011.09607
-
AttentiveNAS: Improving Neural Architecture Search via Attentive Sampling 18 Nov 2020 · 2 repositories · arXiv:2011.09011
-
Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge? 18 Nov 2020 · 7 repositories · arXiv:2011.09533
-
Manual-Label Free 3D Detection via An Open-Source Simulator 16 Nov 2020 · 0 repositories · arXiv:2011.07784
-
Tonic: A Deep Reinforcement Learning Library for Fast Prototyping and Benchmarking 15 Nov 2020 · 1 repository · arXiv:2011.07537Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Query-based Targeted Action-Space Adversarial Policies on Deep Reinforcement Learning Agents 13 Nov 2020 · 1 repository · arXiv:2011.07114