Methods › Reinforcement Learning › Policy Gradient Methods › PPO › Papers, page 7
Proximal Policy Optimization
PPO
Papers archive 2025-07-28
archive papers tagged: 949 · with a code link: 397 · where Syntology ran a sample: 139 (114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (139 of 949 tagged: 114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument)
Page 7 of 10: papers 601 to 700 of 949, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CARLA-GeAR: a Dataset Generator for a Systematic Evaluation of Adversarial Robustness of Vision Models 9 Jun 2022 · 1 repository · arXiv:2206.04365
-
Generalized Data Distribution Iteration 7 Jun 2022 · 0 repositories · arXiv:2206.03192
-
GIN: Graph-based Interaction-aware Constraint Policy Optimization for Autonomous Driving 3 Jun 2022 · 1 repository · arXiv:2206.01488
-
On the Choice of Data for Efficient Training and Validation of End-to-End Driving Models 1 Jun 2022 · 0 repositories · arXiv:2206.00608
-
TransFuser: Imitation with Transformer-Based Sensor Fusion for Autonomous Driving 31 May 2022 · 3 repositories · arXiv:2205.15997Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Efficient Reward Poisoning Attacks on Online Deep Reinforcement Learning 30 May 2022 · 1 repository · arXiv:2205.14842Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Risk of Stochastic Systems for Temporal Logic Specifications 28 May 2022 · 0 repositories · arXiv:2205.14523
-
Quark: Controllable Text Generation with Reinforced Unlearning 26 May 2022 · 1 repository · arXiv:2205.13636Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Learning to Drive Using Sparse Imitation Reinforcement Learning 24 May 2022 · 0 repositories · arXiv:2205.12128
-
An Evaluation Study of Intrinsic Motivation Techniques applied to Reinforcement Learning over Hard Exploration Environments 23 May 2022 · 1 repository · arXiv:2205.11184Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Flexible Diffusion Modeling of Long Videos 23 May 2022 · 1 repository · arXiv:2205.11495
-
Generalization, Mayhems and Limits in Recurrent Proximal Policy Optimization 23 May 2022 · 0 repositories · arXiv:2205.11104
-
The Sufficiency of Off-Policyness and Soft Clipping: PPO is still Insufficient according to an Off-Policy Measure 20 May 2022 · 1 repository · arXiv:2205.10047
-
A2C is a special case of PPO 18 May 2022 · 1 repository · arXiv:2205.09123
-
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing 18 May 2022 · 0 repositories · arXiv:2205.08989
-
Policy Distillation with Selective Input Gradient Regularization for Efficient Interpretability 18 May 2022 · 0 repositories · arXiv:2205.08685
-
Qualitative Differences Between Evolutionary Strategies and Reinforcement Learning Methods for Control of Autonomous Agents 16 May 2022 · 0 repositories · arXiv:2205.07592
-
Cliff Diving: Exploring Reward Surfaces in Reinforcement Learning Environments 14 May 2022 · 0 repositories · arXiv:2205.07015Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
NMR: Neural Manifold Representation for Autonomous Driving 11 May 2022 · 0 repositories · arXiv:2205.05551
-
UnrealNAS: Can We Search Neural Architectures with Unreal Data? 4 May 2022 · 0 repositories · arXiv:2205.02162
-
Processing Network Controls via Deep Reinforcement Learning 1 May 2022 · 0 repositories · arXiv:2205.02119
-
Control-Aware Prediction Objectives for Autonomous Driving 28 Apr 2022 · 0 repositories · arXiv:2204.13319
-
KING: Generating Safety-Critical Driving Scenarios for Robust Imitation via Kinematics Gradients 28 Apr 2022 · 1 repository · arXiv:2204.13683Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Task-Induced Representation Learning 25 Apr 2022 · 0 repositories · arXiv:2204.11827
-
SelfD: Self-Learning Large-Scale Driving Policies From the Web 21 Apr 2022 · 0 repositories · arXiv:2204.10320
-
Comparing Deep Reinforcement Learning Algorithms in Two-Echelon Supply Chains 20 Apr 2022 · 1 repository · arXiv:2204.09603
-
SELMA: SEmantic Large-scale Multimodal Acquisitions in Variable Weather, Daytime and Viewpoints 20 Apr 2022 · 0 repositories · arXiv:2204.09788
-
End-to-end Autonomous Driving with Semantic Depth Cloud Mapping and Multi-agent 12 Apr 2022 · 1 repository · arXiv:2204.05513
-
Proximal Policy Optimization Learning based Control of Congested Freeway Traffic 12 Apr 2022 · 0 repositories · arXiv:2204.05627
-
Scale Invariant Semantic Segmentation with RGB-D Fusion 10 Apr 2022 · 0 repositories · arXiv:2204.04679
-
Accelerating Federated Edge Learning via Topology Optimization 1 Apr 2022 · 0 repositories · arXiv:2204.00489
-
Hysteresis-Based RL: Robustifying Reinforcement Learning-based Control Policies via Hybrid Control 1 Apr 2022 · 2 repositories · arXiv:2204.00654
-
Assessing Evolutionary Terrain Generation Methods for Curriculum Reinforcement Learning 29 Mar 2022 · 0 repositories · arXiv:2203.15172
-
Learning from All Vehicles 22 Mar 2022 · 1 repository · arXiv:2203.11934Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Proximal Policy Optimization-based Transmit Beamforming and Phase-shift Design in an IRS-aided ISAC System for the THz Band 21 Mar 2022 · 0 repositories · arXiv:2203.10819
-
MicroRacer: a didactic environment for Deep Reinforcement Learning 20 Mar 2022 · 1 repository · arXiv:2203.10494
-
V2X-ViT: Vehicle-to-Everything Cooperative Perception with Vision Transformer 20 Mar 2022 · 1 repository · arXiv:2203.10638Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
DTA: Physical Camouflage Attacks using Differentiable Transformation Network 18 Mar 2022 · 0 repositories · arXiv:2203.09831
-
Proximal Policy Optimization with Adaptive Threshold for Symmetric Relative Density Ratio 18 Mar 2022 · 0 repositories · arXiv:2203.09809
-
Conquering Ghosts: Relation Learning for Information Reliability Representation and End-to-End Robust Navigation 14 Mar 2022 · 0 repositories · arXiv:2203.09952
-
SynWoodScape: Synthetic Surround-view Fisheye Camera Dataset for Autonomous Driving 9 Mar 2022 · 0 repositories · arXiv:2203.05056
-
MIRROR: Differentiable Deep Social Projection for Assistive Human-Robot Communication 6 Mar 2022 · 1 repository · arXiv:2203.02877
-
AI-aided Traffic Control Scheme for M2M Communications in the Internet of Vehicles 5 Mar 2022 · 0 repositories · arXiv:2204.03504
-
Risk-Aware Scene Sampling for Dynamic Assurance of Autonomous Systems 28 Feb 2022 · 1 repository · arXiv:2202.13510
-
PanoFlow: Learning 360° Optical Flow for Surrounding Temporal Understanding 27 Feb 2022 · 1 repository · arXiv:2202.13388Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Context-Hierarchy Inverse Reinforcement Learning 25 Feb 2022 · 0 repositories · arXiv:2202.12597
-
Consistent Dropout for Policy Gradient Reinforcement Learning 23 Feb 2022 · 0 repositories · arXiv:2202.11818
-
A-Eye: Driving with the Eyes of AI for Corner Case Generation 22 Feb 2022 · 0 repositories · arXiv:2202.10803
-
Don't Touch What Matters: Task-Aware Lipschitz Data Augmentation for Visual Reinforcement Learning 21 Feb 2022 · 1 repository · arXiv:2202.09982
-
Learning a Shield from Catastrophic Action Effects: Never Repeat the Same Mistake 19 Feb 2022 · 0 repositories · arXiv:2202.09516
-
Multi-task Safe Reinforcement Learning for Navigating Intersections in Dense Traffic 19 Feb 2022 · 0 repositories · arXiv:2202.09644
-
CADRE: A Cascade Deep Reinforcement Learning Framework for Vision-based Autonomous Urban Driving 17 Feb 2022 · 1 repository · arXiv:2202.08557Syntology official (archive's flag): 8 ran · 8 ran (of which 7 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
OmniSyn: Synthesizing 360 Videos with Wide-baseline Panoramas 17 Feb 2022 · 0 repositories · arXiv:2202.08752
-
Multi-Modal Fusion for Sensorimotor Coordination in Steering Angle Prediction 11 Feb 2022 · 1 repository · arXiv:2202.05500
-
skrl: Modular and Flexible Library for Reinforcement Learning 8 Feb 2022 · 1 repository · arXiv:2202.03825
-
Simulation-to-Reality domain adaptation for offline 3D object annotation on pointclouds with correlation alignment 6 Feb 2022 · 0 repositories · arXiv:2202.02666
-
AI-as-a-Service Toolkit for Human-Centered Intelligence in Autonomous Driving 3 Feb 2022 · 0 repositories · arXiv:2202.01645
-
A Machine Learning Smartphone-based Sensing for Driver Behavior Classification 1 Feb 2022 · 0 repositories · arXiv:2202.01893
-
Trust Region Bounds for Decentralized PPO Under Non-stationarity 31 Jan 2022 · 0 repositories · arXiv:2202.00082
-
You May Not Need Ratio Clipping in PPO 31 Jan 2022 · 0 repositories · arXiv:2202.00079
-
Ray Based Distributed Autonomous Vehicle Research Platform 18 Jan 2022 · 0 repositories · arXiv:2201.06835
-
Spatiotemporal Costmap Inference for MPC via Deep Inverse Reinforcement Learning 17 Jan 2022 · 0 repositories · arXiv:2201.06539
-
The 37 Implementation Details of Proximal Policy Optimization 17 Jan 2022 · 0 repositories
-
A Study on Mitigating Hard Boundaries of Decision-Tree-based Uncertainty Estimates for AI Models 10 Jan 2022 · 0 repositories · arXiv:2201.03263
-
Mirror Learning: A Unifying Framework of Policy Optimisation 7 Jan 2022 · 1 repository · arXiv:2201.02373
-
DReyeVR: Democratizing Virtual Reality Driving Simulation for Behavioural & Interaction Research 6 Jan 2022 · 2 repositories · arXiv:2201.01931
-
Towards Robustness of Neural Networks 30 Dec 2021 · 0 repositories · arXiv:2112.15188
-
Modified DDPG car-following model with a real-world human driving experience with CARLA simulator 29 Dec 2021 · 0 repositories · arXiv:2112.14602
-
Multiagent Model-based Credit Assignment for Continuous Control 27 Dec 2021 · 0 repositories · arXiv:2112.13937
-
Intelligent Traffic Light via Policy-based Deep Reinforcement Learning 27 Dec 2021 · 1 repository · arXiv:2112.13817
-
Doppler velocity-based algorithm for Clustering and Velocity Estimation of moving objects 24 Dec 2021 · 0 repositories · arXiv:2112.12984
-
Intersection focused Situation Coverage-based Verification and Validation Framework for Autonomous Vehicles Implemented in CARLA 24 Dec 2021 · 1 repository · arXiv:2112.14706
-
Maximum Entropy Population-Based Training for Zero-Shot Human-AI Coordination 22 Dec 2021 · 3 repositories · arXiv:2112.11701
-
A deep reinforcement learning model for predictive maintenance planning of road assets: Integrating LCA and LCCA 20 Dec 2021 · 0 repositories · arXiv:2112.12589
-
Learning Reward Machines: A Study in Partially Observable Reinforcement Learning 17 Dec 2021 · 0 repositories · arXiv:2112.09477
-
Scientific Discovery and the Cost of Measurement -- Balancing Information and Cost in Reinforcement Learning 14 Dec 2021 · 0 repositories · arXiv:2112.07535
-
Real-time Collision Risk Estimation based on Stochastic Reachability Spaces 8 Dec 2021 · 1 repository
-
PTR-PPO: Proximal Policy Optimization with Prioritized Trajectory Replay 7 Dec 2021 · 0 repositories · arXiv:2112.03798
-
Personalized Federated Learning of Driver Prediction Models for Autonomous Driving 2 Dec 2021 · 0 repositories · arXiv:2112.00956
-
Paris-CARLA-3D: A Real and Synthetic Outdoor Point Cloud Dataset for Challenging Tasks in 3D Mapping 22 Nov 2021 · 0 repositories · arXiv:2111.11348
-
Learning Robust Output Control Barrier Functions from Safe Expert Demonstrations 18 Nov 2021 · 1 repository · arXiv:2111.09971
-
GRI: General Reinforced Imitation and its Application to Vision-Based Autonomous Driving 16 Nov 2021 · 0 repositories · arXiv:2111.08575
-
The Pseudo Projection Operator: Applications of Deep Learning to Projection Based Filtering in Non-Trivial Frequency Regimes 13 Nov 2021 · 0 repositories · arXiv:2111.07140
-
A Comparison of Model-Free and Model Predictive Control for Price Responsive Water Heaters 8 Nov 2021 · 0 repositories · arXiv:2111.04689
-
Coordinated Proximal Policy Optimization 7 Nov 2021 · 1 repository · arXiv:2111.04051
-
AI-based Radio Resource Management and Trajectory Design for PD-NOMA Communication in IRS-UAV Assisted Networks 6 Nov 2021 · 0 repositories · arXiv:2111.03869
-
TND-NAS: Towards Non-differentiable Objectives in Progressive Differentiable NAS Framework 6 Nov 2021 · 0 repositories · arXiv:2111.03892
-
Compressing Sensor Data for Remote Assistance of Autonomous Vehicles using Deep Generative Models 5 Nov 2021 · 1 repository · arXiv:2111.03201
-
Improving RNA Secondary Structure Design using Deep Reinforcement Learning 5 Nov 2021 · 0 repositories · arXiv:2111.04504
-
Learning Distilled Collaboration Graph for Multi-Agent Perception 1 Nov 2021 · 2 repositories · arXiv:2111.00643Syntology community repositories only · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples)
-
Object-Aware Regularization for Addressing Causal Confusion in Imitation Learning 27 Oct 2021 · 1 repository · arXiv:2110.14118Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Deep Reinforcement Learning Approach for Audio-based Navigation and Audio Source Localization in Multi-speaker Environments 25 Oct 2021 · 0 repositories · arXiv:2110.12778
-
CIM-PPO:Proximal Policy Optimization with Liu-Correntropy Induced Metric 20 Oct 2021 · 0 repositories · arXiv:2110.10522
-
Generative Adversarial Imitation Learning for End-to-End Autonomous Driving on Urban Environments 16 Oct 2021 · 0 repositories · arXiv:2110.08586
-
Improving the sample-efficiency of neural architecture search with reinforcement learning 13 Oct 2021 · 1 repository · arXiv:2110.06751
-
Navigation In Urban Environments Amongst Pedestrians Using Multi-Objective Deep Reinforcement Learning 11 Oct 2021 · 0 repositories · arXiv:2110.05205
-
Camera Calibration through Camera Projection Loss 7 Oct 2021 · 2 repositories · arXiv:2110.03479
-
Cycle-Consistent World Models for Domain Independent Latent Imagination 2 Oct 2021 · 0 repositories · arXiv:2110.00808
-
Bitcoin Transaction Strategy Construction Based on Deep Reinforcement Learning 30 Sep 2021 · 0 repositories · arXiv:2109.14789
-
Fight fire with fire: countering bad shortcuts in imitation learning with good shortcuts 29 Sep 2021 · 0 repositories