Methods › Reinforcement Learning › Policy Gradient Methods › PPO › Papers, page 6
Proximal Policy Optimization
PPO
Papers archive 2025-07-28
archive papers tagged: 949 · with a code link: 397 · where Syntology ran a sample: 139 (114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (139 of 949 tagged: 114 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument)
Page 6 of 10: papers 501 to 600 of 949, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Implicit Ray-Transformers for Multi-view Remote Sensing Image Segmentation 15 Mar 2023 · 0 repositories · arXiv:2303.08401
-
Reinforcement Learning-based Wavefront Sensorless Adaptive Optics Approaches for Satellite-to-Ground Laser Communication 13 Mar 2023 · 0 repositories · arXiv:2303.07516
-
A Strategy-Oriented Bayesian Soft Actor-Critic Model 7 Mar 2023 · 0 repositories · arXiv:2303.04193
-
A Reinforcement Learning Approach for Scheduling Problems With Improved Generalization Through Order Swapping 27 Feb 2023 · 0 repositories · arXiv:2302.13941
-
SEO: Safety-Aware Energy Optimization Framework for Multi-Sensor Neural Controllers at the Edge 24 Feb 2023 · 0 repositories · arXiv:2302.12493
-
Behavior Proximal Policy Optimization 22 Feb 2023 · 2 repositories · arXiv:2302.11312Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Dynamic Simplex: Balancing Safety and Performance in Autonomous Cyber Physical Systems 20 Feb 2023 · 1 repository · arXiv:2302.09750Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Towards Co-operative Congestion Mitigation 17 Feb 2023 · 0 repositories · arXiv:2302.09140
-
Cooperative Perception for Safe Control of Autonomous Vehicles under LiDAR Spoofing Attacks 14 Feb 2023 · 0 repositories · arXiv:2302.07341
-
EnergyShield: Provably-Safe Offloading of Neural Network Controllers for Energy Efficiency 13 Feb 2023 · 0 repositories · arXiv:2302.06572
-
Shared Information-Based Safe And Efficient Behavior Planning For Connected Autonomous Vehicles 8 Feb 2023 · 0 repositories · arXiv:2302.04321
-
Sample Dropout: A Simple yet Effective Variance Reduction Technique in Deep Policy Optimization 5 Feb 2023 · 1 repository · arXiv:2302.02299Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
STEPS: Joint Self-supervised Nighttime Image Enhancement and Depth Estimation 2 Feb 2023 · 1 repository · arXiv:2302.01334
-
Bridging Physics-Informed Neural Networks with Reinforcement Learning: Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO) 1 Feb 2023 · 0 repositories · arXiv:2302.00237
-
Learning, Fast and Slow: A Goal-Directed Memory-Based Approach for Dynamic Environments 31 Jan 2023 · 1 repository · arXiv:2301.13758
-
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence 30 Jan 2023 · 1 repository · arXiv:2301.13139
-
Incorporating Recurrent Reinforcement Learning into Model Predictive Control for Adaptive Control in Autonomous Driving 30 Jan 2023 · 0 repositories · arXiv:2301.13313
-
SoftTreeMax: Exponential Variance Reduction in Policy Gradient via Tree Search 30 Jan 2023 · 0 repositories · arXiv:2301.13236
-
Joint action loss for proximal policy optimization 26 Jan 2023 · 1 repository · arXiv:2301.10919
-
schlably: A Python Framework for Deep Reinforcement Learning Based Scheduling Experiments 10 Jan 2023 · 1 repository · arXiv:2301.04182
-
Asynchronous Multi-Agent Reinforcement Learning for Efficient Real-Time Multi-Robot Cooperative Exploration 9 Jan 2023 · 2 repositories · arXiv:2301.03398
-
Tuning Path Tracking Controllers for Autonomous Cars Using Reinforcement Learning 9 Jan 2023 · 0 repositories · arXiv:2301.03363
-
e-Inu: Simulating A Quadruped Robot With Emotional Sentience 3 Jan 2023 · 0 repositories · arXiv:2301.00964
-
Deep Reinforcement Learning for Asset Allocation: Reward Clipping 2 Jan 2023 · 0 repositories · arXiv:2301.05300
-
Simoun: Synergizing Interactive Motion-appearance Understanding for Vision-based Reinforcement Learning 1 Jan 2023 · 0 repositories
-
Stabilizing Visual Reinforcement Learning via Asymmetric Interactive Cooperation 1 Jan 2023 · 0 repositories
-
Towards automating Codenames spymasters with deep reinforcement learning 28 Dec 2022 · 0 repositories · arXiv:2212.14104
-
Hierarchical Deep Reinforcement Learning for Age-of-Information Minimization in IRS-aided and Wireless-powered Wireless Networks 27 Dec 2022 · 0 repositories · arXiv:2212.13390
-
Learning Generalizable Representations for Reinforcement Learning via Adaptive Meta-learner of Behavioral Similarities 26 Dec 2022 · 1 repository · arXiv:2212.13088
-
Lifelong Reinforcement Learning with Modulating Masks 21 Dec 2022 · 4 repositories · arXiv:2212.11110
-
Pre-Trained Image Encoder for Generalizable Visual Reinforcement Learning 17 Dec 2022 · 0 repositories · arXiv:2212.08860
-
Distribution-aware Goal Prediction and Conformant Model-based Planning for Safe Autonomous Driving 16 Dec 2022 · 0 repositories · arXiv:2212.08729
-
Learning for Vehicle-to-Vehicle Cooperative Perception under Lossy Communication 16 Dec 2022 · 1 repository · arXiv:2212.08273
-
Multi-Agent Reinforcement Learning with Shared Resources for Inventory Management 15 Dec 2022 · 0 repositories · arXiv:2212.07684
-
Robust Policy Optimization in Deep Reinforcement Learning 14 Dec 2022 · 1 repository · arXiv:2212.07536Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
PPO-UE: Proximal Policy Optimization via Uncertainty-Aware Exploration 13 Dec 2022 · 0 repositories · arXiv:2212.06343
-
Decentralized cooperative perception for autonomous vehicles: Learning to value the unknown 12 Dec 2022 · 0 repositories · arXiv:2301.01250
-
Reinforcement Learning for Molecular Dynamics Optimization: A Stochastic Pontryagin Maximum Principle Approach 6 Dec 2022 · 1 repository · arXiv:2212.03320
-
Safe Reinforcement Learning with Probabilistic Control Barrier Functions for Ramp Merging 1 Dec 2022 · 0 repositories · arXiv:2212.00618
-
Analyzing Infrastructure LiDAR Placement with Realistic LiDAR Simulation Library 29 Nov 2022 · 2 repositories · arXiv:2211.15975Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Combined Peak Reduction and Self-Consumption Using Proximal Policy Optimization 27 Nov 2022 · 0 repositories · arXiv:2211.14831
-
Multi-task Learning for Camera Calibration 22 Nov 2022 · 2 repositories · arXiv:2211.12432
-
Rationale-aware Autonomous Driving Policy utilizing Safety Force Field implemented on CARLA Simulator 18 Nov 2022 · 0 repositories · arXiv:2211.10237
-
Dynamic Conditional Imitation Learning for Autonomous Driving 17 Nov 2022 · 1 repository · arXiv:2211.11579
-
Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization 11 Nov 2022 · 1 repository · arXiv:2211.06236
-
Estimation of Appearance and Occupancy Information in Birds Eye View from Surround Monocular Images 8 Nov 2022 · 0 repositories · arXiv:2211.04557
-
Decentralized Policy Optimization 6 Nov 2022 · 0 repositories · arXiv:2211.03032
-
Design Process is a Reinforcement Learning Problem 6 Nov 2022 · 1 repository · arXiv:2211.03136
-
DeFIX: Detecting and Fixing Failure Scenarios with Reinforcement Learning in Imitation Learning Based Autonomous Driving 29 Oct 2022 · 2 repositories · arXiv:2210.16567
-
Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms 29 Oct 2022 · 2 repositories · arXiv:2210.16575
-
Many-Objective Reinforcement Learning for Online Testing of DNN-Enabled Systems 27 Oct 2022 · 0 repositories · arXiv:2210.15432
-
PlanT: Explainable Planning Transformers via Object-Level Representations 25 Oct 2022 · 2 repositories · arXiv:2210.14222Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Out of Distribution Reasoning by Weakly-Supervised Disentangled Logic Variational Autoencoder 18 Oct 2022 · 0 repositories · arXiv:2210.09959
-
A Multilevel Reinforcement Learning Framework for PDE-based Control 15 Oct 2022 · 2 repositories · arXiv:2210.08400
-
Model-Based Imitation Learning for Urban Driving 14 Oct 2022 · 1 repository · arXiv:2210.07729Syntology official (archive's flag): 21 ran · 21 ran (of which 13 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 23 harvested samples)
-
Exploring Contextual Representation and Multi-Modality for End-to-End Autonomous Driving 13 Oct 2022 · 0 repositories · arXiv:2210.06758
-
Discovered Policy Optimisation 11 Oct 2022 · 1 repository · arXiv:2210.05639
-
Enhance Sample Efficiency and Robustness of End-to-end Urban Autonomous Driving via Semantic Masked World Model 8 Oct 2022 · 0 repositories · arXiv:2210.04017
-
Real-Time Reinforcement Learning for Vision-Based Robotics Utilizing Local and Remote Computers 5 Oct 2022 · 2 repositories · arXiv:2210.02317Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Spatial-Temporal-Aware Safe Multi-Agent Reinforcement Learning of Connected Autonomous Vehicles in Challenging Scenarios 5 Oct 2022 · 0 repositories · arXiv:2210.02300
-
Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy Optimization 3 Oct 2022 · 3 repositories · arXiv:2210.01241
-
Multi-Agent Chance-Constrained Stochastic Shortest Path with Application to Risk-Aware Intelligent Intersection 3 Oct 2022 · 0 repositories · arXiv:2210.01766
-
IPPO: Obstacle Avoidance for Robotic Manipulators in Joint Space via Improved Proximal Policy Optimization 3 Oct 2022 · 0 repositories · arXiv:2210.00803
-
SoftTreeMax: Policy Gradient with Tree Search 28 Sep 2022 · 0 repositories · arXiv:2209.13966
-
Lamarckian Platform: Pushing the Boundaries of Evolutionary Reinforcement Learning towards Asynchronous Commercial Games 21 Sep 2022 · 0 repositories · arXiv:2209.10055
-
Model-Free Reinforcement Learning for Asset Allocation 21 Sep 2022 · 0 repositories · arXiv:2209.10458
-
Experimental Study on The Effect of Multi-step Deep Reinforcement Learning in POMDPs 12 Sep 2022 · 1 repository · arXiv:2209.04999
-
Normality-Guided Distributional Reinforcement Learning for Continuous Control 28 Aug 2022 · 0 repositories · arXiv:2208.13125
-
Entropy Augmented Reinforcement Learning 19 Aug 2022 · 0 repositories · arXiv:2208.09322
-
Path Planning of Cleaning Robot with Reinforcement Learning 17 Aug 2022 · 0 repositories · arXiv:2208.08211
-
Bayesian Soft Actor-Critic: A Directed Acyclic Strategy Graph Based Deep Reinforcement Learning 11 Aug 2022 · 2 repositories · arXiv:2208.06033
-
Aerial Monocular 3D Object Detection 8 Aug 2022 · 1 repository · arXiv:2208.03974
-
Learning to Generalize with Object-centric Agents in the Open World Survival Game Crafter 5 Aug 2022 · 1 repository · arXiv:2208.03374Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Performance Comparison of Deep RL Algorithms for Energy Systems Optimal Scheduling 1 Aug 2022 · 1 repository · arXiv:2208.00728
-
Adaptive Feature Fusion for Cooperative Perception using LiDAR Point Clouds 30 Jul 2022 · 0 repositories · arXiv:2208.00116
-
Solving the vehicle routing problem with deep reinforcement learning 30 Jul 2022 · 0 repositories · arXiv:2208.00202
-
Safety-Enhanced Autonomous Driving Using Interpretable Sensor Fusion Transformer 28 Jul 2022 · 1 repository · arXiv:2207.14024Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Adaptive Decision Making at the Intersection for Autonomous Vehicles Based on Skill Discovery 24 Jul 2022 · 0 repositories · arXiv:2207.11724
-
Synthetic Dataset Generation for Adversarial Machine Learning Research 21 Jul 2022 · 1 repository · arXiv:2207.10719
-
Resolving Copycat Problems in Visual Imitation Learning via Residual Action Prediction 20 Jul 2022 · 0 repositories · arXiv:2207.09705
-
ANTI-CARLA: An Adversarial Testing Framework for Autonomous Vehicles in CARLA 19 Jul 2022 · 1 repository · arXiv:2208.06309
-
ST-P3: End-to-end Vision-based Autonomous Driving via Spatial-Temporal Feature Learning 15 Jul 2022 · 1 repository · arXiv:2207.07601Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Automated Detection of Label Errors in Semantic Segmentation Datasets via Deep Learning and Uncertainty Quantification 13 Jul 2022 · 1 repository · arXiv:2207.06104Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Towards Global Optimality in Cooperative MARL with the Transformation And Distillation Framework 12 Jul 2022 · 0 repositories · arXiv:2207.11143
-
Keep your Distance: Determining Sampling and Distance Thresholds in Machine Learning Monitoring 11 Jul 2022 · 1 repository · arXiv:2207.05078
-
Game State Learning via Game Scene Augmentation 4 Jul 2022 · 0 repositories · arXiv:2207.01289
-
Data generation using simulation technology to improve perception mechanism of autonomous vehicles 1 Jul 2022 · 0 repositories · arXiv:2207.00191
-
Learning mixture of domain-specific experts via disentangled factors for autonomous driving 28 Jun 2022 · 1 repository
-
IBISCape: A Simulated Benchmark for multi-modal SLAM Systems Evaluation in Large-scale Dynamic Environments 27 Jun 2022 · 1 repository · arXiv:2206.13455
-
Fighting Fire with Fire: Avoiding DNN Shortcuts through Priming 22 Jun 2022 · 0 repositories · arXiv:2206.10816
-
Multi-Agent Car Parking using Reinforcement Learning 22 Jun 2022 · 1 repository · arXiv:2206.13338
-
EnvPool: A Highly Parallel Reinforcement Learning Environment Execution Engine 21 Jun 2022 · 3 repositories · arXiv:2206.10558Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Imitate then Transcend: Multi-Agent Optimal Execution with Dual-Window Denoise PPO 21 Jun 2022 · 0 repositories · arXiv:2206.10736
-
Incorporating Voice Instructions in Model-Based Reinforcement Learning for Self-Driving Cars 21 Jun 2022 · 0 repositories · arXiv:2206.10249
-
A Parametric Class of Approximate Gradient Updates for Policy Optimization 17 Jun 2022 · 0 repositories · arXiv:2206.08499
-
Towards Human-Level Bimanual Dexterous Manipulation with Reinforcement Learning 17 Jun 2022 · 1 repository · arXiv:2206.08686
-
Level 2 Autonomous Driving on a Single Device: Diving into the Devils of Openpilot 16 Jun 2022 · 0 repositories · arXiv:2206.08176
-
Trajectory-guided Control Prediction for End-to-end Autonomous Driving: A Simple yet Strong Baseline 16 Jun 2022 · 1 repository · arXiv:2206.08129
-
Learning Task-Independent Game State Representations from Unlabeled Images 13 Jun 2022 · 0 repositories · arXiv:2206.06490
-
A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games 12 Jun 2022 · 3 repositories · arXiv:2206.05825Syntology official: no sample here; runs from other or unrecorded repositories · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 3 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)