Methods › General › Regularization › Entropy Regularization › Papers, page 9
Entropy Regularization
Papers archive 2025-07-28
archive papers tagged: 1,128 · with a code link: 451 · where Syntology ran a sample: 156 (129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (156 of 1,128 tagged: 129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument)
Page 9 of 12: papers 801 to 900 of 1,128, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Efficient Out-of-Distribution Detection Using Latent Space of β-VAE for Cyber-Physical Systems 26 Aug 2021 · 0 repositories · arXiv:2108.11800
-
Graph Laplacian Diffusion Localization of Connected and Automated Vehicles 24 Aug 2021 · 0 repositories · arXiv:2108.10678
-
Relative Entropy-Regularized Optimal Transport on a Graph: a new algorithm and an experimental comparison 23 Aug 2021 · 0 repositories · arXiv:2108.10004
-
Settling the Variance of Multi-Agent Policy Gradients 19 Aug 2021 · 1 repository · arXiv:2108.08612
-
End-to-End Urban Driving by Imitating a Reinforcement Learning Coach 18 Aug 2021 · 3 repositories · arXiv:2108.08265Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
KITTI-CARLA: a KITTI-like dataset generated by CARLA Simulator 17 Aug 2021 · 1 repository · arXiv:2109.00892
-
Evaluating the Robustness of Semantic Segmentation for Autonomous Driving against Real-World Adversarial Patch Attacks 13 Aug 2021 · 1 repository · arXiv:2108.06179
-
A general class of surrogate functions for stable and efficient reinforcement learning 12 Aug 2021 · 1 repository · arXiv:2108.05828
-
Capture Uncertainties in Deep Neural Networks for Safe Operation of Autonomous Driving Vehicles 11 Aug 2021 · 0 repositories · arXiv:2108.05118
-
CARLA: A Python Library to Benchmark Algorithmic Recourse and Counterfactual Explanation Algorithms 2 Aug 2021 · 4 repositories · arXiv:2108.00783Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples)
-
Value-Based Reinforcement Learning for Continuous Control Robotic Manipulation in Multi-Task Sparse Reward Settings 28 Jul 2021 · 0 repositories · arXiv:2107.13356
-
MarsExplorer: Exploration of Unknown Terrains via Deep Reinforcement Learning and Procedurally Generated Environments 21 Jul 2021 · 2 repositories · arXiv:2107.09996
-
Improving exploration in policy gradient search: Application to symbolic optimization 19 Jul 2021 · 1 repository · arXiv:2107.09158
-
Is attention to bounding boxes all you need for pedestrian action prediction? 16 Jul 2021 · 0 repositories · arXiv:2107.08031
-
Scalable Optimal Transport in High Dimensions for Graph Distances, Embedding Alignment, and More 14 Jul 2021 · 0 repositories · arXiv:2107.06876
-
Distributed Online Service Coordination Using Deep Reinforcement Learning 7 Jul 2021 · 1 repository
-
Structure-aware reinforcement learning for node-overload protection in mobile edge computing 29 Jun 2021 · 0 repositories · arXiv:2107.01025
-
Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation 24 Jun 2021 · 1 repository · arXiv:2106.13281Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Safe Local Motion Planning With Self-Supervised Freespace Forecasting 19 Jun 2021 · 1 repository
-
Multi-modal Scene-compliant User Intention Estimation in Navigation 13 Jun 2021 · 0 repositories · arXiv:2106.06920
-
Keyframe-Focused Visual Imitation Learning 11 Jun 2021 · 0 repositories · arXiv:2106.06452
-
Learning by Watching 10 Jun 2021 · 0 repositories · arXiv:2106.05966
-
Don't Get Yourself into Trouble! Risk-aware Decision-Making for Autonomous Vehicles 8 Jun 2021 · 0 repositories · arXiv:2106.04625
-
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation 8 Jun 2021 · 0 repositories · arXiv:2106.04096
-
Safe Deep Q-Network for Autonomous Vehicles at Unsignalized Intersection 8 Jun 2021 · 0 repositories · arXiv:2106.04561
-
Average-Reward Reinforcement Learning with Trust Region Methods 7 Jun 2021 · 0 repositories · arXiv:2106.03442
-
Cross-Trajectory Representation Learning for Zero-Shot Generalization in RL 4 Jun 2021 · 1 repository · arXiv:2106.02193Syntology official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Heterogeneous Wasserstein Discrepancy for Incomparable Distributions 4 Jun 2021 · 0 repositories · arXiv:2106.02542
-
Lifetime policy reuse and the importance of task capacity 3 Jun 2021 · 1 repository · arXiv:2106.01741
-
An Entropy Regularization Free Mechanism for Policy-based Reinforcement Learning 1 Jun 2021 · 0 repositories · arXiv:2106.00707
-
Clipping Loops for Sample-Efficient Dialogue Policy Optimisation 1 Jun 2021 · 0 repositories
-
Fast Policy Extragradient Methods for Competitive Games with Entropy Regularization 31 May 2021 · 0 repositories · arXiv:2105.15186
-
Urban Traffic Surveillance (UTS): A fully probabilistic 3D tracking approach based on 2D detections 31 May 2021 · 0 repositories · arXiv:2105.14993
-
Pylot: A Modular Platform for Exploring Latency-Accuracy Tradeoffs in Autonomous Vehicles 30 May 2021 · 1 repository
-
Shaped Policy Search for Evolutionary Strategies using Waypoints 30 May 2021 · 0 repositories · arXiv:2105.14639
-
SBEVNet: End-to-End Deep Stereo Layout Estimation 25 May 2021 · 1 repository · arXiv:2105.11705
-
Omni-supervised Point Cloud Segmentation via Gradual Receptive Field Component Reasoning 21 May 2021 · 2 repositories · arXiv:2105.10203Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
A parallel-network continuous quantitative trading model with GARCH and PPO 8 May 2021 · 0 repositories · arXiv:2105.03625
-
On the Linear convergence of Natural Policy Gradient Algorithm 4 May 2021 · 0 repositories · arXiv:2105.01424
-
Learning to drive from a world on rails 3 May 2021 · 1 repository · arXiv:2105.00636Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Investigating the Impact of Multi-LiDAR Placement on Object Detection for Autonomous Driving 2 May 2021 · 1 repository · arXiv:2105.00373Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Network Space Search for Pareto-Efficient Spaces 22 Apr 2021 · 0 repositories · arXiv:2104.11014
-
Multi-Modal Fusion Transformer for End-to-End Autonomous Driving 19 Apr 2021 · 2 repositories · arXiv:2104.09224Syntology official (archive's flag): 3 ran · 9 ran (of which 7 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
End-to-End Interactive Prediction and Planning with Optical Flow Distillation for Autonomous Driving 18 Apr 2021 · 0 repositories · arXiv:2104.08862
-
End-to-end Keyword Spotting using Neural Architecture Search and Quantization 14 Apr 2021 · 0 repositories · arXiv:2104.06666
-
Building Mental Models through Preview of Autopilot Behaviors 12 Apr 2021 · 0 repositories · arXiv:2104.05470
-
A Bayesian Approach to Reinforcement Learning of Vision-Based Vehicular Control 8 Apr 2021 · 1 repository · arXiv:2104.03807
-
A Reinforcement Learning Environment For Job-Shop Scheduling 8 Apr 2021 · 4 repositories · arXiv:2104.03760
-
Risk-Aware Lane Selection on Highway with Dynamic Obstacles 8 Apr 2021 · 0 repositories · arXiv:2104.04105
-
Progressive extension of reinforcement learning action dimension for asymmetric assembly tasks 6 Apr 2021 · 0 repositories · arXiv:2104.04078
-
Weakly-Supervised Image Semantic Segmentation Using Graph Convolutional Networks 31 Mar 2021 · 1 repository · arXiv:2103.16762
-
Flexible MPC-based Conflict Resolution Using Online Adaptive ADMM 25 Mar 2021 · 0 repositories · arXiv:2103.14118
-
Hierarchical Program-Triggered Reinforcement Learning Agents For Automated Driving 25 Mar 2021 · 0 repositories · arXiv:2103.13861
-
Convex Online Video Frame Subset Selection using Multiple Criteria for Data Efficient Autonomous Driving 24 Mar 2021 · 0 repositories · arXiv:2103.13021
-
Self-Supervised Steering Angle Prediction for Vehicle Control Using Visual Odometry 20 Mar 2021 · 0 repositories · arXiv:2103.11204
-
An Energy-Saving Snake Locomotion Gait Policy Obtained Using Deep Reinforcement Learning 8 Mar 2021 · 0 repositories · arXiv:2103.04511
-
Visual Explanation using Attention Mechanism in Actor-Critic-based Deep Reinforcement Learning 6 Mar 2021 · 0 repositories · arXiv:2103.04067
-
The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games 2 Mar 2021 · 19 repositories · arXiv:2103.01955Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A-DeepPixBis: Attentional Angular Margin for Face Anti-Spoofing 1 Mar 2021 · 0 repositories
-
AutoPreview: A Framework for Autopilot Behavior Understanding 25 Feb 2021 · 0 repositories · arXiv:2102.13034
-
Spatio-Temporal Look-Ahead Trajectory Prediction using Memory Neural Network 24 Feb 2021 · 0 repositories · arXiv:2102.12070
-
On Proximal Policy Optimization's Heavy-tailed Gradients 20 Feb 2021 · 0 repositories · arXiv:2102.10264
-
Combining Events and Frames using Recurrent Asynchronous Multimodal Networks for Monocular Depth Prediction 18 Feb 2021 · 1 repository · arXiv:2102.09320Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Domain Adaptation In Reinforcement Learning Via Latent Unified State Representation 10 Feb 2021 · 1 repository · arXiv:2102.05714Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
A review of motion planning algorithms for intelligent robotics 4 Feb 2021 · 0 repositories · arXiv:2102.02376
-
Affordance-based Reinforcement Learning for Urban Driving 15 Jan 2021 · 0 repositories · arXiv:2101.05970
-
Instance-Aware Predictive Navigation in Multi-Agent Environments 14 Jan 2021 · 1 repository · arXiv:2101.05893
-
A Strong On-Policy Competitor To PPO 1 Jan 2021 · 0 repositories
-
Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms 1 Jan 2021 · 0 repositories
-
Deep Coherent Exploration For Continuous Control 1 Jan 2021 · 0 repositories
-
Fast MNAS: Uncertainty-aware Neural Architecture Search with Lifelong Learning 1 Jan 2021 · 0 repositories
-
Grounded Compositional Generalization with Environment Interactions 1 Jan 2021 · 0 repositories
-
Inverse reinforcement learning for autonomous navigation via differentiable semantic mapping and planning 1 Jan 2021 · 0 repositories · arXiv:2101.00186
-
Optimizing Information Bottleneck in Reinforcement Learning: A Stein Variational Approach 1 Jan 2021 · 0 repositories
-
PGPS : Coupling Policy Gradient with Population-based Search 1 Jan 2021 · 0 repositories
-
Policy Optimization in Zero-Sum Markov Games: Fictitious Self-Play Provably Attains Nash Equilibria 1 Jan 2021 · 0 repositories
-
TMCOSS: Thresholded Multi-Criteria Online Subset Selection for Data-Efficient Autonomous Driving 1 Jan 2021 · 0 repositories
-
Warpspeed Computation of Optimal Transport, Graph Distances, and Embedding Alignment 1 Jan 2021 · 0 repositories
-
Towards Understanding Asynchronous Advantage Actor-critic: Convergence and Linear Speedup 31 Dec 2020 · 0 repositories · arXiv:2012.15511
-
Reconfigurable Intelligent Surface Assisted Mobile Edge Computing with Heterogeneous Learning Tasks 25 Dec 2020 · 0 repositories · arXiv:2012.13533
-
myGym: Modular Toolkit for Visuomotor Robotic Tasks 21 Dec 2020 · 0 repositories · arXiv:2012.11643
-
Multi-Modal Depth Estimation Using Convolutional Neural Networks 17 Dec 2020 · 0 repositories · arXiv:2012.09667
-
CARLA Real Traffic Scenarios -- novel training ground and benchmark for autonomous driving 16 Dec 2020 · 0 repositories · arXiv:2012.11329
-
Increasing Data Efficiency of Driving Agent By World Model 14 Dec 2020 · 1 repository
-
IPM Move Planner: AN EFFICIENT EXPLOITING DEEP REINFORCEMENT LEARNING WITH MONTE CARLO TREE SEARCH 14 Dec 2020 · 0 repositories
-
Mobile Robots Exploration via Deep Reinforcement Learning 14 Dec 2020 · 0 repositories
-
Policy Gradient for items Recommendation on Virtual Taobao 14 Dec 2020 · 0 repositories
-
Policy Gradient RL Algorithms as Directed Acyclic Graphs 14 Dec 2020 · 1 repository · arXiv:2012.07763
-
Train a snake with reinforcement learning algorithms 14 Dec 2020 · 0 repositories
-
Simple Copy-Paste is a Strong Data Augmentation Method for Instance Segmentation 13 Dec 2020 · 5 repositories · arXiv:2012.07177Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
[Re] Reimplementation of FixMatch and Investigation on Noisy (Pseudo) Labels and Confirmation Errors of FixMatch 6 Dec 2020 · 1 repository
-
Proximal Policy Optimization Smoothed Algorithm 4 Dec 2020 · 0 repositories · arXiv:2012.02439
-
Domain Generalization via Entropy Regularization 1 Dec 2020 · 1 repository
-
On the Convergence of Smooth Regularized Approximate Value Iteration Schemes 1 Dec 2020 · 0 repositories
-
Promoting Stochasticity for Expressive Policies via a Simple and Efficient Regularization Method 1 Dec 2020 · 0 repositories
-
An End-to-end Deep Reinforcement Learning Approach for the Long-term Short-term Planning on the Frenet Space 26 Nov 2020 · 1 repository · arXiv:2011.13098
-
Enhanced Scene Specificity with Sparse Dynamic Value Estimation 25 Nov 2020 · 0 repositories · arXiv:2011.12574
-
FinRL: A Deep Reinforcement Learning Library for Automated Stock Trading in Quantitative Finance 19 Nov 2020 · 6 repositories · arXiv:2011.09607
-
AttentiveNAS: Improving Neural Architecture Search via Attentive Sampling 18 Nov 2020 · 2 repositories · arXiv:2011.09011
-
Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge? 18 Nov 2020 · 7 repositories · arXiv:2011.09533