Methods › General › Regularization › Entropy Regularization › Papers, page 10
Entropy Regularization
Papers archive 2025-07-28
archive papers tagged: 1,128 · with a code link: 451 · where Syntology ran a sample: 156 (129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (156 of 1,128 tagged: 129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument)
Page 10 of 12: papers 901 to 1,000 of 1,128, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Manual-Label Free 3D Detection via An Open-Source Simulator 16 Nov 2020 · 0 repositories · arXiv:2011.07784
-
Tonic: A Deep Reinforcement Learning Library for Fast Prototyping and Benchmarking 15 Nov 2020 · 1 repository · arXiv:2011.07537Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Query-based Targeted Action-Space Adversarial Policies on Deep Reinforcement Learning Agents 13 Nov 2020 · 1 repository · arXiv:2011.07114
-
Proximal Policy Optimization via Enhanced Exploration Efficiency 11 Nov 2020 · 0 repositories · arXiv:2011.05525
-
Trajectory Planning for Autonomous Vehicles Using Hierarchical Reinforcement Learning 9 Nov 2020 · 1 repository · arXiv:2011.04752
-
Multimodal Trajectory Prediction via Topological Invariance for Navigation at Uncontrolled Intersections 8 Nov 2020 · 1 repository · arXiv:2011.03894
-
Drafting in Collectible Card Games via Reinforcement Learning 7 Nov 2020 · 1 repository
-
Guided Dialogue Policy Learning without Adversarial Learning in the Loop 1 Nov 2020 · 1 repository
-
PILOT: Efficient Planning by Imitation Learning and Optimisation for Safe Autonomous Driving 1 Nov 2020 · 0 repositories · arXiv:2011.00509
-
A Software Architecture for Autonomous Vehicles: Team LRM-B Entry in the First CARLA Autonomous Driving Challenge 23 Oct 2020 · 0 repositories · arXiv:2010.12598
-
Proximal Policy Gradient: PPO with Policy Gradient 20 Oct 2020 · 0 repositories · arXiv:2010.09933
-
Finding Physical Adversarial Examples for Autonomous Driving with Fast and Differentiable Image Compositing 17 Oct 2020 · 1 repository · arXiv:2010.08844
-
Learning Monocular Dense Depth from Events 16 Oct 2020 · 1 repository · arXiv:2010.08350
-
A Learning Approach to Robot-Agnostic Force-Guided High Precision Assembly 15 Oct 2020 · 0 repositories · arXiv:2010.08052
-
Unsupervised Learning of Depth and Ego-Motion from Cylindrical Panoramic Video with Applications for Virtual Reality 14 Oct 2020 · 1 repository · arXiv:2010.07704
-
LM-Reloc: Levenberg-Marquardt Based Direct Visual Relocalization 13 Oct 2020 · 0 repositories · arXiv:2010.06323
-
Smaller World Models for Reinforcement Learning 12 Oct 2020 · 0 repositories · arXiv:2010.05767
-
Automated Concatenation of Embeddings for Structured Prediction 10 Oct 2020 · 2 repositories · arXiv:2010.05006
-
No MCMC for me: Amortized sampling for fast and stable training of energy-based models 8 Oct 2020 · 1 repository · arXiv:2010.04230
-
Proximal Policy Optimization with Relative Pearson Divergence 7 Oct 2020 · 0 repositories · arXiv:2010.03290
-
Neural Mask Generator: Learning to Generate Adaptive Word Maskings for Language Model Adaptation 6 Oct 2020 · 1 repository · arXiv:2010.02705
-
Entropy Regularization for Mean Field Games with Learning 30 Sep 2020 · 0 repositories · arXiv:2010.00145
-
Revisiting Design Choices in Proximal Policy Optimization 23 Sep 2020 · 1 repository · arXiv:2009.10897Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 11 harvested samples)
-
Regularizing Attention Networks for Anomaly Detection in Visual Question Answering 21 Sep 2020 · 0 repositories · arXiv:2009.10054
-
Phasic Policy Gradient 9 Sep 2020 · 3 repositories · arXiv:2009.04416Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples)
-
Data-Driven Transferred Energy Management Strategy for Hybrid Electric Vehicles via Deep Reinforcement Learning 7 Sep 2020 · 0 repositories · arXiv:2009.03289
-
DRLE: Decentralized Reinforcement Learning at the Edge for Traffic Light Control in the IoV 3 Sep 2020 · 1 repository · arXiv:2009.01502
-
Dynamic Scheduling for Stochastic Edge-Cloud Computing Environments using A3C learning and Residual Recurrent Neural Networks 1 Sep 2020 · 1 repository · arXiv:2009.02186
-
Driving Through Ghosts: Behavioral Cloning with False Positives 29 Aug 2020 · 0 repositories · arXiv:2008.12969
-
On the model-based stochastic value gradient for continuous reinforcement learning 28 Aug 2020 · 1 repository · arXiv:2008.12775
-
Domain Adaptation Through Task Distillation 27 Aug 2020 · 1 repository · arXiv:2008.11911
-
Query Focused Multi-document Summarisation of Biomedical Texts: Macquarie Universiy and the Australian National University at BioASQ8b 27 Aug 2020 · 1 repository
-
Cross-regional oil palm tree counting and detection via multi-level attention domain adaptation network 26 Aug 2020 · 1 repository · arXiv:2008.11505
-
Towards Closing the Sim-to-Real Gap in Collaborative Multi-Robot Deep Reinforcement Learning 18 Aug 2020 · 1 repository · arXiv:2008.07875
-
Reinforced Wasserstein Training for Severity-Aware Semantic Segmentation in Autonomous Driving 11 Aug 2020 · 0 repositories · arXiv:2008.04751
-
Physical Adversarial Attack on Vehicle Detector in the Carla Simulator 31 Jul 2020 · 0 repositories · arXiv:2007.16118
-
Queueing Network Controls via Deep Reinforcement Learning 31 Jul 2020 · 1 repository · arXiv:2008.01644Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Lagrangian Duality in Reinforcement Learning 20 Jul 2020 · 0 repositories · arXiv:2007.09998
-
Fast Global Convergence of Natural Policy Gradient Methods with Entropy Regularization 13 Jul 2020 · 0 repositories · arXiv:2007.06558
-
Maximum Entropy Regularization and Chinese Text Recognition 9 Jul 2020 · 0 repositories · arXiv:2007.04651
-
Learning Implicit Credit Assignment for Cooperative Multi-Agent Reinforcement Learning 6 Jul 2020 · 1 repository · arXiv:2007.02529Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Directional Primitives for Uncertainty-Aware Motion Estimation in Urban Environments 1 Jul 2020 · 0 repositories · arXiv:2007.00161
-
Sample Factory: Egocentric 3D Control from Pixels at 100000 FPS with Asynchronous Reinforcement Learning 21 Jun 2020 · 4 repositories · arXiv:2006.11751Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
An operator view of policy gradient methods 19 Jun 2020 · 0 repositories · arXiv:2006.11266
-
Generalization of Agent Behavior through Explicit Representation of Context 18 Jun 2020 · 0 repositories · arXiv:2006.11305
-
Fine-Tuning DARTS for Image Classification 16 Jun 2020 · 0 repositories · arXiv:2006.09042
-
ShieldNN: A Provably Safe NN Filter for Unsafe NN Controllers 16 Jun 2020 · 0 repositories · arXiv:2006.09564
-
Optimistic Distributionally Robust Policy Optimization 14 Jun 2020 · 1 repository · arXiv:2006.07815Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Bonsai-Net: One-Shot Neural Architecture Search via Differentiable Pruners 12 Jun 2020 · 1 repository · arXiv:2006.09264Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploration by Maximizing Rényi Entropy for Reward-Free RL Framework 11 Jun 2020 · 0 repositories · arXiv:2006.06193
-
Rethinking Pre-training and Self-training 11 Jun 2020 · 2 repositories · arXiv:2006.06882
-
Learning Navigation Costs from Demonstration with Semantic Observations 9 Jun 2020 · 0 repositories · arXiv:2006.05043
-
A Comparison of Self-Play Algorithms Under a Generalized Framework 8 Jun 2020 · 0 repositories · arXiv:2006.04471
-
Fast Synthetic LiDAR Rendering via Spherical UV Unwrapping of Equirectangular Z-Buffer Images 8 Jun 2020 · 0 repositories · arXiv:2006.04345
-
Explaining Autonomous Driving by Learning End-to-End Visual Attention 5 Jun 2020 · 0 repositories · arXiv:2006.03347
-
Single-step deep reinforcement learning for open-loop control of laminar and turbulent flows 4 Jun 2020 · 1 repository · arXiv:2006.02979
-
Diversity Actor-Critic: Sample-Aware Entropy Regularization for Sample-Efficient Exploration 2 Jun 2020 · 1 repository · arXiv:2006.01419
-
Exploring Data Aggregation in Policy Learning for Vision-Based Urban Autonomous Driving 1 Jun 2020 · 1 repository
-
Learning Situational Driving 1 Jun 2020 · 0 repositories
-
Severity-Aware Semantic Segmentation With Reinforced Wasserstein Training 1 Jun 2020 · 0 repositories
-
Fast Risk Assessment for Autonomous Vehicles Using Learned Models of Agent Futures 27 May 2020 · 1 repository · arXiv:2005.13458
-
Dynamic Value Estimation for Single-Task Multi-Scene Reinforcement Learning 25 May 2020 · 0 repositories · arXiv:2005.12254
-
Implementation Matters in Deep Policy Gradients: A Case Study on PPO and TRPO 25 May 2020 · 3 repositories · arXiv:2005.12729Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Mirror Descent Policy Optimization 20 May 2020 · 1 repository · arXiv:2005.09814
-
On the Global Convergence Rates of Softmax Policy Gradient Methods 13 May 2020 · 0 repositories · arXiv:2005.06392
-
Smooth Exploration for Robotic Reinforcement Learning 12 May 2020 · 4 repositories · arXiv:2005.05719Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Learning hierarchical behavior and motion planning for autonomous driving 8 May 2020 · 1 repository · arXiv:2005.03863
-
Generalized Entropy Regularization or: There's Nothing Special about Label Smoothing 2 May 2020 · 0 repositories · arXiv:2005.00820
-
Model-based reinforcement learning for biological sequence design 1 May 2020 · 0 repositories
-
Look at the First Sentence: Position Bias in Question Answering 30 Apr 2020 · 1 repository · arXiv:2004.14602
-
Reinforcement Learning with Augmented Data 30 Apr 2020 · 2 repositories · arXiv:2004.14990Syntology 21 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 12 where Syntology's instrument failed) · 3 unverified (of 24 harvested samples) · 21 pointer-only (licence)
-
Mean-Variance Policy Iteration for Risk-Averse Reinforcement Learning 22 Apr 2020 · 1 repository · arXiv:2004.10888
-
ParkPredict: Motion and Intent Prediction of Vehicles in Parking Lots 21 Apr 2020 · 0 repositories · arXiv:2004.10293
-
Solving the scalarization issues of Advantage-based Reinforcement Learning Algorithms 8 Apr 2020 · 1 repository · arXiv:2004.04120
-
Guided Dialog Policy Learning without Adversarial Learning in the Loop 7 Apr 2020 · 1 repository · arXiv:2004.03267Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Evolving Normalization-Activation Layers 6 Apr 2020 · 8 repositories · arXiv:2004.02967Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Leverage the Average: an Analysis of KL Regularization in RL 31 Mar 2020 · 0 repositories · arXiv:2003.14089
-
MTL-NAS: Task-Agnostic Neural Architecture Search towards General-Purpose Multi-Task Learning 31 Mar 2020 · 1 repository · arXiv:2003.14058Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
Obstacle Avoidance and Navigation Utilizing Reinforcement Learning with Reward Shaping 28 Mar 2020 · 1 repository · arXiv:2003.12863
-
Towards Safer Self-Driving Through Great PAIN (Physically Adversarial Intelligent Networks) 24 Mar 2020 · 1 repository · arXiv:2003.10662
-
Robust Deep Reinforcement Learning against Adversarial Perturbations on State Observations 19 Mar 2020 · 4 repositories · arXiv:2003.08938
-
PFPN: Continuous Control of Physically Simulated Characters using Particle Filtering Policy Network 16 Mar 2020 · 1 repository · arXiv:2003.06959
-
Explore and Exploit with Heterotic Line Bundle Models 10 Mar 2020 · 1 repository · arXiv:2003.04817
-
Fast Online Adaptation in Robotics through Meta-Learning Embeddings of Simulated Priors 10 Mar 2020 · 1 repository · arXiv:2003.04663
-
A machine learning environment for evaluating autonomous driving software 7 Mar 2020 · 0 repositories · arXiv:2003.03576
-
Fully Asynchronous Policy Evaluation in Distributed Reinforcement Learning over Networks 1 Mar 2020 · 0 repositories · arXiv:2003.00433
-
A Self-Tuning Actor-Critic Algorithm 28 Feb 2020 · 0 repositories · arXiv:2002.12928
-
A Visual Communication Map for Multi-Agent Deep Reinforcement Learning 27 Feb 2020 · 0 repositories · arXiv:2002.11882
-
Generalized Product Quantization Network for Semi-supervised Image Retrieval 26 Feb 2020 · 2 repositories · arXiv:2002.11281
-
Reinforcement Learning Framework for Deep Brain Stimulation Study 22 Feb 2020 · 1 repository · arXiv:2002.10948
-
First Order Constrained Optimization in Policy Space 16 Feb 2020 · 2 repositories · arXiv:2002.06506
-
Deep RL Agent for a Real-Time Action Strategy Game 15 Feb 2020 · 1 repository · arXiv:2002.06290
-
Temporal-adaptive Hierarchical Reinforcement Learning 6 Feb 2020 · 0 repositories · arXiv:2002.02080
-
Unsupervised Domain Adaptive Object Detection using Forward-Backward Cyclic Adaptation 3 Feb 2020 · 0 repositories · arXiv:2002.00575
-
Integrating Deep Reinforcement Learning with Model-based Path Planners for Automated Driving 2 Feb 2020 · 1 repository · arXiv:2002.00434Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Brain Metastasis Segmentation Network Trained with Robustness to Annotations with Multiple False Negatives 26 Jan 2020 · 0 repositories · arXiv:2001.09501
-
Interpretable End-to-end Urban Autonomous Driving with Latent Deep Reinforcement Learning 23 Jan 2020 · 4 repositories · arXiv:2001.08726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Continuous-action Reinforcement Learning for Playing Racing Games: Comparing SPG to PPO 15 Jan 2020 · 1 repository · arXiv:2001.05270
-
Intelligent Roundabout Insertion using Deep Reinforcement Learning 3 Jan 2020 · 0 repositories · arXiv:2001.00786
-
Learning Representations in Reinforcement Learning: an Information Bottleneck Approach 1 Jan 2020 · 0 repositories