Methods › General › Regularization › Entropy Regularization › Papers, page 7
Entropy Regularization
Papers archive 2025-07-28
archive papers tagged: 1,128 · with a code link: 451 · where Syntology ran a sample: 156 (129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (156 of 1,128 tagged: 129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument)
Page 7 of 12: papers 601 to 700 of 1,128, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
PMR: Prototypical Modal Rebalance for Multimodal Learning 14 Nov 2022 · 0 repositories · arXiv:2211.07089
-
TIER-A: Denoising Learning Framework for Information Extraction 13 Nov 2022 · 0 repositories · arXiv:2211.11527
-
Empirical Risk Minimization with Relative Entropy Regularization 12 Nov 2022 · 0 repositories · arXiv:2211.06617
-
Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization 11 Nov 2022 · 1 repository · arXiv:2211.06236
-
Estimation of Appearance and Occupancy Information in Birds Eye View from Surround Monocular Images 8 Nov 2022 · 0 repositories · arXiv:2211.04557
-
Decentralized Policy Optimization 6 Nov 2022 · 0 repositories · arXiv:2211.03032
-
Design Process is a Reinforcement Learning Problem 6 Nov 2022 · 1 repository · arXiv:2211.03136
-
DeFIX: Detecting and Fixing Failure Scenarios with Reinforcement Learning in Imitation Learning Based Autonomous Driving 29 Oct 2022 · 2 repositories · arXiv:2210.16567
-
Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms 29 Oct 2022 · 2 repositories · arXiv:2210.16575
-
Many-Objective Reinforcement Learning for Online Testing of DNN-Enabled Systems 27 Oct 2022 · 0 repositories · arXiv:2210.15432
-
PlanT: Explainable Planning Transformers via Object-Level Representations 25 Oct 2022 · 2 repositories · arXiv:2210.14222Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Non-Iterative Scribble-Supervised Learning with Pacing Pseudo-Masks for Medical Image Segmentation 20 Oct 2022 · 1 repository · arXiv:2210.10956
-
Out of Distribution Reasoning by Weakly-Supervised Disentangled Logic Variational Autoencoder 18 Oct 2022 · 0 repositories · arXiv:2210.09959
-
A Multilevel Reinforcement Learning Framework for PDE-based Control 15 Oct 2022 · 2 repositories · arXiv:2210.08400
-
Model-Based Imitation Learning for Urban Driving 14 Oct 2022 · 1 repository · arXiv:2210.07729Syntology official (archive's flag): 21 ran · 21 ran (of which 13 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 23 harvested samples)
-
Exploring Contextual Representation and Multi-Modality for End-to-End Autonomous Driving 13 Oct 2022 · 0 repositories · arXiv:2210.06758
-
Point Cloud Scene Completion with Joint Color and Semantic Estimation from Single RGB-D Image 12 Oct 2022 · 0 repositories · arXiv:2210.05891
-
Discovered Policy Optimisation 11 Oct 2022 · 1 repository · arXiv:2210.05639
-
Enhance Sample Efficiency and Robustness of End-to-end Urban Autonomous Driving via Semantic Masked World Model 8 Oct 2022 · 0 repositories · arXiv:2210.04017
-
Real-Time Reinforcement Learning for Vision-Based Robotics Utilizing Local and Remote Computers 5 Oct 2022 · 2 repositories · arXiv:2210.02317Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Spatial-Temporal-Aware Safe Multi-Agent Reinforcement Learning of Connected Autonomous Vehicles in Challenging Scenarios 5 Oct 2022 · 0 repositories · arXiv:2210.02300
-
Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy Optimization 3 Oct 2022 · 3 repositories · arXiv:2210.01241
-
Multi-Agent Chance-Constrained Stochastic Shortest Path with Application to Risk-Aware Intelligent Intersection 3 Oct 2022 · 0 repositories · arXiv:2210.01766
-
IPPO: Obstacle Avoidance for Robotic Manipulators in Joint Space via Improved Proximal Policy Optimization 3 Oct 2022 · 0 repositories · arXiv:2210.00803
-
SoftTreeMax: Policy Gradient with Tree Search 28 Sep 2022 · 0 repositories · arXiv:2209.13966
-
Lamarckian Platform: Pushing the Boundaries of Evolutionary Reinforcement Learning towards Asynchronous Commercial Games 21 Sep 2022 · 0 repositories · arXiv:2209.10055
-
Model-Free Reinforcement Learning for Asset Allocation 21 Sep 2022 · 0 repositories · arXiv:2209.10458
-
Experimental Study on The Effect of Multi-step Deep Reinforcement Learning in POMDPs 12 Sep 2022 · 1 repository · arXiv:2209.04999
-
Normality-Guided Distributional Reinforcement Learning for Continuous Control 28 Aug 2022 · 0 repositories · arXiv:2208.13125
-
Entropy Regularization for Population Estimation 24 Aug 2022 · 0 repositories · arXiv:2208.11747
-
Entropy Augmented Reinforcement Learning 19 Aug 2022 · 0 repositories · arXiv:2208.09322
-
Path Planning of Cleaning Robot with Reinforcement Learning 17 Aug 2022 · 0 repositories · arXiv:2208.08211
-
Bayesian Soft Actor-Critic: A Directed Acyclic Strategy Graph Based Deep Reinforcement Learning 11 Aug 2022 · 2 repositories · arXiv:2208.06033
-
Aerial Monocular 3D Object Detection 8 Aug 2022 · 1 repository · arXiv:2208.03974
-
Learning to Generalize with Object-centric Agents in the Open World Survival Game Crafter 5 Aug 2022 · 1 repository · arXiv:2208.03374Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Performance Comparison of Deep RL Algorithms for Energy Systems Optimal Scheduling 1 Aug 2022 · 1 repository · arXiv:2208.00728
-
Adaptive Feature Fusion for Cooperative Perception using LiDAR Point Clouds 30 Jul 2022 · 0 repositories · arXiv:2208.00116
-
Solving the vehicle routing problem with deep reinforcement learning 30 Jul 2022 · 0 repositories · arXiv:2208.00202
-
Safety-Enhanced Autonomous Driving Using Interpretable Sensor Fusion Transformer 28 Jul 2022 · 1 repository · arXiv:2207.14024Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Adaptive Decision Making at the Intersection for Autonomous Vehicles Based on Skill Discovery 24 Jul 2022 · 0 repositories · arXiv:2207.11724
-
Few-Shot Class-Incremental Learning via Entropy-Regularized Data-Free Replay 22 Jul 2022 · 1 repository · arXiv:2207.11213Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Synthetic Dataset Generation for Adversarial Machine Learning Research 21 Jul 2022 · 1 repository · arXiv:2207.10719
-
Resolving Copycat Problems in Visual Imitation Learning via Residual Action Prediction 20 Jul 2022 · 0 repositories · arXiv:2207.09705
-
ANTI-CARLA: An Adversarial Testing Framework for Autonomous Vehicles in CARLA 19 Jul 2022 · 1 repository · arXiv:2208.06309
-
ST-P3: End-to-end Vision-based Autonomous Driving via Spatial-Temporal Feature Learning 15 Jul 2022 · 1 repository · arXiv:2207.07601Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Automated Detection of Label Errors in Semantic Segmentation Datasets via Deep Learning and Uncertainty Quantification 13 Jul 2022 · 1 repository · arXiv:2207.06104Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Towards Global Optimality in Cooperative MARL with the Transformation And Distillation Framework 12 Jul 2022 · 0 repositories · arXiv:2207.11143
-
Keep your Distance: Determining Sampling and Distance Thresholds in Machine Learning Monitoring 11 Jul 2022 · 1 repository · arXiv:2207.05078
-
Asynchronous Curriculum Experience Replay: A Deep Reinforcement Learning Approach for UAV Autonomous Motion Control in Unknown Dynamic Environments 4 Jul 2022 · 0 repositories · arXiv:2207.01251
-
Game State Learning via Game Scene Augmentation 4 Jul 2022 · 0 repositories · arXiv:2207.01289
-
Data generation using simulation technology to improve perception mechanism of autonomous vehicles 1 Jul 2022 · 0 repositories · arXiv:2207.00191
-
Learning mixture of domain-specific experts via disentangled factors for autonomous driving 28 Jun 2022 · 1 repository
-
IBISCape: A Simulated Benchmark for multi-modal SLAM Systems Evaluation in Large-scale Dynamic Environments 27 Jun 2022 · 1 repository · arXiv:2206.13455
-
Fighting Fire with Fire: Avoiding DNN Shortcuts through Priming 22 Jun 2022 · 0 repositories · arXiv:2206.10816
-
Multi-Agent Car Parking using Reinforcement Learning 22 Jun 2022 · 1 repository · arXiv:2206.13338
-
EnvPool: A Highly Parallel Reinforcement Learning Environment Execution Engine 21 Jun 2022 · 3 repositories · arXiv:2206.10558Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Imitate then Transcend: Multi-Agent Optimal Execution with Dual-Window Denoise PPO 21 Jun 2022 · 0 repositories · arXiv:2206.10736
-
Incorporating Voice Instructions in Model-Based Reinforcement Learning for Self-Driving Cars 21 Jun 2022 · 0 repositories · arXiv:2206.10249
-
A Parametric Class of Approximate Gradient Updates for Policy Optimization 17 Jun 2022 · 0 repositories · arXiv:2206.08499
-
Towards Human-Level Bimanual Dexterous Manipulation with Reinforcement Learning 17 Jun 2022 · 1 repository · arXiv:2206.08686
-
Level 2 Autonomous Driving on a Single Device: Diving into the Devils of Openpilot 16 Jun 2022 · 0 repositories · arXiv:2206.08176
-
Trajectory-guided Control Prediction for End-to-end Autonomous Driving: A Simple yet Strong Baseline 16 Jun 2022 · 1 repository · arXiv:2206.08129
-
Learning Task-Independent Game State Representations from Unlabeled Images 13 Jun 2022 · 0 repositories · arXiv:2206.06490
-
A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games 12 Jun 2022 · 3 repositories · arXiv:2206.05825Syntology official: no sample here; runs from other or unrecorded repositories · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 3 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
CARLA-GeAR: a Dataset Generator for a Systematic Evaluation of Adversarial Robustness of Vision Models 9 Jun 2022 · 1 repository · arXiv:2206.04365
-
Generalized Data Distribution Iteration 7 Jun 2022 · 0 repositories · arXiv:2206.03192
-
GIN: Graph-based Interaction-aware Constraint Policy Optimization for Autonomous Driving 3 Jun 2022 · 1 repository · arXiv:2206.01488
-
Finite-Time Analysis of Entropy-Regularized Neural Natural Actor-Critic Algorithm 2 Jun 2022 · 0 repositories · arXiv:2206.00833
-
On the Choice of Data for Efficient Training and Validation of End-to-End Driving Models 1 Jun 2022 · 0 repositories · arXiv:2206.00608
-
TransFuser: Imitation with Transformer-Based Sensor Fusion for Autonomous Driving 31 May 2022 · 3 repositories · arXiv:2205.15997Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Efficient Reward Poisoning Attacks on Online Deep Reinforcement Learning 30 May 2022 · 1 repository · arXiv:2205.14842Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Risk of Stochastic Systems for Temporal Logic Specifications 28 May 2022 · 0 repositories · arXiv:2205.14523
-
KL-Entropy-Regularized RL with a Generative Model is Minimax Optimal 27 May 2022 · 0 repositories · arXiv:2205.14211
-
Quark: Controllable Text Generation with Reinforced Unlearning 26 May 2022 · 1 repository · arXiv:2205.13636Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Accuracy on In-Domain Samples Matters When Building Out-of-Domain detectors: A Reply to Marek et al. (2021) 24 May 2022 · 1 repository · arXiv:2205.11887
-
Learning to Drive Using Sparse Imitation Reinforcement Learning 24 May 2022 · 0 repositories · arXiv:2205.12128
-
An Evaluation Study of Intrinsic Motivation Techniques applied to Reinforcement Learning over Hard Exploration Environments 23 May 2022 · 1 repository · arXiv:2205.11184Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Flexible Diffusion Modeling of Long Videos 23 May 2022 · 1 repository · arXiv:2205.11495
-
Generalization, Mayhems and Limits in Recurrent Proximal Policy Optimization 23 May 2022 · 0 repositories · arXiv:2205.11104
-
The Sufficiency of Off-Policyness and Soft Clipping: PPO is still Insufficient according to an Off-Policy Measure 20 May 2022 · 1 repository · arXiv:2205.10047
-
A2C is a special case of PPO 18 May 2022 · 1 repository · arXiv:2205.09123
-
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing 18 May 2022 · 0 repositories · arXiv:2205.08989
-
Policy Distillation with Selective Input Gradient Regularization for Efficient Interpretability 18 May 2022 · 0 repositories · arXiv:2205.08685
-
Qualitative Differences Between Evolutionary Strategies and Reinforcement Learning Methods for Control of Autonomous Agents 16 May 2022 · 0 repositories · arXiv:2205.07592
-
Cliff Diving: Exploring Reward Surfaces in Reinforcement Learning Environments 14 May 2022 · 0 repositories · arXiv:2205.07015Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
NMR: Neural Manifold Representation for Autonomous Driving 11 May 2022 · 0 repositories · arXiv:2205.05551
-
UnrealNAS: Can We Search Neural Architectures with Unreal Data? 4 May 2022 · 0 repositories · arXiv:2205.02162
-
Processing Network Controls via Deep Reinforcement Learning 1 May 2022 · 0 repositories · arXiv:2205.02119
-
Loss Function Entropy Regularization for Diverse Decision Boundaries 30 Apr 2022 · 0 repositories · arXiv:2205.00224
-
Control-Aware Prediction Objectives for Autonomous Driving 28 Apr 2022 · 0 repositories · arXiv:2204.13319
-
KING: Generating Safety-Critical Driving Scenarios for Robust Imitation via Kinematics Gradients 28 Apr 2022 · 1 repository · arXiv:2204.13683Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Task-Induced Representation Learning 25 Apr 2022 · 0 repositories · arXiv:2204.11827
-
SelfD: Self-Learning Large-Scale Driving Policies From the Web 21 Apr 2022 · 0 repositories · arXiv:2204.10320
-
Comparing Deep Reinforcement Learning Algorithms in Two-Echelon Supply Chains 20 Apr 2022 · 1 repository · arXiv:2204.09603
-
SELMA: SEmantic Large-scale Multimodal Acquisitions in Variable Weather, Daytime and Viewpoints 20 Apr 2022 · 0 repositories · arXiv:2204.09788
-
End-to-end Autonomous Driving with Semantic Depth Cloud Mapping and Multi-agent 12 Apr 2022 · 1 repository · arXiv:2204.05513
-
Proximal Policy Optimization Learning based Control of Congested Freeway Traffic 12 Apr 2022 · 0 repositories · arXiv:2204.05627
-
RL-CoSeg : A Novel Image Co-Segmentation Algorithm with Deep Reinforcement Learning 12 Apr 2022 · 0 repositories · arXiv:2204.05951
-
Scale Invariant Semantic Segmentation with RGB-D Fusion 10 Apr 2022 · 0 repositories · arXiv:2204.04679
-
Semantic Exploration from Language Abstractions and Pretrained Representations 8 Apr 2022 · 0 repositories · arXiv:2204.05080