Methods › General › Regularization › Entropy Regularization › Papers, page 8
Entropy Regularization
Papers archive 2025-07-28
archive papers tagged: 1,128 · with a code link: 451 · where Syntology ran a sample: 156 (129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (156 of 1,128 tagged: 129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument)
Page 8 of 12: papers 701 to 800 of 1,128, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Accelerating Federated Edge Learning via Topology Optimization 1 Apr 2022 · 0 repositories · arXiv:2204.00489
-
Hysteresis-Based RL: Robustifying Reinforcement Learning-based Control Policies via Hybrid Control 1 Apr 2022 · 2 repositories · arXiv:2204.00654
-
Semi-FairVAE: Semi-supervised Fair Representation Learning with Adversarial Variational Autoencoder 1 Apr 2022 · 0 repositories · arXiv:2204.00536
-
Is Word Error Rate a good evaluation metric for Speech Recognition in Indic Languages? 30 Mar 2022 · 0 repositories · arXiv:2203.16601
-
Assessing Evolutionary Terrain Generation Methods for Curriculum Reinforcement Learning 29 Mar 2022 · 0 repositories · arXiv:2203.15172
-
Your Policy Regularizer is Secretly an Adversary 23 Mar 2022 · 0 repositories · arXiv:2203.12592
-
Learning from All Vehicles 22 Mar 2022 · 1 repository · arXiv:2203.11934Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Linear convergence of a policy gradient method for some finite horizon continuous time control problems 22 Mar 2022 · 0 repositories · arXiv:2203.11758
-
Proximal Policy Optimization-based Transmit Beamforming and Phase-shift Design in an IRS-aided ISAC System for the THz Band 21 Mar 2022 · 0 repositories · arXiv:2203.10819
-
MicroRacer: a didactic environment for Deep Reinforcement Learning 20 Mar 2022 · 1 repository · arXiv:2203.10494
-
V2X-ViT: Vehicle-to-Everything Cooperative Perception with Vision Transformer 20 Mar 2022 · 1 repository · arXiv:2203.10638Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
DTA: Physical Camouflage Attacks using Differentiable Transformation Network 18 Mar 2022 · 0 repositories · arXiv:2203.09831
-
Proximal Policy Optimization with Adaptive Threshold for Symmetric Relative Density Ratio 18 Mar 2022 · 0 repositories · arXiv:2203.09809
-
Conquering Ghosts: Relation Learning for Information Reliability Representation and End-to-End Robust Navigation 14 Mar 2022 · 0 repositories · arXiv:2203.09952
-
SynWoodScape: Synthetic Surround-view Fisheye Camera Dataset for Autonomous Driving 9 Mar 2022 · 0 repositories · arXiv:2203.05056
-
MIRROR: Differentiable Deep Social Projection for Assistive Human-Robot Communication 6 Mar 2022 · 1 repository · arXiv:2203.02877
-
AI-aided Traffic Control Scheme for M2M Communications in the Internet of Vehicles 5 Mar 2022 · 0 repositories · arXiv:2204.03504
-
Risk-Aware Scene Sampling for Dynamic Assurance of Autonomous Systems 28 Feb 2022 · 1 repository · arXiv:2202.13510
-
PanoFlow: Learning 360° Optical Flow for Surrounding Temporal Understanding 27 Feb 2022 · 1 repository · arXiv:2202.13388Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Context-Hierarchy Inverse Reinforcement Learning 25 Feb 2022 · 0 repositories · arXiv:2202.12597
-
Consistent Dropout for Policy Gradient Reinforcement Learning 23 Feb 2022 · 0 repositories · arXiv:2202.11818
-
A-Eye: Driving with the Eyes of AI for Corner Case Generation 22 Feb 2022 · 0 repositories · arXiv:2202.10803
-
A Self-Supervised Descriptor for Image Copy Detection 21 Feb 2022 · 2 repositories · arXiv:2202.10261Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Don't Touch What Matters: Task-Aware Lipschitz Data Augmentation for Visual Reinforcement Learning 21 Feb 2022 · 1 repository · arXiv:2202.09982
-
Learning a Shield from Catastrophic Action Effects: Never Repeat the Same Mistake 19 Feb 2022 · 0 repositories · arXiv:2202.09516
-
Multi-task Safe Reinforcement Learning for Navigating Intersections in Dense Traffic 19 Feb 2022 · 0 repositories · arXiv:2202.09644
-
CADRE: A Cascade Deep Reinforcement Learning Framework for Vision-based Autonomous Urban Driving 17 Feb 2022 · 1 repository · arXiv:2202.08557Syntology official (archive's flag): 8 ran · 8 ran (of which 7 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
OmniSyn: Synthesizing 360 Videos with Wide-baseline Panoramas 17 Feb 2022 · 0 repositories · arXiv:2202.08752
-
Multi-Modal Fusion for Sensorimotor Coordination in Steering Angle Prediction 11 Feb 2022 · 1 repository · arXiv:2202.05500
-
Empirical Risk Minimization with Relative Entropy Regularization: Optimality and Sensitivity Analysis 9 Feb 2022 · 0 repositories · arXiv:2202.04385
-
Revisiting QMIX: Discriminative Credit Assignment by Gradient Entropy Regularization 9 Feb 2022 · 0 repositories · arXiv:2202.04427
-
skrl: Modular and Flexible Library for Reinforcement Learning 8 Feb 2022 · 1 repository · arXiv:2202.03825
-
Simulation-to-Reality domain adaptation for offline 3D object annotation on pointclouds with correlation alignment 6 Feb 2022 · 0 repositories · arXiv:2202.02666
-
AI-as-a-Service Toolkit for Human-Centered Intelligence in Autonomous Driving 3 Feb 2022 · 0 repositories · arXiv:2202.01645
-
A Machine Learning Smartphone-based Sensing for Driver Behavior Classification 1 Feb 2022 · 0 repositories · arXiv:2202.01893
-
Trust Region Bounds for Decentralized PPO Under Non-stationarity 31 Jan 2022 · 0 repositories · arXiv:2202.00082
-
You May Not Need Ratio Clipping in PPO 31 Jan 2022 · 0 repositories · arXiv:2202.00079
-
Do You Need the Entropy Reward (in Practice)? 28 Jan 2022 · 2 repositories · arXiv:2201.12434
-
Neuro-Symbolic Entropy Regularization 25 Jan 2022 · 0 repositories · arXiv:2201.11250
-
Ray Based Distributed Autonomous Vehicle Research Platform 18 Jan 2022 · 0 repositories · arXiv:2201.06835
-
Spatiotemporal Costmap Inference for MPC via Deep Inverse Reinforcement Learning 17 Jan 2022 · 0 repositories · arXiv:2201.06539
-
The 37 Implementation Details of Proximal Policy Optimization 17 Jan 2022 · 0 repositories
-
A Study on Mitigating Hard Boundaries of Decision-Tree-based Uncertainty Estimates for AI Models 10 Jan 2022 · 0 repositories · arXiv:2201.03263
-
Mirror Learning: A Unifying Framework of Policy Optimisation 7 Jan 2022 · 1 repository · arXiv:2201.02373
-
DReyeVR: Democratizing Virtual Reality Driving Simulation for Behavioural & Interaction Research 6 Jan 2022 · 2 repositories · arXiv:2201.01931
-
Towards Robustness of Neural Networks 30 Dec 2021 · 0 repositories · arXiv:2112.15188
-
Modified DDPG car-following model with a real-world human driving experience with CARLA simulator 29 Dec 2021 · 0 repositories · arXiv:2112.14602
-
Multiagent Model-based Credit Assignment for Continuous Control 27 Dec 2021 · 0 repositories · arXiv:2112.13937
-
Intelligent Traffic Light via Policy-based Deep Reinforcement Learning 27 Dec 2021 · 1 repository · arXiv:2112.13817
-
Doppler velocity-based algorithm for Clustering and Velocity Estimation of moving objects 24 Dec 2021 · 0 repositories · arXiv:2112.12984
-
Intersection focused Situation Coverage-based Verification and Validation Framework for Autonomous Vehicles Implemented in CARLA 24 Dec 2021 · 1 repository · arXiv:2112.14706
-
Maximum Entropy Population-Based Training for Zero-Shot Human-AI Coordination 22 Dec 2021 · 3 repositories · arXiv:2112.11701
-
A deep reinforcement learning model for predictive maintenance planning of road assets: Integrating LCA and LCCA 20 Dec 2021 · 0 repositories · arXiv:2112.12589
-
Learning Reward Machines: A Study in Partially Observable Reinforcement Learning 17 Dec 2021 · 0 repositories · arXiv:2112.09477
-
Scientific Discovery and the Cost of Measurement -- Balancing Information and Cost in Reinforcement Learning 14 Dec 2021 · 0 repositories · arXiv:2112.07535
-
Real-time Collision Risk Estimation based on Stochastic Reachability Spaces 8 Dec 2021 · 1 repository
-
PTR-PPO: Proximal Policy Optimization with Prioritized Trajectory Replay 7 Dec 2021 · 0 repositories · arXiv:2112.03798
-
Personalized Federated Learning of Driver Prediction Models for Autonomous Driving 2 Dec 2021 · 0 repositories · arXiv:2112.00956
-
Paris-CARLA-3D: A Real and Synthetic Outdoor Point Cloud Dataset for Challenging Tasks in 3D Mapping 22 Nov 2021 · 0 repositories · arXiv:2111.11348
-
Learning Robust Output Control Barrier Functions from Safe Expert Demonstrations 18 Nov 2021 · 1 repository · arXiv:2111.09971
-
GRI: General Reinforced Imitation and its Application to Vision-Based Autonomous Driving 16 Nov 2021 · 0 repositories · arXiv:2111.08575
-
The Pseudo Projection Operator: Applications of Deep Learning to Projection Based Filtering in Non-Trivial Frequency Regimes 13 Nov 2021 · 0 repositories · arXiv:2111.07140
-
A Comparison of Model-Free and Model Predictive Control for Price Responsive Water Heaters 8 Nov 2021 · 0 repositories · arXiv:2111.04689
-
Coordinated Proximal Policy Optimization 7 Nov 2021 · 1 repository · arXiv:2111.04051
-
AI-based Radio Resource Management and Trajectory Design for PD-NOMA Communication in IRS-UAV Assisted Networks 6 Nov 2021 · 0 repositories · arXiv:2111.03869
-
TND-NAS: Towards Non-differentiable Objectives in Progressive Differentiable NAS Framework 6 Nov 2021 · 0 repositories · arXiv:2111.03892
-
Compressing Sensor Data for Remote Assistance of Autonomous Vehicles using Deep Generative Models 5 Nov 2021 · 1 repository · arXiv:2111.03201
-
Improving RNA Secondary Structure Design using Deep Reinforcement Learning 5 Nov 2021 · 0 repositories · arXiv:2111.04504
-
Understanding Entropic Regularization in GANs 2 Nov 2021 · 0 repositories · arXiv:2111.01387
-
Learning Distilled Collaboration Graph for Multi-Agent Perception 1 Nov 2021 · 2 repositories · arXiv:2111.00643Syntology community repositories only · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples)
-
Object-Aware Regularization for Addressing Causal Confusion in Imitation Learning 27 Oct 2021 · 1 repository · arXiv:2110.14118Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
EnTRPO: Trust Region Policy Optimization Method with Entropy Regularization 26 Oct 2021 · 0 repositories · arXiv:2110.13373
-
Learning Collaborative Policies to Solve NP-hard Routing Problems 26 Oct 2021 · 1 repository · arXiv:2110.13987Syntology official (archive's flag): 6 ran · 6 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
A Deep Reinforcement Learning Approach for Audio-based Navigation and Audio Source Localization in Multi-speaker Environments 25 Oct 2021 · 0 repositories · arXiv:2110.12778
-
A Distributed Deep Reinforcement Learning Technique for Application Placement in Edge and Fog Computing Environments 24 Oct 2021 · 0 repositories · arXiv:2110.12415
-
CIM-PPO:Proximal Policy Optimization with Liu-Correntropy Induced Metric 20 Oct 2021 · 0 repositories · arXiv:2110.10522
-
Generative Adversarial Imitation Learning for End-to-End Autonomous Driving on Urban Environments 16 Oct 2021 · 0 repositories · arXiv:2110.08586
-
Improving the sample-efficiency of neural architecture search with reinforcement learning 13 Oct 2021 · 1 repository · arXiv:2110.06751
-
Navigation In Urban Environments Amongst Pedestrians Using Multi-Objective Deep Reinforcement Learning 11 Oct 2021 · 0 repositories · arXiv:2110.05205
-
Camera Calibration through Camera Projection Loss 7 Oct 2021 · 2 repositories · arXiv:2110.03479
-
The Benefits of Being Categorical Distributional: Uncertainty-aware Regularized Exploration in Reinforcement Learning 7 Oct 2021 · 0 repositories · arXiv:2110.03155
-
A new weakly supervised approach for ALS point cloud semantic segmentation 4 Oct 2021 · 0 repositories · arXiv:2110.01462
-
Cycle-Consistent World Models for Domain Independent Latent Imagination 2 Oct 2021 · 0 repositories · arXiv:2110.00808
-
Bitcoin Transaction Strategy Construction Based on Deep Reinforcement Learning 30 Sep 2021 · 0 repositories · arXiv:2109.14789
-
Fight fire with fire: countering bad shortcuts in imitation learning with good shortcuts 29 Sep 2021 · 0 repositories
-
Generalized Maximum Entropy Reinforcement Learning via Reward Shaping 29 Sep 2021 · 0 repositories
-
Joint Self-Supervised Learning for Vision-based Reinforcement Learning 29 Sep 2021 · 0 repositories
-
P4O: Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization 29 Sep 2021 · 0 repositories
-
Sample Efficient Stochastic Policy Extragradient Algorithm for Zero-Sum Markov Game 29 Sep 2021 · 0 repositories
-
A Step Towards Efficient Evaluation of Complex Perception Tasks in Simulation 28 Sep 2021 · 0 repositories · arXiv:2110.02739
-
Fast nonlinear risk assessment for autonomous vehicles using learned conditional probabilistic models of agent futures 21 Sep 2021 · 1 repository · arXiv:2109.09975
-
Stochastic MPC with Multi-modal Predictions for Traffic Intersections 20 Sep 2021 · 0 repositories · arXiv:2109.09792
-
POAR: Efficient Policy Optimization via Online Abstract State Representation Learning 17 Sep 2021 · 1 repository · arXiv:2109.08642
-
OPV2V: An Open Benchmark Dataset and Fusion Pipeline for Perception with Vehicle-to-Vehicle Communication 16 Sep 2021 · 2 repositories · arXiv:2109.07644
-
Border-SegGCN: Improving Semantic Segmentation by Refining the Border Outline using Graph Convolutional Network 11 Sep 2021 · 0 repositories · arXiv:2109.05353
-
NEAT: Neural Attention Fields for End-to-End Autonomous Driving 9 Sep 2021 · 1 repository · arXiv:2109.04456Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
MACRPO: Multi-Agent Cooperative Recurrent Policy Optimization 2 Sep 2021 · 1 repository · arXiv:2109.00882
-
roadscene2vec: A Tool for Extracting and Embedding Road Scene-Graphs 2 Sep 2021 · 1 repository · arXiv:2109.01183
-
Deep Reinforcement Learning at the Edge of the Statistical Precipice 30 Aug 2021 · 3 repositories · arXiv:2108.13264Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 4 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
WAD: A Deep Reinforcement Learning Agent for Urban Autonomous Driving 27 Aug 2021 · 0 repositories · arXiv:2108.12134