Methods › General › Regularization › Entropy Regularization › Papers, page 6
Entropy Regularization
Papers archive 2025-07-28
archive papers tagged: 1,128 · with a code link: 451 · where Syntology ran a sample: 156 (129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (156 of 1,128 tagged: 129 with a run with no instrument failure, 27 where every run was a failure of Syntology's instrument)
Page 6 of 12: papers 501 to 600 of 1,128, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Learning to Generate Better Than Your LLM 20 Jun 2023 · 1 repository · arXiv:2306.11816
-
Safe, Efficient, Comfort, and Energy-saving Automated Driving through Roundabout Based on Deep Reinforcement Learning 20 Jun 2023 · 0 repositories · arXiv:2306.11465
-
A Study on Quantifying Sim2Real Image Gap in Autonomous Driving Simulations Using Lane Segmentation Attention Map Similarity 18 Jun 2023 · 0 repositories · arXiv:2306.10491
-
Coaching a Teachable Student 16 Jun 2023 · 1 repository · arXiv:2306.10014Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples)
-
Hidden Biases of End-to-End Driving Models 13 Jun 2023 · 1 repository · arXiv:2306.07957Syntology official (archive's flag): 13 ran · 13 ran (of which 10 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples)
-
Analysis of the Relative Entropy Asymmetry in the Regularization of Empirical Risk Minimization 12 Jun 2023 · 0 repositories · arXiv:2306.07123
-
Enhancing Topic Extraction in Recommender Systems with Entropy Regularization 12 Jun 2023 · 0 repositories · arXiv:2306.07403
-
RLtools: A Fast, Portable Deep Reinforcement Learning Library for Continuous Control 6 Jun 2023 · 1 repository · arXiv:2306.03530
-
Fine-Tuning Language Models with Advantage-Induced Policy Alignment 4 Jun 2023 · 1 repository · arXiv:2306.02231Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Deep Q-Learning versus Proximal Policy Optimization: Performance Comparison in a Material Sorting Task 2 Jun 2023 · 0 repositories · arXiv:2306.01451
-
ReLU to the Rescue: Improve Your On-Policy Actor-Critic with Positive Advantages 2 Jun 2023 · 1 repository · arXiv:2306.01460
-
Identifiability and Generalizability in Constrained Inverse Reinforcement Learning 1 Jun 2023 · 1 repository · arXiv:2306.00629Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
Normalization Enhances Generalization in Visual Reinforcement Learning 1 Jun 2023 · 1 repository · arXiv:2306.00656Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Latent Exploration for Reinforcement Learning 31 May 2023 · 1 repository · arXiv:2305.20065
-
Exploring the Promise and Limits of Real-Time Recurrent Learning 30 May 2023 · 1 repository · arXiv:2305.19044
-
DoMo-AC: Doubly Multi-step Off-policy Actor-Critic Algorithm 29 May 2023 · 0 repositories · arXiv:2305.18501
-
Resilience in Platoons of Cooperative Heterogeneous Vehicles: Self-organization Strategies and Provably-correct Design 27 May 2023 · 0 repositories · arXiv:2305.17443
-
All Points Matter: Entropy-Regularized Distribution Alignment for Weakly-supervised 3D Segmentation 25 May 2023 · 1 repository · arXiv:2305.15832Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 13 where Syntology's instrument failed) · 17 unverified (of 33 harvested samples)
-
Generating Synergistic Formulaic Alpha Collections via Reinforcement Learning 25 May 2023 · 1 repository · arXiv:2306.12964
-
Realistically distributing object placements in synthetic training data improves the performance of vision-based object detection models 24 May 2023 · 1 repository · arXiv:2305.14621
-
AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback 22 May 2023 · 2 repositories · arXiv:2305.14387
-
Learning Pedestrian Actions to Ensure Safe Autonomous Driving 22 May 2023 · 0 repositories · arXiv:2305.13051
-
Actor-Critic Methods using Physics-Informed Neural Networks: Control of a 1D PDE Model for Fluid-Cooled Battery Packs 18 May 2023 · 1 repository · arXiv:2305.10952
-
Sharing Lifelong Reinforcement Learning Knowledge via Modulating Masks 18 May 2023 · 2 repositories · arXiv:2305.10997
-
ReasonNet: End-to-End Driving with Temporal and Global Reasoning 17 May 2023 · 0 repositories · arXiv:2305.10507
-
SLiC-HF: Sequence Likelihood Calibration with Human Feedback 17 May 2023 · 0 repositories · arXiv:2305.10425
-
A Theoretical Analysis of Optimistic Proximal Policy Optimization in Linear Markov Decision Processes 15 May 2023 · 0 repositories · arXiv:2305.08841
-
Dynamically Conservative Self-Driving Planner for Long-Tail Cases 12 May 2023 · 0 repositories · arXiv:2305.07497
-
Policy Gradient Algorithms Implicitly Optimize by Continuation 11 May 2023 · 0 repositories · arXiv:2305.06851
-
Think Twice before Driving: Towards Scalable Decoders for End-to-End Autonomous Driving 10 May 2023 · 1 repository · arXiv:2305.06242Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Reducing the Cost of Cycle-Time Tuning for Real-World Policy Optimization 9 May 2023 · 1 repository · arXiv:2305.05760Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Local Optimization Achieves Global Optimality in Multi-Agent Reinforcement Learning 8 May 2023 · 1 repository · arXiv:2305.04819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
CARLA-BSP: a simulated dataset with pedestrians 29 Apr 2023 · 2 repositories · arXiv:2305.00204
-
Adversarial Policy Optimization in Deep Reinforcement Learning 27 Apr 2023 · 0 repositories · arXiv:2304.14533
-
Can Agents Run Relay Race with Strangers? Generalization of RL to Out-of-Distribution Trajectories 26 Apr 2023 · 0 repositories · arXiv:2304.13424
-
An End-to-End Vehicle Trajcetory Prediction Framework 19 Apr 2023 · 0 repositories · arXiv:2304.09764
-
Bridging RL Theory and Practice with the Effective Horizon 19 Apr 2023 · 1 repository · arXiv:2304.09853
-
Benchmarking the Physical-world Adversarial Robustness of Vehicle Detection 11 Apr 2023 · 0 repositories · arXiv:2304.05098
-
RRHF: Rank Responses to Align Language Models with Human Feedback without tears 11 Apr 2023 · 1 repository · arXiv:2304.05302Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
LANe: Lighting-Aware Neural Fields for Compositional Scene Synthesis 6 Apr 2023 · 0 repositories · arXiv:2304.03280
-
AutoRL Hyperparameter Landscapes 5 Apr 2023 · 1 repository · arXiv:2304.02396Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
PAC-Based Formal Verification for Out-of-Distribution Data Detection 4 Apr 2023 · 0 repositories · arXiv:2304.01592
-
Understanding Reinforcement Learning Algorithms: The Progress from Basic Q-learning to Proximal Policy Optimization 31 Mar 2023 · 0 repositories · arXiv:2304.00026
-
Specification-Guided Data Aggregation for Semantically Aware Imitation Learning 29 Mar 2023 · 0 repositories · arXiv:2303.17010
-
Model-Based Reinforcement Learning with Isolated Imaginations 27 Mar 2023 · 1 repository · arXiv:2303.14889
-
Implicit Ray-Transformers for Multi-view Remote Sensing Image Segmentation 15 Mar 2023 · 0 repositories · arXiv:2303.08401
-
Reinforcement Learning-based Wavefront Sensorless Adaptive Optics Approaches for Satellite-to-Ground Laser Communication 13 Mar 2023 · 0 repositories · arXiv:2303.07516
-
Twin Contrastive Learning with Noisy Labels 13 Mar 2023 · 1 repository · arXiv:2303.06930
-
Self-NeRF: A Self-Training Pipeline for Few-Shot Neural Radiance Fields 10 Mar 2023 · 0 repositories · arXiv:2303.05775
-
Intriguing Property and Counterfactual Explanation of GAN for Remote Sensing Image Generation 9 Mar 2023 · 1 repository · arXiv:2303.05240
-
A Strategy-Oriented Bayesian Soft Actor-Critic Model 7 Mar 2023 · 0 repositories · arXiv:2303.04193
-
Uncoupled and Convergent Learning in Two-Player Zero-Sum Markov Games with Bandit Feedback 5 Mar 2023 · 0 repositories · arXiv:2303.02738
-
Double A3C: Deep Reinforcement Learning on OpenAI Gym Games 4 Mar 2023 · 0 repositories · arXiv:2303.02271
-
A Reinforcement Learning Approach for Scheduling Problems With Improved Generalization Through Order Swapping 27 Feb 2023 · 0 repositories · arXiv:2302.13941
-
SEO: Safety-Aware Energy Optimization Framework for Multi-Sensor Neural Controllers at the Edge 24 Feb 2023 · 0 repositories · arXiv:2302.12493
-
Behavior Proximal Policy Optimization 22 Feb 2023 · 2 repositories · arXiv:2302.11312Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Dynamic Simplex: Balancing Safety and Performance in Autonomous Cyber Physical Systems 20 Feb 2023 · 1 repository · arXiv:2302.09750Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Towards Co-operative Congestion Mitigation 17 Feb 2023 · 0 repositories · arXiv:2302.09140
-
Cooperative Perception for Safe Control of Autonomous Vehicles under LiDAR Spoofing Attacks 14 Feb 2023 · 0 repositories · arXiv:2302.07341
-
EnergyShield: Provably-Safe Offloading of Neural Network Controllers for Energy Efficiency 13 Feb 2023 · 0 repositories · arXiv:2302.06572
-
Shared Information-Based Safe And Efficient Behavior Planning For Connected Autonomous Vehicles 8 Feb 2023 · 0 repositories · arXiv:2302.04321
-
Sample Dropout: A Simple yet Effective Variance Reduction Technique in Deep Policy Optimization 5 Feb 2023 · 1 repository · arXiv:2302.02299Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
STEPS: Joint Self-supervised Nighttime Image Enhancement and Depth Estimation 2 Feb 2023 · 1 repository · arXiv:2302.01334
-
Bridging Physics-Informed Neural Networks with Reinforcement Learning: Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO) 1 Feb 2023 · 0 repositories · arXiv:2302.00237
-
Learning, Fast and Slow: A Goal-Directed Memory-Based Approach for Dynamic Environments 31 Jan 2023 · 1 repository · arXiv:2301.13758
-
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence 30 Jan 2023 · 1 repository · arXiv:2301.13139
-
Fast Computation of Optimal Transport via Entropy-Regularized Extragradient Methods 30 Jan 2023 · 0 repositories · arXiv:2301.13006
-
Incorporating Recurrent Reinforcement Learning into Model Predictive Control for Adaptive Control in Autonomous Driving 30 Jan 2023 · 0 repositories · arXiv:2301.13313
-
SoftTreeMax: Exponential Variance Reduction in Policy Gradient via Tree Search 30 Jan 2023 · 0 repositories · arXiv:2301.13236
-
Joint action loss for proximal policy optimization 26 Jan 2023 · 1 repository · arXiv:2301.10919
-
PDVN: A Patch-based Dual-view Network for Face Liveness Detection using Light Field Focal Stack 17 Jan 2023 · 0 repositories
-
Asynchronous Multi-Agent Reinforcement Learning for Efficient Real-Time Multi-Robot Cooperative Exploration 9 Jan 2023 · 2 repositories · arXiv:2301.03398
-
Tuning Path Tracking Controllers for Autonomous Cars Using Reinforcement Learning 9 Jan 2023 · 0 repositories · arXiv:2301.03363
-
e-Inu: Simulating A Quadruped Robot With Emotional Sentience 3 Jan 2023 · 0 repositories · arXiv:2301.00964
-
Deep Reinforcement Learning for Asset Allocation: Reward Clipping 2 Jan 2023 · 0 repositories · arXiv:2301.05300
-
Simoun: Synergizing Interactive Motion-appearance Understanding for Vision-based Reinforcement Learning 1 Jan 2023 · 0 repositories
-
Stabilizing Visual Reinforcement Learning via Asymmetric Interactive Cooperation 1 Jan 2023 · 0 repositories
-
Towards automating Codenames spymasters with deep reinforcement learning 28 Dec 2022 · 0 repositories · arXiv:2212.14104
-
Hierarchical Deep Reinforcement Learning for Age-of-Information Minimization in IRS-aided and Wireless-powered Wireless Networks 27 Dec 2022 · 0 repositories · arXiv:2212.13390
-
Learning Generalizable Representations for Reinforcement Learning via Adaptive Meta-learner of Behavioral Similarities 26 Dec 2022 · 1 repository · arXiv:2212.13088
-
Alignment Entropy Regularization 22 Dec 2022 · 0 repositories · arXiv:2212.12442
-
Lifelong Reinforcement Learning with Modulating Masks 21 Dec 2022 · 4 repositories · arXiv:2212.11110
-
Pre-Trained Image Encoder for Generalizable Visual Reinforcement Learning 17 Dec 2022 · 0 repositories · arXiv:2212.08860
-
Distribution-aware Goal Prediction and Conformant Model-based Planning for Safe Autonomous Driving 16 Dec 2022 · 0 repositories · arXiv:2212.08729
-
Learning for Vehicle-to-Vehicle Cooperative Perception under Lossy Communication 16 Dec 2022 · 1 repository · arXiv:2212.08273
-
Multi-Agent Reinforcement Learning with Shared Resources for Inventory Management 15 Dec 2022 · 0 repositories · arXiv:2212.07684
-
Robust Policy Optimization in Deep Reinforcement Learning 14 Dec 2022 · 1 repository · arXiv:2212.07536Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
PPO-UE: Proximal Policy Optimization via Uncertainty-Aware Exploration 13 Dec 2022 · 0 repositories · arXiv:2212.06343
-
Decentralized cooperative perception for autonomous vehicles: Learning to value the unknown 12 Dec 2022 · 0 repositories · arXiv:2301.01250
-
Reinforcement Learning for Molecular Dynamics Optimization: A Stochastic Pontryagin Maximum Principle Approach 6 Dec 2022 · 1 repository · arXiv:2212.03320
-
Resilience Evaluation of Entropy Regularized Logistic Networks with Probabilistic Cost 5 Dec 2022 · 0 repositories · arXiv:2212.02060
-
Safe Reinforcement Learning with Probabilistic Control Barrier Functions for Ramp Merging 1 Dec 2022 · 0 repositories · arXiv:2212.00618
-
Handling Missing Data via Max-Entropy Regularized Graph Autoencoder 30 Nov 2022 · 0 repositories · arXiv:2211.16771
-
Analyzing Infrastructure LiDAR Placement with Realistic LiDAR Simulation Library 29 Nov 2022 · 2 repositories · arXiv:2211.15975Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Combined Peak Reduction and Self-Consumption Using Proximal Policy Optimization 27 Nov 2022 · 0 repositories · arXiv:2211.14831
-
Homology-constrained vector quantization entropy regularizer 25 Nov 2022 · 1 repository · arXiv:2211.14363Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Mutual Information Learned Regressor: an Information-theoretic Viewpoint of Training Regression Systems 23 Nov 2022 · 0 repositories · arXiv:2211.12685
-
Multi-task Learning for Camera Calibration 22 Nov 2022 · 2 repositories · arXiv:2211.12432
-
Rationale-aware Autonomous Driving Policy utilizing Safety Force Field implemented on CARLA Simulator 18 Nov 2022 · 0 repositories · arXiv:2211.10237
-
Dynamic Conditional Imitation Learning for Autonomous Driving 17 Nov 2022 · 1 repository · arXiv:2211.11579