Methods › Computer Vision › Convolutions › SAC › Papers, page 2
Switchable Atrous Convolution
SAC
Papers archive 2025-07-28
archive papers tagged: 168 · with a code link: 69 · where Syntology ran a sample: 20 (17 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (20 of 168 tagged: 17 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 168 of 168, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Performance Comparison of Deep RL Algorithms for Energy Systems Optimal Scheduling 1 Aug 2022 · 1 repository · arXiv:2208.00728
-
Value Function Decomposition for Iterative Design of Reinforcement Learning Agents 24 Jun 2022 · 0 repositories · arXiv:2206.13901
-
Equivariant Reinforcement Learning for Quadrotor UAV 2 Jun 2022 · 0 repositories · arXiv:2206.01233
-
Efficient Reward Poisoning Attacks on Online Deep Reinforcement Learning 30 May 2022 · 1 repository · arXiv:2205.14842Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Spatial Autoregressive Coding for Graph Neural Recommendation 19 May 2022 · 0 repositories · arXiv:2205.09489
-
Building Decision Forest via Deep Reinforcement Learning 1 Apr 2022 · 0 repositories · arXiv:2204.00306
-
MicroRacer: a didactic environment for Deep Reinforcement Learning 20 Mar 2022 · 1 repository · arXiv:2203.10494
-
AI-based Robust Resource Allocation in End-to-End Network Slicing under Demand and CSI Uncertainties 10 Feb 2022 · 0 repositories · arXiv:2202.05131
-
skrl: Modular and Flexible Library for Reinforcement Learning 8 Feb 2022 · 1 repository · arXiv:2202.03825
-
Super-Reparametrizations of Weighted CSPs: Properties and Optimization Perspective 6 Jan 2022 · 0 repositories · arXiv:2201.02018
-
Soft Actor-Critic with Cross-Entropy Policy Optimization 21 Dec 2021 · 1 repository · arXiv:2112.11115
-
Stochastic Planner-Actor-Critic for Unsupervised Deformable Image Registration 14 Dec 2021 · 1 repository · arXiv:2112.07415
-
Segment and Complete: Defending Object Detectors against Adversarial Patch Attacks with Robust Patch Detection 8 Dec 2021 · 1 repository · arXiv:2112.04532Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Target Entropy Annealing for Discrete Soft Actor-Critic 6 Dec 2021 · 0 repositories · arXiv:2112.02852
-
Mastering Atari Games with Limited Data 30 Oct 2021 · 3 repositories · arXiv:2111.00210Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Recurrent Off-policy Baselines for Memory-based Continuous Control 25 Oct 2021 · 1 repository · arXiv:2110.12628
-
Balancing Value Underestimation and Overestimation with Realistic Actor-Critic 19 Oct 2021 · 1 repository · arXiv:2110.09712
-
Continuous Control with Action Quantization from Demonstrations 19 Oct 2021 · 1 repository · arXiv:2110.10149
-
Dropout Q-Functions for Doubly Efficient Reinforcement Learning 5 Oct 2021 · 2 repositories · arXiv:2110.02034Syntology official (archive's flag): 1 ran · 6 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Parallel Actors and Learners: A Framework for Generating Scalable RL Implementations 3 Oct 2021 · 0 repositories · arXiv:2110.01101
-
CausalDyna: Improving Generalization of Dyna-style Reinforcement Learning via Counterfactual-Based Data Augmentation 29 Sep 2021 · 0 repositories
-
Experience Replay More When It's a Key Transition in Deep Reinforcement Learning 29 Sep 2021 · 0 repositories
-
Explanation-Aware Experience Replay in Rule-Dense Environments 29 Sep 2021 · 1 repository · arXiv:2109.14711
-
Faster Reinforcement Learning with Value Target Lower Bounding 29 Sep 2021 · 0 repositories
-
Learning Controllable Elements Oriented Representations for Reinforcement Learning 29 Sep 2021 · 0 repositories
-
Meta Attention For Off-Policy Actor-Critic 29 Sep 2021 · 0 repositories
-
OVD-Explorer: A General Information-theoretic Exploration Approach for Reinforcement Learning 29 Sep 2021 · 0 repositories
-
SPP-RL: State Planning Policy Reinforcement Learning 29 Sep 2021 · 0 repositories
-
Improved Soft Actor-Critic: Mixing Prioritized Off-Policy Samples with On-Policy Experience 24 Sep 2021 · 0 repositories · arXiv:2109.11767
-
Soft Actor-Critic With Integer Actions 17 Sep 2021 · 0 repositories · arXiv:2109.08512
-
Deep Reinforcement Learning at the Edge of the Statistical Precipice 30 Aug 2021 · 3 repositories · arXiv:2108.13264Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 4 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
WAD: A Deep Reinforcement Learning Agent for Urban Autonomous Driving 27 Aug 2021 · 0 repositories · arXiv:2108.12134
-
Value-Based Reinforcement Learning for Continuous Control Robotic Manipulation in Multi-Task Sparse Reward Settings 28 Jul 2021 · 0 repositories · arXiv:2107.13356
-
Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation 24 Jun 2021 · 1 repository · arXiv:2106.13281Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
A Learning-based Optimal Market Bidding Strategy for Price-Maker Energy Storage 4 Jun 2021 · 0 repositories · arXiv:2106.02396
-
Deep Reinforcement Learning-based UAV Navigation and Control: A Soft Actor-Critic with Hindsight Experience Replay Approach 2 Jun 2021 · 0 repositories · arXiv:2106.01016
-
Towards Deeper Deep Reinforcement Learning with Spectral Normalization 2 Jun 2021 · 0 repositories · arXiv:2106.01151
-
Context-Based Soft Actor Critic for Environments with Non-stationary Dynamics 7 May 2021 · 1 repository · arXiv:2105.03310
-
Development of a Soft Actor Critic Deep Reinforcement Learning Approach for Harnessing Energy Flexibility in a Large Office Building 25 Apr 2021 · 0 repositories · arXiv:2104.12125
-
ACERAC: Efficient reinforcement learning in fine time discretization 8 Apr 2021 · 0 repositories · arXiv:2104.04004
-
Co-Adaptation of Algorithmic and Implementational Innovations in Inference-based Deep Reinforcement Learning 31 Mar 2021 · 1 repository · arXiv:2103.17258
-
Investigating Value of Curriculum Reinforcement Learning in Autonomous Driving Under Diverse Road and Weather Conditions 14 Mar 2021 · 0 repositories · arXiv:2103.07903
-
Low-Precision Reinforcement Learning: Running Soft Actor-Critic in Half Precision 26 Feb 2021 · 0 repositories · arXiv:2102.13565
-
Exploring Supervised and Unsupervised Rewards in Machine Translation 22 Feb 2021 · 1 repository · arXiv:2102.11403
-
Multi-Stage Transmission Line Flow Control Using Centralized and Decentralized Reinforcement Learning Agents 16 Feb 2021 · 0 repositories · arXiv:2102.08430
-
Q-Value Weighted Regression: Reinforcement Learning with Limited Data 12 Feb 2021 · 1 repository · arXiv:2102.06782Syntology 0 ran · 1 unverified (of 1 harvested sample)
-
OffCon³: What is state of the art anyway? 27 Jan 2021 · 1 repository · arXiv:2101.11331
-
The Semantic Adjacency Criterion in Time Intervals Mining 11 Jan 2021 · 0 repositories · arXiv:2101.03842
-
CAT-SAC: Soft Actor-Critic with Curiosity-Aware Entropy Temperature 1 Jan 2021 · 0 repositories
-
Deep Coherent Exploration For Continuous Control 1 Jan 2021 · 0 repositories
-
PGPS : Coupling Policy Gradient with Population-based Search 1 Jan 2021 · 0 repositories
-
myGym: Modular Toolkit for Visuomotor Robotic Tasks 21 Dec 2020 · 0 repositories · arXiv:2012.11643
-
Policy Gradient RL Algorithms as Directed Acyclic Graphs 14 Dec 2020 · 1 repository · arXiv:2012.07763
-
Virtual Autonomous Driving with Reinforcement Learning 14 Dec 2020 · 0 repositories
-
OPAC: Opportunistic Actor-Critic 11 Dec 2020 · 0 repositories · arXiv:2012.06555
-
Efficient Reservoir Management through Deep Reinforcement Learning 7 Dec 2020 · 0 repositories · arXiv:2012.03822
-
FinRL: A Deep Reinforcement Learning Library for Automated Stock Trading in Quantitative Finance 19 Nov 2020 · 6 repositories · arXiv:2011.09607
-
Tonic: A Deep Reinforcement Learning Library for Fast Prototyping and Benchmarking 15 Nov 2020 · 1 repository · arXiv:2011.07537Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
MESA: Boost Ensemble Imbalanced Learning with MEta-SAmpler 17 Oct 2020 · 2 repositories · arXiv:2010.08830Syntology official (archive's flag): 3 ran · 15 ran (of which 6 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 1 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 18 harvested samples) · 2 pointer-only (licence)
-
Using Soft Actor-Critic for Low-Level UAV Control 5 Oct 2020 · 1 repository · arXiv:2010.02293
-
A framework for reinforcement learning with autocorrelated actions 10 Sep 2020 · 1 repository · arXiv:2009.04777
-
Measuring the Credibility of Student Attendance Data in Higher Education for Data Mining 1 Sep 2020 · 0 repositories · arXiv:2009.00679
-
Market-making with reinforcement-learning (SAC) 27 Aug 2020 · 2 repositories · arXiv:2008.12275
-
Maximum Mutation Reinforcement Learning for Scalable Control 24 Jul 2020 · 2 repositories · arXiv:2007.13690
-
Predictive Information Accelerates Learning in RL 24 Jul 2020 · 1 repository · arXiv:2007.12401Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Meta-SAC: Auto-tune the Entropy Temperature of Soft Actor-Critic via Metagradient 3 Jul 2020 · 1 repository · arXiv:2007.01932Syntology official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 6 harvested samples)
-
Band-limited Soft Actor Critic Model 19 Jun 2020 · 1 repository · arXiv:2006.11431
-
DetectoRS: Detecting Objects with Recursive Feature Pyramid and Switchable Atrous Convolution 3 Jun 2020 · 6 repositories · arXiv:2006.02334Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)