Methods › General › Skip Connections › Residual Connection › Papers, page 208
Residual Connection
Papers archive 2025-07-28
archive papers tagged: 28,401 · with a code link: 12,847 · where Syntology ran a sample: 3,897 (3,291 with a run with no instrument failure, 606 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,897 of 28,401 tagged: 3,291 with a run with no instrument failure, 606 where every run was a failure of Syntology's instrument)
Page 208 of 285: papers 20,701 to 20,800 of 28,401, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Contrastive Quant: Quantization Makes Stronger Contrastive Learning 29 Sep 2021 · 0 repositories
-
Convolutional Neural Network Compression through Generalized Kronecker Product Decomposition 29 Sep 2021 · 0 repositories · arXiv:2109.14710
-
Cross-Architecture Distillation Using Bidirectional CMOW Embeddings 29 Sep 2021 · 0 repositories
-
Crossformer: Transformer with Alternated Cross-Layer Guidance 29 Sep 2021 · 0 repositories
-
D²ETR: Decoder-Only DETR with Computationally Efficient Cross-Scale Attention 29 Sep 2021 · 0 repositories
-
Disentangling Properties of Contrastive Methods 29 Sep 2021 · 0 repositories
-
Distributional Decision Transformer for Hindsight Information Matching 29 Sep 2021 · 0 repositories
-
Distributional Generalization: Structure Beyond Test Error 29 Sep 2021 · 0 repositories
-
Efficient Packing: Towards 2x NLP Speed-Up without Loss of Accuracy for BERT 29 Sep 2021 · 0 repositories
-
Efficient Point Transformer for Large-scale 3D Scene Understanding 29 Sep 2021 · 0 repositories
-
Embedding models through the lens of Stable Coloring 29 Sep 2021 · 0 repositories
-
Equivariant Transformers for Neural Network based Molecular Potentials 29 Sep 2021 · 0 repositories
-
ERNIE-SPARSE: Robust Efficient Transformer Through Hierarchically Unifying Isolated Information 29 Sep 2021 · 0 repositories
-
ESCo: Towards Provably Effective and Scalable Contrastive Representation Learning 29 Sep 2021 · 0 repositories
-
Fairness in Representation for Multilingual NLP: Insights from Controlled Experiments on Conditional Language Modeling 29 Sep 2021 · 0 repositories
-
Federated Contrastive Learning for Privacy-Preserving Unpaired Image-to-Image Translation 29 Sep 2021 · 0 repositories
-
GenTAL: Generative Denoising Skip-gram Transformer for Unsupervised Binary Code Similarity Detection 29 Sep 2021 · 0 repositories
-
Geometry-Entangled Visual Semantic Transformer for Image Captioning 29 Sep 2021 · 0 repositories · arXiv:2109.14137
-
Gradient Broadcast Adaptation: Defending against the backdoor attack in pre-trained models 29 Sep 2021 · 0 repositories
-
GRAPHIX: A Pre-trained Graph Edit Model for Automated Program Repair 29 Sep 2021 · 0 repositories
-
Grounding Language Representation with Visual Object Information via Cross Modal Pretraining 29 Sep 2021 · 0 repositories
-
Group-based Interleaved Pipeline Parallelism for Large-scale DNN Training 29 Sep 2021 · 1 repository
-
Guiding Transformers to Process in Steps 29 Sep 2021 · 0 repositories
-
HFSP: A Hardware-friendly Soft Pruning Framework for Vision Transformers 29 Sep 2021 · 0 repositories
-
Hierarchical Character Tagger for Short Text Spelling Error Correction 29 Sep 2021 · 0 repositories · arXiv:2109.14259
-
HoloFormer: Deep Compression of Pre-Trained Transforms via Unified Optimization of N:M Sparsity and Integer Quantization 29 Sep 2021 · 0 repositories
-
How does BERT address polysemy of Korean adverbial postpositions -ey, -eyse, and -(u)lo? 29 Sep 2021 · 0 repositories
-
HydraSum - Disentangling Stylistic Features in Text Summarization using Multi-Decoder Models 29 Sep 2021 · 0 repositories
-
Illiterate DALL·E Learns to Compose 29 Sep 2021 · 0 repositories
-
Improved Xception with Dual Attention Mechanism and Feature Fusion for Face Forgery Detection 29 Sep 2021 · 1 repository · arXiv:2109.14136
-
Improving Sentiment Classification Using 0-Shot Generated Labels for Custom Transformer Embeddings 29 Sep 2021 · 0 repositories
-
In defense of dual-encoders for neural ranking 29 Sep 2021 · 0 repositories
-
Isotropic Contextual Representations through Variational Regularization 29 Sep 2021 · 0 repositories
-
Language Model Pre-training Improves Generalization in Policy Learning 29 Sep 2021 · 0 repositories
-
Learning Rate Grafting: Transferability of Optimizer Tuning 29 Sep 2021 · 0 repositories
-
Learning to Schedule Learning rate with Graph Neural Networks 29 Sep 2021 · 0 repositories
-
Learning Visual-Linguistic Adequacy, Fidelity, and Fluency for Novel Object Captioning 29 Sep 2021 · 0 repositories
-
LMSA: Low-relation Mutil-head Self-Attention Mechanism in Visual Transformer 29 Sep 2021 · 0 repositories
-
Localizing Objects with Self-Supervised Transformers and no Labels 29 Sep 2021 · 2 repositories · arXiv:2109.14279
-
MaiT: integrating spatial locality into image transformers with attention masks 29 Sep 2021 · 1 repository
-
Mapping Language Models to Grounded Conceptual Spaces 29 Sep 2021 · 0 repositories
-
Mix-MaxEnt: Creating High Entropy Barriers To Improve Accuracy and Uncertainty Estimates of Deterministic Neural Networks 29 Sep 2021 · 0 repositories
-
MLP-based architecture with variable length input for automatic speech recognition 29 Sep 2021 · 0 repositories
-
Modeling label correlations implicitly through latent label encodings for multi-label text classification 29 Sep 2021 · 0 repositories
-
Modeling Variable Space with Residual Tensor Networks for Multivariate Time Series 29 Sep 2021 · 0 repositories
-
NAS-Bench-Zero: A Large Scale Dataset for Understanding Zero-Shot Neural Architecture Search 29 Sep 2021 · 0 repositories
-
Non-Autoregressive Models are Better Multilingual Translators 29 Sep 2021 · 0 repositories
-
Not All Regions are Worthy to be Distilled: Region-aware Knowledge Distillation Towards Efficient Image-to-Image Translation 29 Sep 2021 · 0 repositories
-
Offline Pre-trained Multi-Agent Decision Transformer 29 Sep 2021 · 0 repositories
-
Offline Reinforcement Learning for Large Scale Language Action Spaces 29 Sep 2021 · 0 repositories
-
Personalized Heterogeneous Federated Learning with Gradient Similarity 29 Sep 2021 · 0 repositories
-
Planning in Stochastic Environments with a Learned Model 29 Sep 2021 · 2 repositories
-
Policy improvement by planning with Gumbel 29 Sep 2021 · 2 repositories
-
Pretraining for Language Conditioned Imitation with Transformers 29 Sep 2021 · 0 repositories
-
Privacy-preserving Task-Agnostic Vision Transformer for Image Processing 29 Sep 2021 · 1 repository
-
Rank4Class: Examining Multiclass Classification through the Lens of Learning to Rank 29 Sep 2021 · 0 repositories
-
Robot Intent Recognition Method Based on State Grid Business Office 29 Sep 2021 · 0 repositories
-
ScaLA: Speeding-Up Fine-tuning of Pre-trained Transformer Networks via Efficient and Scalable Adversarial Perturbation 29 Sep 2021 · 0 repositories
-
Scale Efficiently: Insights from Pretraining and Finetuning Transformers 29 Sep 2021 · 0 repositories
-
Scaling the Depth of Vision Transformers via the Fourier Domain Analysis 29 Sep 2021 · 0 repositories
-
Scaling-up Diverse Orthogonal Convolutional Networks by a Paraunitary Framework 29 Sep 2021 · 0 repositories
-
Semi-supervised Offline Reinforcement Learning with Pre-trained Decision Transformers 29 Sep 2021 · 0 repositories
-
SeqPATE: Differentially Private Text Generation via Knowledge Distillation 29 Sep 2021 · 0 repositories
-
SiT: Simulation Transformer for Particle-based Physics Simulation 29 Sep 2021 · 0 repositories
-
Spanning Tree-based Graph Generation for Molecules 29 Sep 2021 · 0 repositories
-
Sparse Attention with Learning to Hash 29 Sep 2021 · 0 repositories
-
Sparse Unbalanced GAN Training with In-Time Over-Parameterization 29 Sep 2021 · 0 repositories
-
Specialized Transformers: Faster, Smaller and more Accurate NLP Models 29 Sep 2021 · 0 repositories
-
Subdimensional Expansion Using Attention-Based Learning For Multi-Agent Path Finding 29 Sep 2021 · 1 repository · arXiv:2109.14695
-
Temporal Action Localization with Global Segmentation Mask Transformers 29 Sep 2021 · 0 repositories
-
Test Time Robustification of Deep Models via Adaptation and Augmentation 29 Sep 2021 · 0 repositories
-
Topic Aware Neural Language Model: Domain Adaptation of Unconditional Text Generation Models 29 Sep 2021 · 0 repositories
-
Training sequence labeling models using prior knowledge 29 Sep 2021 · 0 repositories
-
Transliteration: A Simple Technique For Improving Multilingual Language Modeling 29 Sep 2021 · 0 repositories
-
TransSlowDown: Efficiency Attacks on Neural Machine Translation Systems 29 Sep 2021 · 0 repositories
-
TransTCN: An Attention-based TCN Framework for Sequential Modeling 29 Sep 2021 · 0 repositories
-
Tuformer: Data-Driven Design of Expressive Transformer by Tucker Tensor Representation 29 Sep 2021 · 0 repositories
-
UFO-ViT: High Performance Linear Vision Transformer without Softmax 29 Sep 2021 · 1 repository · arXiv:2109.14382
-
Understanding ResNet from a Discrete Dynamical System Perspective 29 Sep 2021 · 0 repositories
-
Understanding the Role of Self Attention for Efficient Speech Recognition 29 Sep 2021 · 0 repositories
-
Video Forgery Detection Using Multiple Cues on Fusion of EfficientNet and Swin Transformer 29 Sep 2021 · 0 repositories
-
VUT: Versatile UI Transformer for Multimodal Multi-Task User Interface Modeling 29 Sep 2021 · 0 repositories
-
A hierarchical residual network with compact triplet-center loss for sketch recognition 28 Sep 2021 · 0 repositories · arXiv:2109.13536
-
Fine-tuning Vision Transformers for the Prediction of State Variables in Ising Models 28 Sep 2021 · 0 repositories · arXiv:2109.13925
-
How Different Text-preprocessing Techniques Using The BERT Model Affect The Gender Profiling of Authors 28 Sep 2021 · 0 repositories · arXiv:2109.13890
-
Nana-HDR: A Non-attentive Non-autoregressive Hybrid Model for TTS 28 Sep 2021 · 0 repositories · arXiv:2109.13673
-
RAFT: A Real-World Few-Shot Text Classification Benchmark 28 Sep 2021 · 1 repository · arXiv:2109.14076
-
Single-dataset Experts for Multi-dataset Question Answering 28 Sep 2021 · 1 repository · arXiv:2109.13880Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
What to Prioritize? Natural Language Processing for the Development of a Modern Bug Tracking Solution in Hardware Development 28 Sep 2021 · 0 repositories · arXiv:2109.13825
-
Focus! Rating XAI Methods and Finding Biases 28 Sep 2021 · 1 repository · arXiv:2109.15035
-
Compressive Visual Representations 27 Sep 2021 · 1 repository · arXiv:2109.12909Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
Effective Use of Graph Convolution Network and Contextual Sub-Tree forCommodity News Event Extraction 27 Sep 2021 · 1 repository · arXiv:2109.12781
-
Fast-MD: Fast Multi-Decoder End-to-End Speech Translation with Non-Autoregressive Hidden Intermediates 27 Sep 2021 · 1 repository · arXiv:2109.12804
-
GANiry: Bald-to-Hairy Translation Using CycleGAN 27 Sep 2021 · 1 repository · arXiv:2109.13126
-
Improving Stack Overflow question title generation with copying enhanced CodeBERT model and bi-modal information 27 Sep 2021 · 1 repository · arXiv:2109.13073
-
Integrated Training for Sequence-to-Sequence Models Using Non-Autoregressive Transformer 27 Sep 2021 · 0 repositories · arXiv:2109.12950
-
Optimising for Interpretability: Convolutional Dynamic Alignment Networks 27 Sep 2021 · 1 repository · arXiv:2109.13004
-
PASS: An ImageNet replacement for self-supervised pretraining without humans 27 Sep 2021 · 1 repository · arXiv:2109.13228Syntology official: no sample here; runs from other or unrecorded repositories · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 9 pointer-only (licence)
-
Patterns of Lexical Ambiguity in Contextualised Language Models 27 Sep 2021 · 0 repositories · arXiv:2109.13032
-
SAU: Smooth activation function using convolution with approximate identities 27 Sep 2021 · 0 repositories · arXiv:2109.13210