Methods › General › Skip Connections › Residual Connection › Papers, page 235
Residual Connection
Papers archive 2025-07-28
archive papers tagged: 28,401 · with a code link: 12,847 · where Syntology ran a sample: 3,897 (3,291 with a run with no instrument failure, 606 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,897 of 28,401 tagged: 3,291 with a run with no instrument failure, 606 where every run was a failure of Syntology's instrument)
Page 235 of 285: papers 23,401 to 23,500 of 28,401, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fully Non-autoregressive Neural Machine Translation: Tricks of the Trade 31 Dec 2020 · 1 repository · arXiv:2012.15833
-
KART: Parameterization of Privacy Leakage Scenarios from Pre-trained Language Models 31 Dec 2020 · 1 repository · arXiv:2101.00036
-
Fast WordPiece Tokenization 31 Dec 2020 · 1 repository · arXiv:2012.15524
-
Making Pre-trained Language Models Better Few-shot Learners 31 Dec 2020 · 9 repositories · arXiv:2012.15723Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 9 harvested samples) · 7 pointer-only (licence)
-
MiniLMv2: Multi-Head Self-Attention Relation Distillation for Compressing Pretrained Transformers 31 Dec 2020 · 2 repositories · arXiv:2012.15828
-
Rethinking Semantic Segmentation from a Sequence-to-Sequence Perspective with Transformers 31 Dec 2020 · 5 repositories · arXiv:2012.15840
-
Revisiting Robust Neural Machine Translation: A Transformer Case Study 31 Dec 2020 · 0 repositories · arXiv:2012.15710
-
Studying Strategically: Learning to Mask for Closed-book QA 31 Dec 2020 · 0 repositories · arXiv:2012.15856
-
The Pile: An 800GB Dataset of Diverse Text for Language Modeling 31 Dec 2020 · 22 repositories · arXiv:2101.00027Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
TransTrack: Multiple Object Tracking with Transformer 31 Dec 2020 · 2 repositories · arXiv:2012.15460
-
Unified Mandarin TTS Front-end Based on Distilled BERT Model 31 Dec 2020 · 1 repository · arXiv:2012.15404
-
UNKs Everywhere: Adapting Multilingual Language Models to New Scripts 31 Dec 2020 · 2 repositories · arXiv:2012.15562
-
Verb Knowledge Injection for Multilingual Event Processing 31 Dec 2020 · 0 repositories · arXiv:2012.15421
-
XLM-T: Scaling up Multilingual Machine Translation with Pretrained Cross-lingual Transformer Encoders 31 Dec 2020 · 0 repositories · arXiv:2012.15547
-
ECONET: Effective Continual Pretraining of Language Models for Event Temporal Reasoning 30 Dec 2020 · 2 repositories · arXiv:2012.15283
-
Deriving Contextualised Semantic Features from BERT (and Other Transformer Model) Embeddings 30 Dec 2020 · 0 repositories · arXiv:2012.15353
-
Improving BERT with Syntax-aware Local Attention 30 Dec 2020 · 1 repository · arXiv:2012.15150
-
Optimizing Deeper Transformers on Small Datasets 30 Dec 2020 · 1 repository · arXiv:2012.15355
-
Out of Order: How Important Is The Sequential Order of Words in a Sentence in Natural Language Understanding Tasks? 30 Dec 2020 · 0 repositories · arXiv:2012.15180
-
SemGloVe: Semantic Co-occurrences for GloVe from BERT 30 Dec 2020 · 3 repositories · arXiv:2012.15197
-
Transformer for Image Quality Assessment 30 Dec 2020 · 0 repositories · arXiv:2101.01097
-
UnNatural Language Inference 30 Dec 2020 · 1 repository · arXiv:2101.00010
-
A Hierarchical Transformer with Speaker Modeling for Emotion Recognition in Conversation 29 Dec 2020 · 1 repository · arXiv:2012.14781
-
Ensembled ResUnet for Anatomical Brain Barriers Segmentation 29 Dec 2020 · 0 repositories · arXiv:2012.14567
-
Hybrid Micro/Macro Level Convolution for Heterogeneous Graph Learning 29 Dec 2020 · 1 repository · arXiv:2012.14722
-
Kaleidoscope: An Efficient, Learnable Representation For All Structured Linear Maps 29 Dec 2020 · 2 repositories · arXiv:2012.14966Syntology official (archive's flag): 16 ran · 16 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 7 where Syntology's instrument failed) · 7 unverified (of 23 harvested samples)
-
LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding 29 Dec 2020 · 9 repositories · arXiv:2012.14740
-
Robust Dialogue Utterance Rewriting as Sequence Tagging 29 Dec 2020 · 1 repository · arXiv:2012.14535
-
Code Summarization with Structure-induced Transformer 29 Dec 2020 · 1 repository · arXiv:2012.14710
-
A Paragraph-level Multi-task Learning Model for Scientific Fact-Verification 28 Dec 2020 · 1 repository · arXiv:2012.14500
-
BURT: BERT-inspired Universal Representation from Learning Meaningful Segment 28 Dec 2020 · 0 repositories · arXiv:2012.14320
-
Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models 28 Dec 2020 · 2 repositories · arXiv:2012.14252
-
Red Dragon AI at TextGraphs 2020 Shared Task: LIT : LSTM-Interleaved Transformer for Multi-Hop Explanation Ranking 28 Dec 2020 · 1 repository · arXiv:2012.14164
-
Syntax-Enhanced Pre-trained Model 28 Dec 2020 · 1 repository · arXiv:2012.14116
-
TransPose: Keypoint Localization via Transformer 28 Dec 2020 · 1 repository · arXiv:2012.14214
-
A multi-task learning network using shared BERT models for aspect-based sentiment analysis 27 Dec 2020 · 0 repositories
-
ALP-KD: Attention-Based Layer Projection for Knowledge Distillation 27 Dec 2020 · 0 repositories · arXiv:2012.14022
-
An Embarrassingly Simple Model for Dialogue Relation Extraction 27 Dec 2020 · 1 repository · arXiv:2012.13873
-
Inserting Information Bottlenecks for Attribution in Transformers 27 Dec 2020 · 1 repository · arXiv:2012.13838
-
Learning Light-Weight Translation Models from Deep Transformer 27 Dec 2020 · 1 repository · arXiv:2012.13866Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
MeDAL: Medical Abbreviation Disambiguation Dataset for Natural Language Understanding Pretraining 27 Dec 2020 · 1 repository · arXiv:2012.13978
-
Portfolio Optimization with 2D Relative-Attentional Gated Transformer 27 Dec 2020 · 0 repositories · arXiv:2101.03138
-
SG-Net: Syntax Guided Transformer for Language Representation 27 Dec 2020 · 0 repositories · arXiv:2012.13915
-
Direct Quantization for Training Highly Accurate Low Bit-width Deep Neural Networks 26 Dec 2020 · 0 repositories · arXiv:2012.13762
-
Sparse Adversarial Attack to Object Detection 26 Dec 2020 · 1 repository · arXiv:2012.13692
-
Revisiting Edge Detection in Convolutional Neural Networks 25 Dec 2020 · 0 repositories · arXiv:2012.13576
-
Appearance-Invariant 6-DoF Visual Localization using Generative Adversarial Networks 24 Dec 2020 · 0 repositories · arXiv:2012.13191
-
Detecting Hateful Memes Using a Multimodal Deep Ensemble 24 Dec 2020 · 1 repository · arXiv:2012.13235
-
Global Context Networks 24 Dec 2020 · 3 repositories · arXiv:2012.13375
-
I like fish, especially dolphins: Addressing Contradictions in Dialogue Modeling 24 Dec 2020 · 0 repositories · arXiv:2012.13391
-
On the Granularity of Explanations in Model Agnostic NLP Interpretability 24 Dec 2020 · 1 repository · arXiv:2012.13189
-
To what extent do human explanations of model behavior align with actual model behavior? 24 Dec 2020 · 0 repositories · arXiv:2012.13354
-
A Survey on Visual Transformer 23 Dec 2020 · 0 repositories · arXiv:2012.12556
-
Bridging Textual and Tabular Data for Cross-Domain Text-to-SQL Semantic Parsing 23 Dec 2020 · 2 repositories · arXiv:2012.12627
-
Detecting Hate Speech in Memes Using Multimodal Deep Learning Approaches: Prize-winning solution to Hateful Memes Challenge 23 Dec 2020 · 1 repository · arXiv:2012.12975
-
Future-Guided Incremental Transformer for Simultaneous Translation 23 Dec 2020 · 0 repositories · arXiv:2012.12465
-
ICMSC: Intra- and Cross-modality Semantic Consistency for Unsupervised Domain Adaptation on Hip Joint Bone Segmentation 23 Dec 2020 · 0 repositories · arXiv:2012.12570
-
SWA Object Detection 23 Dec 2020 · 2 repositories · arXiv:2012.12645
-
Applying Wav2vec2.0 to Speech Recognition in Various Low-resource Languages 22 Dec 2020 · 0 repositories · arXiv:2012.12121
-
Domain Adaptation of NMT models for English-Hindi Machine Translation Task at AdapMT ICON 2020 22 Dec 2020 · 0 repositories · arXiv:2012.12112
-
FcaNet: Frequency Channel Attention Networks 22 Dec 2020 · 7 repositories · arXiv:2012.11879
-
Improved Biomedical Word Embeddings in the Transformer Era 22 Dec 2020 · 1 repository · arXiv:2012.11808
-
Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning 22 Dec 2020 · 2 repositories · arXiv:2012.13255Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Learning to Retrieve Entity-Aware Knowledge and Generate Responses with Copy Mechanism for Task-Oriented Dialogue Systems 22 Dec 2020 · 1 repository · arXiv:2012.11937
-
Molecular CT: Unifying Geometry and Representation Learning for Molecules at Different Scales 22 Dec 2020 · 0 repositories · arXiv:2012.11816
-
Multi-Head Self-Attention with Role-Guided Masks 22 Dec 2020 · 1 repository · arXiv:2012.12366
-
Prediction of Chronic Kidney Disease Using Deep Neural Network 22 Dec 2020 · 0 repositories · arXiv:2012.12089
-
Recognizing Emotion Cause in Conversations 22 Dec 2020 · 1 repository · arXiv:2012.11820
-
Uncertainty and Surprisal Jointly Deliver the Punchline: Exploiting Incongruity-Based Features for Humor Recognition 22 Dec 2020 · 0 repositories · arXiv:2012.12007
-
Universal Approximation Properties for an ODENet and a ResNet: Mathematical Analysis and Numerical Experiments 22 Dec 2020 · 0 repositories · arXiv:2101.10229
-
3D Object Detection with Pointformer 21 Dec 2020 · 1 repository · arXiv:2012.11409
-
A Graph Reasoning Network for Multi-turn Response Selection via Customized Pre-training 21 Dec 2020 · 0 repositories · arXiv:2012.11099
-
Cross-domain Retrieval in the Legal and Patent Domains: a Reproducibility Study 21 Dec 2020 · 1 repository · arXiv:2012.11405
-
Domain specific BERT representation for Named Entity Recognition of lab protocol 21 Dec 2020 · 1 repository · arXiv:2012.11145
-
Encoding Syntactic Knowledge in Transformer Encoder for Intent Detection and Slot Filling 21 Dec 2020 · 0 repositories · arXiv:2012.11689
-
RealFormer: Transformer Likes Residual Attention 21 Dec 2020 · 5 repositories · arXiv:2012.11747
-
Infrared image pedestrian target detection based on Yolov3 and migration learning 21 Dec 2020 · 0 repositories · arXiv:2012.11185
-
Leveraging ParsBERT and Pretrained mT5 for Persian Abstractive Text Summarization 21 Dec 2020 · 1 repository · arXiv:2012.11204
-
OBoW: Online Bag-of-Visual-Words Generation for Self-Supervised Learning 21 Dec 2020 · 3 repositories · arXiv:2012.11552Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Small-Footprint Wake Up Word Recognition in Noisy Environments Employing Competing-Words-Based Feature 21 Dec 2020 · 0 repositories
-
Sub-Linear Memory: How to Make Performers SLiM 21 Dec 2020 · 2 repositories · arXiv:2012.11346
-
Towards Incorporating Entity-specific Knowledge Graph Information in Predicting Drug-Drug Interactions 21 Dec 2020 · 0 repositories · arXiv:2012.11142
-
Adaptive Bi-directional Attention: Exploring Multi-Granularity Representations for Machine Reading Comprehension 20 Dec 2020 · 0 repositories · arXiv:2012.10877
-
Color Channel Perturbation Attacks for Fooling Convolutional Neural Networks and A Defense Against Such Attacks 20 Dec 2020 · 1 repository · arXiv:2012.14456
-
Breaking Writer's Block: Low-cost Fine-tuning of Natural Language Generation Models 19 Dec 2020 · 0 repositories · arXiv:2101.03216
-
Enabling Retrain-free Deep Neural Network Pruning using Surrogate Lagrangian Relaxation 18 Dec 2020 · 0 repositories · arXiv:2012.10079
-
An Empirical Study of Using Pre-trained BERT Models for Vietnamese Relation Extraction Task at VLSP 2020 18 Dec 2020 · 1 repository · arXiv:2012.10275
-
HateXplain: A Benchmark Dataset for Explainable Hate Speech Detection 18 Dec 2020 · 6 repositories · arXiv:2012.10289
-
NeurST: Neural Speech Translation Toolkit 18 Dec 2020 · 1 repository · arXiv:2012.10018
-
Object Detection based on OcSaFPN in Aerial Images with Noise 18 Dec 2020 · 0 repositories · arXiv:2012.09859
-
On Modality Bias in the TVQA Dataset 18 Dec 2020 · 1 repository · arXiv:2012.10210
-
Toward Streaming ASR with Non-Autoregressive Insertion-based Model 18 Dec 2020 · 0 repositories · arXiv:2012.10128
-
A Generalization of Transformer Networks to Graphs 17 Dec 2020 · 3 repositories · arXiv:2012.09699Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A White Box Analysis of ColBERT 17 Dec 2020 · 0 repositories · arXiv:2012.09650
-
BERT Goes Shopping: Comparing Distributional Models for Product Representations 17 Dec 2020 · 1 repository · arXiv:2012.09807
-
End-to-end Deep Object Tracking with Circular Loss Function for Rotated Bounding Box 17 Dec 2020 · 0 repositories · arXiv:2012.09771
-
Literature Retrieval for Precision Medicine with Neural Matching and Faceted Summarization 17 Dec 2020 · 1 repository · arXiv:2012.09355
-
MASKER: Masked Keyword Regularization for Reliable Text Classification 17 Dec 2020 · 1 repository · arXiv:2012.09392
-
MIX : a Multi-task Learning Approach to Solve Open-Domain Question Answering 17 Dec 2020 · 0 repositories · arXiv:2012.09766
-
Parallel WaveNet conditioned on VAE latent vectors 17 Dec 2020 · 0 repositories · arXiv:2012.09703