Methods › Natural Language Processing › Autoregressive Transformers › Transformer › Papers, page 115
Transformer
Papers archive 2025-07-28
archive papers tagged: 13,999 · with a code link: 6,572 · where Syntology ran a sample: 2,248 (1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,248 of 13,999 tagged: 1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument)
Page 115 of 140: papers 11,401 to 11,500 of 13,999, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Actions Speak Louder than Listening: Evaluating Music Style Transfer based on Editing Experience 25 Oct 2021 · 1 repository · arXiv:2110.12855
-
DocTr: Document Image Transformer for Geometric Unwarping and Illumination Correction 25 Oct 2021 · 2 repositories · arXiv:2110.12942
-
Gophormer: Ego-Graph Transformer for Node Classification 25 Oct 2021 · 0 repositories · arXiv:2110.13094
-
History Aware Multimodal Transformer for Vision-and-Language Navigation 25 Oct 2021 · 1 repository · arXiv:2110.13309Syntology 7 ran (of which 3 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 15 harvested samples)
-
IconQA: A New Benchmark for Abstract Diagram Understanding and Visual Language Reasoning 25 Oct 2021 · 1 repository · arXiv:2110.13214Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
MVT: Multi-view Vision Transformer for 3D Object Recognition 25 Oct 2021 · 2 repositories · arXiv:2110.13083Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Paradigm Shift in Language Modeling: Revisiting CNN for Modeling Sanskrit Originated Bengali and Hindi Language 25 Oct 2021 · 0 repositories · arXiv:2110.13032
-
The Nuts and Bolts of Adopting Transformer in GANs 25 Oct 2021 · 0 repositories · arXiv:2110.13107
-
CvT-ASSD: Convolutional vision-Transformer Based Attentive Single Shot MultiBox Detector 24 Oct 2021 · 1 repository · arXiv:2110.12364
-
Grafting Transformer on Automatically Designed Convolutional Neural Network for Hyperspectral Image Classification 21 Oct 2021 · 1 repository · arXiv:2110.11084
-
Transformer Acceleration with Dynamic Sparse Attention 21 Oct 2021 · 0 repositories · arXiv:2110.11299
-
Vis-TOP: Visual Transformer Overlay Processor 21 Oct 2021 · 0 repositories · arXiv:2110.10957
-
AFTer-UNet: Axial Fusion Transformer UNet for Medical Image Segmentation 20 Oct 2021 · 0 repositories · arXiv:2110.10403
-
AniFormer: Data-driven 3D Animation with Transformer 20 Oct 2021 · 1 repository · arXiv:2110.10533
-
Continual Learning in Multilingual NMT via Language-Specific Embeddings 20 Oct 2021 · 0 repositories · arXiv:2110.10478
-
ESOD:Edge-based Task Scheduling for Object Detection 20 Oct 2021 · 0 repositories · arXiv:2110.11342
-
Few-Shot Temporal Action Localization with Query Adaptive Transformer 20 Oct 2021 · 1 repository · arXiv:2110.10552
-
SEA: Graph Shell Attention in Graph Neural Networks 20 Oct 2021 · 0 repositories · arXiv:2110.10674
-
Toward Accurate and Reliable Iris Segmentation Using Uncertainty Learning 20 Oct 2021 · 0 repositories · arXiv:2110.10334
-
VLDeformer: Vision-Language Decomposed Transformer for Fast Cross-Modal Retrieval 20 Oct 2021 · 0 repositories · arXiv:2110.11338
-
A Picture is Worth a Thousand Words: A Unified System for Diverse Captions and Rich Images Generation 19 Oct 2021 · 1 repository · arXiv:2110.09756
-
Accelerating Framework of Transformer by Hardware Design and Model Compression Co-Optimization 19 Oct 2021 · 0 repositories · arXiv:2110.10030
-
Bilateral-ViT for Robust Fovea Localization 19 Oct 2021 · 0 repositories · arXiv:2110.09860
-
DetectorNet: Transformer-enhanced Spatial Temporal Graph Neural Network for Traffic Prediction 19 Oct 2021 · 0 repositories · arXiv:2111.00869
-
Generating Symbolic Reasoning Problems with Transformer GANs 19 Oct 2021 · 1 repository · arXiv:2110.10054Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Inductive Biases and Variable Creation in Self-Attention Mechanisms 19 Oct 2021 · 0 repositories · arXiv:2110.10090
-
Permutation invariant graph-to-sequence model for template-free retrosynthesis and reaction prediction 19 Oct 2021 · 1 repository · arXiv:2110.09681Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Spatial-Temporal Transformer for 3D Point Cloud Sequences 19 Oct 2021 · 0 repositories · arXiv:2110.09783
-
SSAST: Self-Supervised Audio Spectrogram Transformer 19 Oct 2021 · 3 repositories · arXiv:2110.09784Syntology official (archive's flag): 1 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
Unifying Multimodal Transformer for Bi-directional Image and Text Generation 19 Oct 2021 · 1 repository · arXiv:2110.09753Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Compositional Attention: Disentangling Search and Retrieval 18 Oct 2021 · 3 repositories · arXiv:2110.09419
-
HRFormer: High-Resolution Transformer for Dense Prediction 18 Oct 2021 · 1 repository · arXiv:2110.09408Syntology official (archive's flag): 7 ran · 7 ran (of which 7 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 7 samples that ran constructed an object rather than computing a result (of 11 harvested samples)
-
SentimentArcs: A Novel Method for Self-Supervised Sentiment Analysis of Time Series Shows SOTA Transformers Can Struggle Finding Narrative Arcs 18 Oct 2021 · 1 repository · arXiv:2110.09454Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Sequential Modeling with Multiple Attributes for Watchlist Recommendation in E-Commerce 18 Oct 2021 · 1 repository · arXiv:2110.11072
-
3D-RETR: End-to-End Single and Multi-View 3D Reconstruction with Transformers 17 Oct 2021 · 1 repository · arXiv:2110.08861
-
CAE-Transformer: Transformer-based Model to Predict Invasiveness of Lung Adenocarcinoma Subsolid Nodules from Non-thin Section 3D CT Scans 17 Oct 2021 · 0 repositories · arXiv:2110.08721
-
Siamese Transformer Pyramid Networks for Real-Time UAV Tracking 17 Oct 2021 · 1 repository · arXiv:2110.08822
-
A Good Prompt Is Worth Millions of Parameters: Low-resource Prompt-based Learning for Vision-Language Models 16 Oct 2021 · 1 repository · arXiv:2110.08484
-
Alleviating the Inequality of Attention Heads for Neural Machine Translation 16 Oct 2021 · 0 repositories
-
ASFormer: Transformer for Action Segmentation 16 Oct 2021 · 1 repository · arXiv:2110.08568Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Attention Temperature Matters in Abstractive Summarization Distillation 16 Oct 2021 · 0 repositories
-
COVID-19 Detection in Chest X-ray Images Using Swin-Transformer and Transformer in Transformer 16 Oct 2021 · 0 repositories · arXiv:2110.08427
-
Emotion Flip Reasoning in Multiparty Conversations 16 Oct 2021 · 0 repositories
-
Hierarchical Transformer Networks for Long-sequence and Multiple Clinical Documents Classification 16 Oct 2021 · 0 repositories
-
ODE Transformer: An Ordinary Differential Equation-Inspired Model for Sequence Generation 16 Oct 2021 · 0 repositories
-
Improving Transformers with Probabilistic Attention Keys 16 Oct 2021 · 1 repository · arXiv:2110.08678Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Combining CNNs With Transformer for Multimodal 3D MRI Brain Tumor Segmentation With Self-Supervised Pretraining 15 Oct 2021 · 0 repositories · arXiv:2110.07919
-
On Learning the Transformer Kernel 15 Oct 2021 · 1 repository · arXiv:2110.08323
-
StreaMulT: Streaming Multimodal Transformer for Heterogeneous and Arbitrary Long Sequential Data 15 Oct 2021 · 0 repositories · arXiv:2110.08021
-
Causal Transformers Perform Below Chance on Recursive Nested Constructions, Unlike Humans 14 Oct 2021 · 0 repositories · arXiv:2110.07240
-
Evaluating Off-the-Shelf Machine Listening and Natural Language Models for Automated Audio Captioning 14 Oct 2021 · 0 repositories · arXiv:2110.07410
-
Integrating Fréchet distance and AI reveals the evolutionary trajectory and origin of SARS-CoV-2 14 Oct 2021 · 0 repositories · arXiv:2110.07696
-
Identifying Introductions in Podcast Episodes from Automatically Generated Transcripts 14 Oct 2021 · 1 repository · arXiv:2110.07096
-
Improved Drug-target Interaction Prediction with Intermolecular Graph Transformer 14 Oct 2021 · 0 repositories · arXiv:2110.07347
-
Non-Autoregressive Translation with Layer-Wise Prediction and Deep Supervision 14 Oct 2021 · 2 repositories · arXiv:2110.07515Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Sub-word Level Lip Reading With Visual Attention 14 Oct 2021 · 0 repositories · arXiv:2110.07603
-
The Neural Data Router: Adaptive Control Flow in Transformers Improves Systematic Generalization 14 Oct 2021 · 1 repository · arXiv:2110.07732Syntology official (archive's flag): 7 ran · 7 ran (of which 5 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
Transformer for Polyp Detection 14 Oct 2021 · 0 repositories · arXiv:2111.07918
-
CLIP4Caption: CLIP for Video Caption 13 Oct 2021 · 0 repositories · arXiv:2110.06615
-
Leveraging redundancy in attention with Reuse Transformers 13 Oct 2021 · 1 repository · arXiv:2110.06821
-
Semantics-aware Attention Improves Neural Machine Translation 13 Oct 2021 · 0 repositories · arXiv:2110.06920
-
Study of positional encoding approaches for Audio Spectrogram Transformers 13 Oct 2021 · 1 repository · arXiv:2110.06999
-
The Dawn of Quantum Natural Language Processing 13 Oct 2021 · 2 repositories · arXiv:2110.06510
-
Transform and Bitstream Domain Image Classification 13 Oct 2021 · 0 repositories · arXiv:2110.06740
-
Yformer: U-Net Inspired Transformer Architecture for Far Horizon Time Series Forecasting 13 Oct 2021 · 1 repository · arXiv:2110.08255
-
Attention-guided Generative Models for Extractive Question Answering 12 Oct 2021 · 0 repositories · arXiv:2110.06393
-
Contrastive Learning for Representation Degeneration Problem in Sequential Recommendation 12 Oct 2021 · 2 repositories · arXiv:2110.05730Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
DiscoDVT: Generating Long Text with Discourse-Aware Discrete Variational Transformer 12 Oct 2021 · 1 repository · arXiv:2110.05999
-
HETFORMER: Heterogeneous Transformer with Sparse Attention for Long-Text Extractive Summarization 12 Oct 2021 · 1 repository · arXiv:2110.06388
-
LightSeq2: Accelerated Training for Transformer-based Models on GPUs 12 Oct 2021 · 1 repository · arXiv:2110.05722
-
Mention Memory: incorporating textual knowledge into Transformers through entity mention attention 12 Oct 2021 · 1 repository · arXiv:2110.06176
-
Relative Molecule Self-Attention Transformer 12 Oct 2021 · 1 repository · arXiv:2110.05841Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Rescoring Sequence-to-Sequence Models for Text Line Recognition with CTC-Prefixes 12 Oct 2021 · 1 repository · arXiv:2110.05909
-
Satellite Image Semantic Segmentation 12 Oct 2021 · 1 repository · arXiv:2110.05812
-
StARformer: Transformer with State-Action-Reward Representations for Visual Reinforcement Learning 12 Oct 2021 · 1 repository · arXiv:2110.06206Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Adaptive Multi-view and Temporal Fusing Transformer for 3D Human Pose Estimation 11 Oct 2021 · 0 repositories · arXiv:2110.05092
-
Investigating Transfer Learning Capabilities of Vision Transformers and CNNs by Fine-Tuning a Single Trainable Block 11 Oct 2021 · 1 repository · arXiv:2110.05270
-
Multi-View Self-Attention Based Transformer for Speaker Recognition 11 Oct 2021 · 0 repositories · arXiv:2110.05036
-
SRU++: Pioneering Fast Recurrence with Attention for Speech Recognition 11 Oct 2021 · 0 repositories · arXiv:2110.05571
-
Unsupervised Source Separation via Bayesian Inference in the Latent Domain 11 Oct 2021 · 1 repository · arXiv:2110.05313
-
DCT: Dynamic Compressive Transformer for Modeling Unbounded Sequence 10 Oct 2021 · 0 repositories · arXiv:2110.04821
-
Multi-Channel End-to-End Neural Diarization with Distributed Microphones 10 Oct 2021 · 0 repositories · arXiv:2110.04694
-
Global Vision Transformer Pruning with Hessian-Aware Saliency 10 Oct 2021 · 1 repository · arXiv:2110.04869Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Automatic Text Extractive Summarization Based on Graph and Pre-trained Language Model Attention 10 Oct 2021 · 0 repositories · arXiv:2110.04878
-
SP-GPT2: Semantics Improvement in Vietnamese Poetry Generation 10 Oct 2021 · 2 repositories · arXiv:2110.15723
-
SuperShaper: Task-Agnostic Super Pre-training of BERT Models with Variable Hidden Dimensions 10 Oct 2021 · 0 repositories · arXiv:2110.04711
-
Vector-quantized Image Modeling with Improved VQGAN 9 Oct 2021 · 5 repositories · arXiv:2110.04627Syntology 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
The Layout Generation Algorithm of Graphic Design Based on Transformer-CVAE 8 Oct 2021 · 0 repositories · arXiv:2110.06794
-
Adversarial Token Attacks on Vision Transformers 8 Oct 2021 · 0 repositories · arXiv:2110.04337
-
Context-LGM: Leveraging Object-Context Relation for Context-Aware Object Recognition 8 Oct 2021 · 0 repositories · arXiv:2110.04042
-
Towards Learning (Dis)-Similarity of Source Code from Program Contrasts 8 Oct 2021 · 0 repositories · arXiv:2110.03868
-
Knowledge-Enhanced Hierarchical Graph Transformer Network for Multi-Behavior Recommendation 8 Oct 2021 · 1 repository · arXiv:2110.04000
-
Learning Adaptive Control Flow in Transformers for Improved Systematic Generalization 8 Oct 2021 · 0 repositories
-
M6-10T: A Sharing-Delinking Paradigm for Efficient Multi-Trillion Parameter Pretraining 8 Oct 2021 · 0 repositories · arXiv:2110.03888
-
Multiplex Behavioral Relation Learning for Recommendation via Memory Augmented Transformer Network 8 Oct 2021 · 1 repository · arXiv:2110.04002
-
RPT: Toward Transferable Model on Heterogeneous Researcher Data via Pre-Training 8 Oct 2021 · 1 repository · arXiv:2110.07336
-
Taming Sparsely Activated Transformer with Stochastic Experts 8 Oct 2021 · 1 repository · arXiv:2110.04260Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
ViDT: An Efficient and Effective Fully Transformer-based Object Detector 8 Oct 2021 · 1 repository · arXiv:2110.03921
-
Attention is All You Need? Good Embeddings with Statistics are enough:Large Scale Audio Understanding without Transformers/ Convolutions/ BERTs/ Mixers/ Attention/ RNNs or .... 7 Oct 2021 · 0 repositories · arXiv:2110.03183
-
End-to-End Supermask Pruning: Learning to Prune Image Captioning Models 7 Oct 2021 · 1 repository · arXiv:2110.03298Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 1 pointer-only (licence)