Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 159
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 159 of 190: papers 15,801 to 15,900 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Improving Transformers with Probabilistic Attention Keys 16 Oct 2021 · 1 repository · arXiv:2110.08678Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 16 Oct 2021 · 0 repositories
-
Combining CNNs With Transformer for Multimodal 3D MRI Brain Tumor Segmentation With Self-Supervised Pretraining 15 Oct 2021 · 0 repositories · arXiv:2110.07919
-
Hierarchical Curriculum Learning for AMR Parsing 15 Oct 2021 · 1 repository · arXiv:2110.07855
-
Kronecker Decomposition for GPT Compression 15 Oct 2021 · 0 repositories · arXiv:2110.08152
-
On Learning the Transformer Kernel 15 Oct 2021 · 1 repository · arXiv:2110.08323
-
StreaMulT: Streaming Multimodal Transformer for Heterogeneous and Arbitrary Long Sequential Data 15 Oct 2021 · 0 repositories · arXiv:2110.08021
-
Building Chinese Biomedical Language Models via Multi-Level Text Discrimination 14 Oct 2021 · 1 repository · arXiv:2110.07244
-
Causal Transformers Perform Below Chance on Recursive Nested Constructions, Unlike Humans 14 Oct 2021 · 0 repositories · arXiv:2110.07240
-
Interpreting the Robustness of Neural NLP Models to Textual Perturbations 14 Oct 2021 · 0 repositories · arXiv:2110.07159
-
Can Machines Learn Morality? The Delphi Experiment 14 Oct 2021 · 1 repository · arXiv:2110.07574
-
Evaluating Off-the-Shelf Machine Listening and Natural Language Models for Automated Audio Captioning 14 Oct 2021 · 0 repositories · arXiv:2110.07410
-
Integrating Fréchet distance and AI reveals the evolutionary trajectory and origin of SARS-CoV-2 14 Oct 2021 · 0 repositories · arXiv:2110.07696
-
Identifying Introductions in Podcast Episodes from Automatically Generated Transcripts 14 Oct 2021 · 1 repository · arXiv:2110.07096
-
Improved Drug-target Interaction Prediction with Intermolecular Graph Transformer 14 Oct 2021 · 0 repositories · arXiv:2110.07347
-
LFPT5: A Unified Framework for Lifelong Few-shot Language Learning Based on Prompt Tuning of T5 14 Oct 2021 · 1 repository · arXiv:2110.07298Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples)
-
CaPE: Contrastive Parameter Ensembling for Reducing Hallucination in Abstractive Summarization 14 Oct 2021 · 0 repositories · arXiv:2110.07166
-
Non-Autoregressive Translation with Layer-Wise Prediction and Deep Supervision 14 Oct 2021 · 2 repositories · arXiv:2110.07515Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing 14 Oct 2021 · 6 repositories · arXiv:2110.07205Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Sub-word Level Lip Reading With Visual Attention 14 Oct 2021 · 0 repositories · arXiv:2110.07603
-
Symbolic Knowledge Distillation: from General Language Models to Commonsense Models 14 Oct 2021 · 1 repository · arXiv:2110.07178
-
The Neural Data Router: Adaptive Control Flow in Transformers Improves Systematic Generalization 14 Oct 2021 · 1 repository · arXiv:2110.07732Syntology official (archive's flag): 7 ran · 7 ran (of which 5 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
Transformer for Polyp Detection 14 Oct 2021 · 0 repositories · arXiv:2111.07918
-
CLIP4Caption: CLIP for Video Caption 13 Oct 2021 · 0 repositories · arXiv:2110.06615
-
Leveraging Generative Models for Covert Messaging: Challenges and Tradeoffs for "Dead-Drop" Deployments 13 Oct 2021 · 0 repositories · arXiv:2110.07009
-
Language Modelling via Learning to Rank 13 Oct 2021 · 0 repositories · arXiv:2110.06961
-
Leveraging redundancy in attention with Reuse Transformers 13 Oct 2021 · 1 repository · arXiv:2110.06821
-
Scaling Laws for the Few-Shot Adaptation of Pre-trained Image Classifiers 13 Oct 2021 · 0 repositories · arXiv:2110.06990
-
Semantics-aware Attention Improves Neural Machine Translation 13 Oct 2021 · 0 repositories · arXiv:2110.06920
-
Study of positional encoding approaches for Audio Spectrogram Transformers 13 Oct 2021 · 1 repository · arXiv:2110.06999
-
The Dawn of Quantum Natural Language Processing 13 Oct 2021 · 2 repositories · arXiv:2110.06510
-
Transform and Bitstream Domain Image Classification 13 Oct 2021 · 0 repositories · arXiv:2110.06740
-
Yformer: U-Net Inspired Transformer Architecture for Far Horizon Time Series Forecasting 13 Oct 2021 · 1 repository · arXiv:2110.08255
-
Attention-guided Generative Models for Extractive Question Answering 12 Oct 2021 · 0 repositories · arXiv:2110.06393
-
Contrastive Learning for Representation Degeneration Problem in Sequential Recommendation 12 Oct 2021 · 2 repositories · arXiv:2110.05730Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
DiscoDVT: Generating Long Text with Discourse-Aware Discrete Variational Transformer 12 Oct 2021 · 1 repository · arXiv:2110.05999
-
HETFORMER: Heterogeneous Transformer with Sparse Attention for Long-Text Extractive Summarization 12 Oct 2021 · 1 repository · arXiv:2110.06388
-
LightSeq2: Accelerated Training for Transformer-based Models on GPUs 12 Oct 2021 · 1 repository · arXiv:2110.05722
-
LiST: Lite Prompted Self-training Makes Parameter-Efficient Few-shot Learners 12 Oct 2021 · 1 repository · arXiv:2110.06274Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Mention Memory: incorporating textual knowledge into Transformers through entity mention attention 12 Oct 2021 · 1 repository · arXiv:2110.06176
-
Relative Molecule Self-Attention Transformer 12 Oct 2021 · 1 repository · arXiv:2110.05841Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Rescoring Sequence-to-Sequence Models for Text Line Recognition with CTC-Prefixes 12 Oct 2021 · 1 repository · arXiv:2110.05909
-
Satellite Image Semantic Segmentation 12 Oct 2021 · 1 repository · arXiv:2110.05812
-
StARformer: Transformer with State-Action-Reward Representations for Visual Reinforcement Learning 12 Oct 2021 · 1 repository · arXiv:2110.06206Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Adaptive Multi-view and Temporal Fusing Transformer for 3D Human Pose Estimation 11 Oct 2021 · 0 repositories · arXiv:2110.05092
-
Investigating Transfer Learning Capabilities of Vision Transformers and CNNs by Fine-Tuning a Single Trainable Block 11 Oct 2021 · 1 repository · arXiv:2110.05270
-
Multi-Task Learning for Situated Multi-Domain End-to-End Dialogue Systems 11 Oct 2021 · 0 repositories · arXiv:2110.05221
-
Multi-View Self-Attention Based Transformer for Speaker Recognition 11 Oct 2021 · 0 repositories · arXiv:2110.05036
-
SRU++: Pioneering Fast Recurrence with Attention for Speech Recognition 11 Oct 2021 · 0 repositories · arXiv:2110.05571
-
Unsupervised Source Separation via Bayesian Inference in the Latent Domain 11 Oct 2021 · 1 repository · arXiv:2110.05313
-
DCT: Dynamic Compressive Transformer for Modeling Unbounded Sequence 10 Oct 2021 · 0 repositories · arXiv:2110.04821
-
Multi-Channel End-to-End Neural Diarization with Distributed Microphones 10 Oct 2021 · 0 repositories · arXiv:2110.04694
-
Global Vision Transformer Pruning with Hessian-Aware Saliency 10 Oct 2021 · 1 repository · arXiv:2110.04869Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Automatic Text Extractive Summarization Based on Graph and Pre-trained Language Model Attention 10 Oct 2021 · 0 repositories · arXiv:2110.04878
-
SP-GPT2: Semantics Improvement in Vietnamese Poetry Generation 10 Oct 2021 · 2 repositories · arXiv:2110.15723
-
SuperShaper: Task-Agnostic Super Pre-training of BERT Models with Variable Hidden Dimensions 10 Oct 2021 · 0 repositories · arXiv:2110.04711
-
Yuan 1.0: Large-Scale Pre-trained Language Model in Zero-Shot and Few-Shot Learning 10 Oct 2021 · 1 repository · arXiv:2110.04725
-
Vector-quantized Image Modeling with Improved VQGAN 9 Oct 2021 · 5 repositories · arXiv:2110.04627Syntology 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
The Layout Generation Algorithm of Graphic Design Based on Transformer-CVAE 8 Oct 2021 · 0 repositories · arXiv:2110.06794
-
Adversarial Token Attacks on Vision Transformers 8 Oct 2021 · 0 repositories · arXiv:2110.04337
-
Context-LGM: Leveraging Object-Context Relation for Context-Aware Object Recognition 8 Oct 2021 · 0 repositories · arXiv:2110.04042
-
Towards Learning (Dis)-Similarity of Source Code from Program Contrasts 8 Oct 2021 · 0 repositories · arXiv:2110.03868
-
HydraSum: Disentangling Stylistic Features in Text Summarization using Multi-Decoder Models 8 Oct 2021 · 1 repository · arXiv:2110.04400Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples)
-
KG-FiD: Infusing Knowledge Graph in Fusion-in-Decoder for Open-Domain Question Answering 8 Oct 2021 · 0 repositories · arXiv:2110.04330
-
Knowledge-Enhanced Hierarchical Graph Transformer Network for Multi-Behavior Recommendation 8 Oct 2021 · 1 repository · arXiv:2110.04000
-
Learning Adaptive Control Flow in Transformers for Improved Systematic Generalization 8 Oct 2021 · 0 repositories
-
M6-10T: A Sharing-Delinking Paradigm for Efficient Multi-Trillion Parameter Pretraining 8 Oct 2021 · 0 repositories · arXiv:2110.03888
-
Multiplex Behavioral Relation Learning for Recommendation via Memory Augmented Transformer Network 8 Oct 2021 · 1 repository · arXiv:2110.04002
-
RPT: Toward Transferable Model on Heterogeneous Researcher Data via Pre-Training 8 Oct 2021 · 1 repository · arXiv:2110.07336
-
Taming Sparsely Activated Transformer with Stochastic Experts 8 Oct 2021 · 1 repository · arXiv:2110.04260Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
ViDT: An Efficient and Effective Fully Transformer-based Object Detector 8 Oct 2021 · 1 repository · arXiv:2110.03921
-
A Comparative Study of Transformer-Based Language Models on Extractive Question Answering 7 Oct 2021 · 0 repositories · arXiv:2110.03142
-
Attention is All You Need? Good Embeddings with Statistics are enough:Large Scale Audio Understanding without Transformers/ Convolutions/ BERTs/ Mixers/ Attention/ RNNs or .... 7 Oct 2021 · 0 repositories · arXiv:2110.03183
-
Cross-Language Learning for Entity Matching 7 Oct 2021 · 1 repository · arXiv:2110.03338
-
End-to-End Supermask Pruning: Learning to Prune Image Captioning Models 7 Oct 2021 · 1 repository · arXiv:2110.03298Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Generative Pre-Trained Transformer for Cardiac Abnormality Detection 7 Oct 2021 · 0 repositories · arXiv:2110.04071
-
Investigating Growth at Risk Using a Multi-country Non-parametric Quantile Factor Model 7 Oct 2021 · 0 repositories · arXiv:2110.03411
-
Layer-wise Pruning of Transformer Attention Heads for Efficient Language Modeling 7 Oct 2021 · 1 repository · arXiv:2110.03252
-
Minimum word error training for non-autoregressive Transformer-based code-switching ASR 7 Oct 2021 · 0 repositories · arXiv:2110.03573
-
Adversarial Robustness Comparison of Vision Transformer and MLP-Mixer to CNNs 6 Oct 2021 · 1 repository · arXiv:2110.02797
-
Anomaly Transformer: Time Series Anomaly Detection with Association Discrepancy 6 Oct 2021 · 3 repositories · arXiv:2110.02642Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Dynamically Decoding Source Domain Knowledge for Domain Generalization 6 Oct 2021 · 0 repositories · arXiv:2110.03027
-
Geometric Transformers for Protein Interface Contact Prediction 6 Oct 2021 · 2 repositories · arXiv:2110.02423
-
How BPE Affects Memorization in Transformers 6 Oct 2021 · 0 repositories · arXiv:2110.02782
-
Learning to Iteratively Solve Routing Problems with Dual-Aspect Collaborative Transformer 6 Oct 2021 · 2 repositories · arXiv:2110.02544Syntology official (archive's flag): 7 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples)
-
On Neurons Invariant to Sentence Structural Changes in Neural Machine Translation 6 Oct 2021 · 1 repository · arXiv:2110.03067
-
PoNet: Pooling Network for Efficient Token Mixing in Long Sequences 6 Oct 2021 · 1 repository · arXiv:2110.02442Syntology official: no sample here; runs from other or unrecorded repositories · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Semantic Prediction: Which One Should Come First, Recognition or Prediction? 6 Oct 2021 · 1 repository · arXiv:2110.02829
-
Exploiting Twitter as Source of Large Corpora of Weakly Similar Pairs for Semantic Sentence Embeddings 5 Oct 2021 · 1 repository · arXiv:2110.02030
-
Leveraging the Inductive Bias of Large Language Models for Abstract Textual Reasoning 5 Oct 2021 · 0 repositories · arXiv:2110.02370
-
Sicilian Translator: A Recipe for Low-Resource NMT 5 Oct 2021 · 1 repository · arXiv:2110.01938
-
Sound Event Detection Transformer: An Event-based End-to-End Model for Sound Event Detection 5 Oct 2021 · 1 repository · arXiv:2110.02011
-
Top-N: Equivariant set and graph generation without exchangeability 5 Oct 2021 · 1 repository · arXiv:2110.02096Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Word Acquisition in Neural Language Models 5 Oct 2021 · 1 repository · arXiv:2110.02406
-
Molformer: Motif-based Transformer on 3D Heterogeneous Molecular Graphs 4 Oct 2021 · 2 repositories · arXiv:2110.01191Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
A free lunch from ViT:Adaptive Attention Multi-scale Fusion Transformer for Fine-grained Visual Recognition 4 Oct 2021 · 0 repositories · arXiv:2110.01240
-
DeepA2: A Modular Framework for Deep Argument Analysis with Pretrained Neural Text2Text Language Models 4 Oct 2021 · 1 repository · arXiv:2110.01509
-
Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA 4 Oct 2021 · 0 repositories · arXiv:2110.01552
-
VTAMIQ: Transformers for Attention Modulated Image Quality Assessment 4 Oct 2021 · 1 repository · arXiv:2110.01655
-
Adversarial Examples Generation for Reducing Implicit Gender Bias in Pre-trained Models 3 Oct 2021 · 0 repositories · arXiv:2110.01094