Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 156
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 156 of 190: papers 15,501 to 15,600 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
End-to-End Referring Video Object Segmentation with Multimodal Transformers 29 Nov 2021 · 2 repositories · arXiv:2111.14821Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
Mixed Precision Low-bit Quantization of Neural Network Language Models for Speech Recognition 29 Nov 2021 · 0 repositories · arXiv:2112.11438
-
Mixed Precision of Quantization of Transformer Language Models for Speech Recognition 29 Nov 2021 · 0 repositories · arXiv:2112.11540
-
On the rate of convergence of a classifier based on a Transformer encoder 29 Nov 2021 · 0 repositories · arXiv:2111.14574
-
Point-BERT: Pre-training 3D Point Cloud Transformers with Masked Point Modeling 29 Nov 2021 · 3 repositories · arXiv:2111.14819Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Recurrent Vision Transformer for Solving Visual Reasoning Problems 29 Nov 2021 · 0 repositories · arXiv:2111.14576
-
Searching the Search Space of Vision Transformer 29 Nov 2021 · 2 repositories · arXiv:2111.14725
-
Sparse DETR: Efficient End-to-End Object Detection with Learnable Sparsity 29 Nov 2021 · 1 repository · arXiv:2111.14330Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 7 harvested samples)
-
TransMVSNet: Global Context-aware Multi-view Stereo Network with Transformers 29 Nov 2021 · 1 repository · arXiv:2111.14600
-
Context Matters in Semantically Controlled Language Generation for Task-oriented Dialogue Systems 28 Nov 2021 · 0 repositories · arXiv:2111.14119
-
FastTrees: Parallel Latent Tree-Induction for Faster Sequence Encoding 28 Nov 2021 · 1 repository · arXiv:2111.14031
-
Multi-domain Integrative Swin Transformer network for Sparse-View Tomographic Reconstruction 28 Nov 2021 · 0 repositories · arXiv:2111.14831
-
ORCHARD: A Benchmark For Measuring Systematic Generalization of Multi-Hierarchical Reasoning 28 Nov 2021 · 1 repository · arXiv:2111.14034
-
Abusive and Threatening Language Detection in Urdu using Boosting based and BERT based models: A Comparative Approach 27 Nov 2021 · 1 repository · arXiv:2111.14830
-
Exploring Transformer Based Models to Identify Hate Speech and Offensive Content in English and Indo-Aryan Languages 27 Nov 2021 · 0 repositories · arXiv:2111.13974
-
FQ-ViT: Post-Training Quantization for Fully Quantized Vision Transformer 27 Nov 2021 · 1 repository · arXiv:2111.13824Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
GEOSCAN: Global Earth Observation using Swarm of Coordinated Autonomous Nanosats 27 Nov 2021 · 0 repositories · arXiv:2111.15627
-
Learning A 3D-CNN and Transformer Prior for Hyperspectral Image Super-Resolution 27 Nov 2021 · 0 repositories · arXiv:2111.13923
-
A Robust Volumetric Transformer for Accurate 3D Tumor Segmentation 26 Nov 2021 · 1 repository · arXiv:2111.13300Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Exploiting full Resolution Feature Context for Liver Tumor and Vessel Segmentation via Integrate Framework: Application to Liver Tumor and Vessel 3D Reconstruction under embedded microprocessor 26 Nov 2021 · 1 repository · arXiv:2111.13299
-
GMFlow: Learning Optical Flow via Global Matching 26 Nov 2021 · 4 repositories · arXiv:2111.13680Syntology official (archive's flag): 11 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 19 harvested samples) · 7 pointer-only (licence)
-
SWAT: Spatial Structure Within and Among Tokens 26 Nov 2021 · 1 repository · arXiv:2111.13677
-
Domain Prompt Learning for Efficiently Adapting CLIP to Unseen Domains 25 Nov 2021 · 1 repository · arXiv:2111.12853Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Attend to Who You Are: Supervising Self-Attention for Keypoint Detection and Instance-Aware Association 25 Nov 2021 · 1 repository · arXiv:2111.12892
-
BoxeR: Box-Attention for 2D and 3D Transformers 25 Nov 2021 · 1 repository · arXiv:2111.13087
-
Exploiting Both Domain-specific and Invariant Knowledge via a Win-win Transformer for Unsupervised Domain Adaptation 25 Nov 2021 · 0 repositories · arXiv:2111.12941
-
Global Interaction Modelling in Vision Transformer via Super Tokens 25 Nov 2021 · 0 repositories · arXiv:2111.13156
-
Scene Representation Transformer: Geometry-Free Novel View Synthesis Through Set-Latent Scene Representations 25 Nov 2021 · 1 repository · arXiv:2111.13152Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Transformer-based Korean Pretrained Language Models: A Survey on Three Years of Progress 25 Nov 2021 · 0 repositories · arXiv:2112.03014
-
TunBERT: Pretrained Contextualized Text Representation for Tunisian Dialect 25 Nov 2021 · 0 repositories · arXiv:2111.13138
-
Attention-based Dual-stream Vision Transformer for Radar Gait Recognition 24 Nov 2021 · 0 repositories · arXiv:2111.12290
-
MHFormer: Multi-Hypothesis Transformer for 3D Human Pose Estimation 24 Nov 2021 · 1 repository · arXiv:2111.12707Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
MorphMLP: An Efficient MLP-Like Backbone for Spatial-Temporal Representation Learning 24 Nov 2021 · 2 repositories · arXiv:2111.12527Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Sparse is Enough in Scaling Transformers 24 Nov 2021 · 0 repositories · arXiv:2111.12763
-
Unleashing Transformers: Parallel Token Prediction with Discrete Absorbing Diffusion for Fast High-Resolution Image Generation from Vector-Quantized Codes 24 Nov 2021 · 3 repositories · arXiv:2111.12701
-
Utilizing Resource-Rich Language Datasets for End-to-End Scene Text Recognition in Resource-Poor Languages 24 Nov 2021 · 0 repositories · arXiv:2111.12276
-
VIOLET : End-to-End Video-Language Transformers with Masked Visual-token Modeling 24 Nov 2021 · 1 repository · arXiv:2111.12681Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Deep Point Cloud Reconstruction 23 Nov 2021 · 0 repositories · arXiv:2111.11704
-
Multi-Person 3D Motion Prediction with Multi-Range Transformers 23 Nov 2021 · 1 repository · arXiv:2111.12073
-
Pruning Self-attentions into Convolutional Layers in Single Path 23 Nov 2021 · 3 repositories · arXiv:2111.11802Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
SimpleTRON: Simple Transformer with O(N) Complexity 23 Nov 2021 · 0 repositories · arXiv:2111.15588
-
S-SimCSE: Sampled Sub-networks for Contrastive Learning of Sentence Embedding 23 Nov 2021 · 0 repositories · arXiv:2111.11750
-
Self-Supervised Pre-Training for Transformer-Based Person Re-Identification 23 Nov 2021 · 3 repositories · arXiv:2111.12084Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
U-shape Transformer for Underwater Image Enhancement 23 Nov 2021 · 2 repositories · arXiv:2111.11843Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Variational Learning for Unsupervised Knowledge Grounded Dialogs 23 Nov 2021 · 1 repository · arXiv:2112.00653
-
Benchmarking Detection Transfer Learning with Vision Transformers 22 Nov 2021 · 2 repositories · arXiv:2111.11429
-
DBIA: Data-free Backdoor Injection Attack against Transformer Networks 22 Nov 2021 · 1 repository · arXiv:2111.11870
-
ExT5: Towards Extreme Multi-Task Scaling for Transfer Learning 22 Nov 2021 · 4 repositories · arXiv:2111.10952Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
L-Verse: Bidirectional Generation Between Image and Text 22 Nov 2021 · 1 repository · arXiv:2111.11133
-
MetaFormer Is Actually What You Need for Vision 22 Nov 2021 · 18 repositories · arXiv:2111.11418Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Patch Vestiges in the Adversarial Examples Against Vision Transformer Can Be Leveraged for Adversarial Detection 22 Nov 2021 · 0 repositories
-
Lightweight Transformer Backbone for Medical Object Detection 22 Nov 2021 · 0 repositories · arXiv:2111.11546
-
Semi-Supervised Vision Transformers 22 Nov 2021 · 1 repository · arXiv:2111.11067
-
CpT: Convolutional Point Transformer for 3D Point Cloud Processing 21 Nov 2021 · 0 repositories · arXiv:2111.10866
-
DuDoTrans: Dual-Domain Transformer Provides More Attention for Sinogram Restoration in Sparse-View CT Reconstruction 21 Nov 2021 · 0 repositories · arXiv:2111.10790
-
Efficient Softmax Approximation for Deep Neural Networks with Attention Mechanism 21 Nov 2021 · 0 repositories · arXiv:2111.10770
-
Are Vision Transformers Robust to Patch Perturbations? 20 Nov 2021 · 0 repositories · arXiv:2111.10659
-
Data Processing Matters: SRPH-Konvergen AI's Machine Translation System for WMT'21 20 Nov 2021 · 0 repositories · arXiv:2111.10513
-
Discrete Representations Strengthen Vision Transformer Robustness 20 Nov 2021 · 1 repository · arXiv:2111.10493
-
Advancing High-Resolution Video-Language Representation with Large-Scale Video Transcriptions 19 Nov 2021 · 1 repository · arXiv:2111.10337
-
DyFormer: A Scalable Dynamic Graph Transformer with Provable Benefits on Generalization Ability 19 Nov 2021 · 0 repositories · arXiv:2111.10447
-
Generalized Decision Transformer for Offline Hindsight Information Matching 19 Nov 2021 · 1 repository · arXiv:2111.10364Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Lexicon-based Methods vs. BERT for Text Sentiment Analysis 19 Nov 2021 · 0 repositories · arXiv:2111.10097
-
TransMorph: Transformer for unsupervised medical image registration 19 Nov 2021 · 2 repositories · arXiv:2111.10480Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
PatchCensor: Patch Robustness Certification for Transformers via Exhaustive Testing 19 Nov 2021 · 0 repositories · arXiv:2111.10481
-
ClipCap: CLIP Prefix for Image Captioning 18 Nov 2021 · 4 repositories · arXiv:2111.09734Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
How much do language models copy from their training data? Evaluating linguistic novelty in text generation using RAVEN 18 Nov 2021 · 0 repositories · arXiv:2111.09509
-
Multimodal Emotion Recognition on RAVDESS Dataset Using Transfer Learning 18 Nov 2021 · 0 repositories
-
Reference-based Magnetic Resonance Image Reconstruction Using Texture Transformer 18 Nov 2021 · 0 repositories · arXiv:2111.09492
-
Restormer: Efficient Transformer for High-Resolution Image Restoration 18 Nov 2021 · 13 repositories · arXiv:2111.09881Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
RoBERTuito: a pre-trained language model for social media text in Spanish 18 Nov 2021 · 1 repository · arXiv:2111.09453Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Swin Transformer V2: Scaling Up Capacity and Resolution 18 Nov 2021 · 23 repositories · arXiv:2111.09883Syntology community repositories only · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 3 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 13 unverified (of 30 harvested samples) · 4 pointer-only (licence)
-
You Only Sample (Almost) Once: Linear Cost Self-Attention Via Bernoulli Sampling 18 Nov 2021 · 1 repository · arXiv:2111.09714Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Guiding Generative Language Models for Data Augmentation in Few-Shot Text Classification 17 Nov 2021 · 0 repositories · arXiv:2111.09064
-
A Comparative Study on Transfer Learning and Distance Metrics in Semantic Clustering over the COVID-19 Tweets 16 Nov 2021 · 0 repositories · arXiv:2111.08658
-
A Deep Generative XAI Framework for Natural Language Inference Explanations Generation 16 Nov 2021 · 0 repositories
-
Active Dialogue Simulation in Conversational Systems 16 Nov 2021 · 0 repositories
-
All Birds with One Stone: Multi-task Learning for Inference with One Forward Pass 16 Nov 2021 · 0 repositories
-
An Empirical Study of Document-to-document Neural Machine Translation 16 Nov 2021 · 0 repositories
-
An Information Theoretic Measurement of Topical Relevance in Learner Essays 16 Nov 2021 · 0 repositories
-
ANNA: Enhanced Language Representation for Question Answering 16 Nov 2021 · 0 repositories
-
Attention-based Multi-hypothesis Fusion for Speech Summarization 16 Nov 2021 · 2 repositories · arXiv:2111.08201
-
BARCOR: Towards A Unified Framework for Conversational Recommendation 16 Nov 2021 · 0 repositories
-
Building Chinese Biomedical Language Models via Multi-Level Text Discrimination 16 Nov 2021 · 1 repository
-
CalBERT - Code-mixed Adaptive Language representations using BERT 16 Nov 2021 · 0 repositories
-
Causal Transformers: Improving the Robustness on Spurious Correlations 16 Nov 2021 · 0 repositories
-
Challenges for Open-domain Targeted Sentiment Analysis 16 Nov 2021 · 0 repositories
-
Coloring the Blank Slate: Pre-training Imparts a Hierarchical Inductive Bias to Sequence-to-sequence Models 16 Nov 2021 · 0 repositories
-
Compressing Sentence Representation via Homomorphic Projective Distillation 16 Nov 2021 · 0 repositories
-
Constructing Phrase-level Semantic Labels to Form Multi-GrainedSupervision for Image-Text Retrieval 16 Nov 2021 · 0 repositories
-
CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation 16 Nov 2021 · 0 repositories
-
CST5: Data augmentation for Code-Switched Semantic Parsing 16 Nov 2021 · 0 repositories
-
DAML-ST5: Low Resource Style Transfer via Domain Adaptive Meta Learning 16 Nov 2021 · 0 repositories
-
Data Augmentation and Learned Layer Aggregation for Improved Multilingual Language Understanding in Dialogue 16 Nov 2021 · 0 repositories
-
Data Augmentation for Intent Classification with Generic Large Language Models 16 Nov 2021 · 0 repositories
-
Efficient Long Sequence Encoding via Synchronization 16 Nov 2021 · 0 repositories
-
ElitePLM: An Empirical Study on General Language Ability Evaluation of Pretrained Language Models 16 Nov 2021 · 0 repositories
-
ELLE: Efficient Lifelong Pre-training for Emerging Data 16 Nov 2021 · 0 repositories
-
Empirical Analysis of Training Strategies of Transformer-based Japanese Chit-chat Systems 16 Nov 2021 · 0 repositories
-
End-To-End Sign Language Translation via Multitask Learning 16 Nov 2021 · 0 repositories