Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 147
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 147 of 190: papers 14,601 to 14,700 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Inferring Implicit Relations in Complex Questions with Language Models 28 Apr 2022 · 1 repository · arXiv:2204.13778
-
Lightweight Bimodal Network for Single-Image Super-Resolution via Symmetric CNN and Recursive Transformer 28 Apr 2022 · 1 repository · arXiv:2204.13286
-
On the Effect of Pretraining Corpora on In-context Learning by a Large-scale Language Model 28 Apr 2022 · 0 repositories · arXiv:2204.13509
-
One Model to Synthesize Them All: Multi-contrast Multi-scale Transformer for Missing Data Imputation 28 Apr 2022 · 0 repositories · arXiv:2204.13738
-
Symmetric Transformer-based Network for Unsupervised Image Registration 28 Apr 2022 · 1 repository · arXiv:2204.13575
-
Tag-assisted Multimodal Sentiment Analysis under Uncertain Missing Modalities 28 Apr 2022 · 1 repository · arXiv:2204.13707
-
Tailor: A Prompt-Based Approach to Attribute-Based Controlled Text Generation 28 Apr 2022 · 0 repositories · arXiv:2204.13362
-
Transformers in Time-series Analysis: A Tutorial 28 Apr 2022 · 0 repositories · arXiv:2205.01138
-
UM6P-CS at SemEval-2022 Task 11: Enhancing Multilingual and Code-Mixed Complex Named Entity Recognition via Pseudo Labels using Multilingual Transformer 28 Apr 2022 · 0 repositories · arXiv:2204.13515
-
An End-to-End Dialogue Summarization System for Sales Calls 27 Apr 2022 · 0 repositories · arXiv:2204.12951
-
CATrans: Context and Affinity Transformer for Few-Shot Segmentation 27 Apr 2022 · 0 repositories · arXiv:2204.12817
-
Building Knowledge-Grounded Dialogue Systems with Graph-Based Semantic Modeling 27 Apr 2022 · 0 repositories · arXiv:2204.12681
-
Modern Baselines for SPARQL Semantic Parsing 27 Apr 2022 · 1 repository · arXiv:2204.12793
-
Ultra Fast Speech Separation Model with Teacher Student Learning 27 Apr 2022 · 0 repositories · arXiv:2204.12777
-
Adaptive Split-Fusion Transformer 26 Apr 2022 · 1 repository · arXiv:2204.12196
-
Crystal Transformer: Self-learning neural language model for Generative and Tinkering Design of Materials 25 Apr 2022 · 0 repositories · arXiv:2204.11953
-
OCFormer: One-Class Transformer Network for Image Classification 25 Apr 2022 · 0 repositories · arXiv:2204.11449
-
Performer: A Novel PPG-to-ECG Reconstruction Transformer for a Digital Biomarker of Cardiovascular Disease Detection 25 Apr 2022 · 0 repositories · arXiv:2204.11795
-
Predicting Real-time Scientific Experiments Using Transformer models and Reinforcement Learning 25 Apr 2022 · 1 repository · arXiv:2204.11718
-
SwinFuse: A Residual Swin Transformer Fusion Network for Infrared and Visible Images 25 Apr 2022 · 1 repository · arXiv:2204.11436
-
Local Gaussian process extrapolation for BART models with applications to causal inference 23 Apr 2022 · 0 repositories · arXiv:2204.10963
-
Time Series Forecasting (TSF) Using Various Deep Learning Models 23 Apr 2022 · 0 repositories · arXiv:2204.11115
-
DFAM-DETR: Deformable feature based attention mechanism DETR on slender object detection 22 Apr 2022 · 0 repositories · arXiv:2204.10667
-
Diverse Instance Discovery: Vision-Transformer for Instance-Aware Multi-Label Image Recognition 22 Apr 2022 · 0 repositories · arXiv:2204.10731
-
End-to-end symbolic regression with transformers 22 Apr 2022 · 3 repositories · arXiv:2204.10532Syntology official (archive's flag): 1 ran · 4 ran (of which 1 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Hierarchical Label-wise Attention Transformer Model for Explainable ICD Coding 22 Apr 2022 · 1 repository · arXiv:2204.10716
-
Hypergraph Transformer: Weakly-supervised Multi-hop Reasoning for Knowledge-based Visual Question Answering 22 Apr 2022 · 1 repository · arXiv:2204.10448
-
Spatiality-guided Transformer for 3D Dense Captioning on Point Clouds 22 Apr 2022 · 1 repository · arXiv:2204.10688
-
Unified Pretraining Framework for Document Understanding 22 Apr 2022 · 0 repositories · arXiv:2204.10939
-
BTranspose: Bottleneck Transformers for Human Pose Estimation with Self-Supervised Pre-Training 21 Apr 2022 · 0 repositories · arXiv:2204.10209
-
Measuring artificial intelligence: a systematic assessment and implications for governance 21 Apr 2022 · 0 repositories · arXiv:2204.10304
-
Multi-Tier Platform for Cognizing Massive Electroencephalogram 21 Apr 2022 · 0 repositories · arXiv:2204.09840
-
SinTra: Learning an inspiration model from a single multi-track music segment 21 Apr 2022 · 1 repository · arXiv:2204.09917
-
Transformer-Guided Convolutional Neural Network for Cross-View Geolocalization 21 Apr 2022 · 0 repositories · arXiv:2204.09967
-
Detecting Unintended Memorization in Language-Model-Fused ASR 20 Apr 2022 · 0 repositories · arXiv:2204.09606
-
Human-Object Interaction Detection via Disentangled Transformer 20 Apr 2022 · 0 repositories · arXiv:2204.09290
-
NFormer: Robust Person Re-identification with Neighbor Transformer 20 Apr 2022 · 1 repository · arXiv:2204.09331Syntology official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 6 harvested samples)
-
Reinforced Structured State-Evolution for Vision-Language Navigation 20 Apr 2022 · 1 repository · arXiv:2204.09280
-
Towards Arabic Sentence Simplification via Classification and Generative Approaches 20 Apr 2022 · 0 repositories · arXiv:2204.09292
-
DiffMD: A Geometric Diffusion Model for Molecular Dynamics Simulations 19 Apr 2022 · 0 repositories · arXiv:2204.08672
-
Blockwise Streaming Transformer for Spoken Language Understanding and Simultaneous Speech Translation 19 Apr 2022 · 0 repositories · arXiv:2204.08920
-
CodexDB: Generating Code for Processing SQL Queries using GPT-3 Codex 19 Apr 2022 · 0 repositories · arXiv:2204.08941
-
CTCNet: A CNN-Transformer Cooperation Network for Face Image Super-Resolution 19 Apr 2022 · 1 repository · arXiv:2204.08696
-
DecBERT: Enhancing the Language Understanding of BERT with Causal Attention Masks 19 Apr 2022 · 0 repositories · arXiv:2204.08688
-
Impact of Tokenization on Language Models: An Analysis for Turkish 19 Apr 2022 · 0 repositories · arXiv:2204.08832
-
MANIQA: Multi-dimension Attention Network for No-Reference Image Quality Assessment 19 Apr 2022 · 2 repositories · arXiv:2204.08958Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Multi-View Spatial-Temporal Network for Continuous Sign Language Recognition 19 Apr 2022 · 0 repositories · arXiv:2204.08747
-
Not All Tokens Are Equal: Human-centric Visual Analysis via Token Clustering Transformer 19 Apr 2022 · 1 repository · arXiv:2204.08680Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Self-Calibrated Efficient Transformer for Lightweight Super-Resolution 19 Apr 2022 · 1 repository · arXiv:2204.08913
-
Application of Transfer Learning and Ensemble Learning in Image-level Classification for Breast Histopathology 18 Apr 2022 · 0 repositories · arXiv:2204.08311
-
BSRT: Improving Burst Super-Resolution with Swin Transformer and Flow-Guided Deformable Alignment 18 Apr 2022 · 1 repository · arXiv:2204.08332Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
Dynamic Position Encoding for Transformers 18 Apr 2022 · 0 repositories · arXiv:2204.08142
-
MASSIVE: A 1M-Example Multilingual Natural Language Understanding Dataset with 51 Typologically-Diverse Languages 18 Apr 2022 · 6 repositories · arXiv:2204.08582Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Temporally Efficient Vision Transformer for Video Instance Segmentation 18 Apr 2022 · 3 repositories · arXiv:2204.08412Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Visio-Linguistic Brain Encoding 18 Apr 2022 · 0 repositories · arXiv:2204.08261
-
Zero-shot Entity and Tweet Characterization with Designed Conditional Prompts and Contexts 18 Apr 2022 · 0 repositories · arXiv:2204.08405
-
An Extendable, Efficient and Effective Transformer-based Object Detector 17 Apr 2022 · 1 repository · arXiv:2204.07962Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Continual Hippocampus Segmentation with Transformers 17 Apr 2022 · 1 repository · arXiv:2204.08043
-
MST++: Multi-stage Spectral-wise Transformer for Efficient Spectral Reconstruction 17 Apr 2022 · 3 repositories · arXiv:2204.07908Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
ParkPredict+: Multimodal Intent and Motion Prediction for Vehicles in Parking Lots with CNN and Transformer 17 Apr 2022 · 1 repository · arXiv:2204.10777
-
VDTR: Video Deblurring with Transformer 17 Apr 2022 · 1 repository · arXiv:2204.08023
-
A Hierarchical N-Gram Framework for Zero-Shot Link Prediction 16 Apr 2022 · 1 repository · arXiv:2204.10293
-
Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks 16 Apr 2022 · 10 repositories · arXiv:2204.07705Syntology community repositories only · 17 ran (of which 1 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 1 where Syntology's instrument failed) · 11 unverified (of 28 harvested samples) · 4 pointer-only (licence)
-
BLISS: Robust Sequence-to-Sequence Learning via Self-Supervised Input Representation 16 Apr 2022 · 0 repositories · arXiv:2204.07837
-
Towards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasks 16 Apr 2022 · 1 repository · arXiv:2204.07780
-
WordAlchemy: A transformer-based Reverse Dictionary 16 Apr 2022 · 0 repositories · arXiv:2204.10181
-
ERGO: Event Relational Graph Transformer for Document-level Event Causality Identification 15 Apr 2022 · 0 repositories · arXiv:2204.07434
-
Image Captioning In the Transformer Age 15 Apr 2022 · 1 repository · arXiv:2204.07374
-
mGPT: Few-Shot Learners Go Multilingual 15 Apr 2022 · 1 repository · arXiv:2204.07580
-
Mixture of Experts for Biomedical Question Answering 15 Apr 2022 · 0 repositories · arXiv:2204.07469
-
ML_LTU at SemEval-2022 Task 4: T5 Towards Identifying Patronizing and Condescending Language 15 Apr 2022 · 0 repositories · arXiv:2204.07432
-
MVSTER: Epipolar Transformer for Efficient Multi-View Stereo 15 Apr 2022 · 1 repository · arXiv:2204.07346Syntology official (archive's flag): 10 ran · 10 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 7 unverified (of 17 harvested samples)
-
On the Role of Pre-trained Language Models in Word Ordering: A Case Study with BART 15 Apr 2022 · 1 repository · arXiv:2204.07367
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 15 Apr 2022 · 1 repository · arXiv:2204.07483
-
Resource-Aware Distributed Submodular Maximization: A Paradigm for Multi-Robot Decision-Making 15 Apr 2022 · 0 repositories · arXiv:2204.07520
-
ResT V2: Simpler, Faster and Stronger 15 Apr 2022 · 2 repositories · arXiv:2204.07366
-
Text Revision by On-the-Fly Representation Optimization 15 Apr 2022 · 1 repository · arXiv:2204.07359
-
Unconditional Image-Text Pair Generation with Multimodal Cross Quantizer 15 Apr 2022 · 1 repository · arXiv:2204.07537
-
Activation Regression for Continuous Domain Generalization with Applications to Crop Classification 14 Apr 2022 · 1 repository · arXiv:2204.07030
-
Analysing similarities between legal court documents using natural language processing approaches based on Transformers 14 Apr 2022 · 0 repositories · arXiv:2204.07182
-
Causal Transformer for Estimating Counterfactual Outcomes 14 Apr 2022 · 1 repository · arXiv:2204.07258
-
Challenges for Open-domain Targeted Sentiment Analysis 14 Apr 2022 · 0 repositories · arXiv:2204.06893
-
DeiT III: Revenge of the ViT 14 Apr 2022 · 12 repositories · arXiv:2204.07118
-
Generative power of a protein language model trained on multiple sequence alignments 14 Apr 2022 · 1 repository · arXiv:2204.07110
-
GPT-NeoX-20B: An Open-Source Autoregressive Language Model 14 Apr 2022 · 11 repositories · arXiv:2204.06745Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Hierarchical Embedded Bayesian Additive Regression Trees 14 Apr 2022 · 0 repositories · arXiv:2204.07207
-
MiniViT: Compressing Vision Transformers with Weight Multiplexing 14 Apr 2022 · 2 repositories · arXiv:2204.07154Syntology official: harvested for another paper · 0 ran · 4 unverified (of 4 harvested samples)
-
Residual Swin Transformer Channel Attention Network for Image Demosaicing 14 Apr 2022 · 0 repositories · arXiv:2204.07098
-
Rows from Many Sources: Enriching row completions from Wikidata with a pre-trained Language Model 14 Apr 2022 · 0 repositories · arXiv:2204.07014
-
Deep Relation Learning for Regression and Its Application to Brain Age Estimation 13 Apr 2022 · 0 repositories · arXiv:2204.06598
-
Fix Bugs with Transformer through a Neural-Symbolic Edit Grammar 13 Apr 2022 · 0 repositories · arXiv:2204.06643
-
Formal Language Recognition by Hard Attention Transformers: Perspectives from Circuit Complexity 13 Apr 2022 · 0 repositories · arXiv:2204.06618
-
Recognition of Freely Selected Keypoints on Human Limbs 13 Apr 2022 · 0 repositories · arXiv:2204.06326
-
Building Markovian Generative Architectures over Pretrained LM Backbones for Efficient Task-Oriented Dialog Systems 13 Apr 2022 · 2 repositories · arXiv:2204.06452
-
TangoBERT: Reducing Inference Cost by using Cascaded Architecture 13 Apr 2022 · 0 repositories · arXiv:2204.06271
-
Are Multimodal Transformers Robust to Missing Modality? 12 Apr 2022 · 0 repositories · arXiv:2204.05454
-
Explore More Guidance: A Task-aware Instruction Network for Sign Language Translation Enhanced with Data Augmentation 12 Apr 2022 · 1 repository · arXiv:2204.05953
-
Few-shot Learning with Noisy Labels 12 Apr 2022 · 1 repository · arXiv:2204.05494
-
HiTPR: Hierarchical Transformer for Place Recognition in Point Cloud 12 Apr 2022 · 0 repositories · arXiv:2204.05481
-
L3Cube-MahaNER: A Marathi Named Entity Recognition Dataset and BERT models 12 Apr 2022 · 1 repository · arXiv:2204.06029