Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 146
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 146 of 190: papers 14,501 to 14,600 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Entity Alignment with Reliable Path Reasoning and Relation-Aware Heterogeneous Graph Transformer 18 May 2022 · 0 repositories · arXiv:2205.08806
-
Evaluation of Transfer Learning for Polish with a Text-to-Text Model 18 May 2022 · 0 repositories · arXiv:2205.08808
-
Leveraging Pseudo-labeled Data to Improve Direct Speech-to-Speech Translation 18 May 2022 · 1 repository · arXiv:2205.08993Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking 18 May 2022 · 0 repositories · arXiv:2205.11245
-
Persian Natural Language Inference: A Meta-learning approach 18 May 2022 · 1 repository · arXiv:2205.08755
-
Transformer based multiple instance learning for weakly supervised histopathology image segmentation 18 May 2022 · 1 repository · arXiv:2205.08878
-
U-Former: Improving Monaural Speech Enhancement with Multi-head Self and Cross Attention 18 May 2022 · 1 repository · arXiv:2205.08681
-
Attention-aware contrastive learning for predicting T cell receptor-antigen binding specificity 17 May 2022 · 0 repositories · arXiv:2206.11255
-
M6-Rec: Generative Pretrained Language Models are Open-Ended Recommender Systems 17 May 2022 · 0 repositories · arXiv:2205.08084
-
MATrIX -- Modality-Aware Transformer for Information eXtraction 17 May 2022 · 0 repositories · arXiv:2205.08094
-
MulT: An End-to-End Multitask Learning Transformer 17 May 2022 · 1 repository · arXiv:2205.08303Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
POViT: Vision Transformer for Multi-objective Design and Characterization of Nanophotonic Devices 17 May 2022 · 0 repositories · arXiv:2205.09045
-
SKILL: Structured Knowledge Infusion for Large Language Models 17 May 2022 · 0 repositories · arXiv:2205.08184
-
The Power of Fragmentation: A Hierarchical Transformer Model for Structural Segmentation in Symbolic Music Generation 17 May 2022 · 0 repositories · arXiv:2205.08579
-
Vision Transformer Adapter for Dense Predictions 17 May 2022 · 2 repositories · arXiv:2205.08534
-
CONSENT: Context Sensitive Transformer for Bold Words Classification 16 May 2022 · 0 repositories · arXiv:2205.07683
-
Heroes, Villains, and Victims, and GPT-3: Automated Extraction of Character Roles Without Training Data 16 May 2022 · 0 repositories · arXiv:2205.07557
-
The AI Teacher Test: Measuring the Pedagogical Ability of Blender and GPT-3 in Educational Dialogues 16 May 2022 · 1 repository · arXiv:2205.07540
-
Transformers in 3D Point Clouds: A Survey 16 May 2022 · 0 repositories · arXiv:2205.07417
-
What GPT Knows About Who is Who 16 May 2022 · 1 repository · arXiv:2205.07407
-
Downstream Transformer Generation of Question-Answer Pairs with Preprocessing and Postprocessing Pipelines 15 May 2022 · 1 repository · arXiv:2205.07387
-
Hero-Gang Neural Model For Named Entity Recognition 15 May 2022 · 1 repository · arXiv:2205.07177
-
Proxyless Neural Architecture Adaptation for Supervised Learning and Self-Supervised Learning 15 May 2022 · 0 repositories · arXiv:2205.07168
-
Transkimmer: Transformer Learns to Layer-wise Skim 15 May 2022 · 1 repository · arXiv:2205.07324Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Video Frame Interpolation with Transformer 15 May 2022 · 1 repository · arXiv:2205.07230
-
Dense residual Transformer for image denoising 14 May 2022 · 0 repositories · arXiv:2205.06944
-
Naturalistic Causal Probing for Morpho-Syntax 14 May 2022 · 1 repository · arXiv:2205.07043
-
RASAT: Integrating Relational Structures into Pretrained Seq2Seq Model for Text-to-SQL 14 May 2022 · 1 repository · arXiv:2205.06983Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Transformer Scale Gate for Semantic Segmentation 14 May 2022 · 0 repositories · arXiv:2205.07056
-
A microstructure estimation Transformer inspired by sparse representation for diffusion MRI 13 May 2022 · 1 repository · arXiv:2205.06450
-
Local Attention Graph-based Transformer for Multi-target Genetic Alteration Prediction 13 May 2022 · 1 repository · arXiv:2205.06672
-
ViT5: Pretrained Text-to-Text Transformer for Vietnamese Language Generation 13 May 2022 · 1 repository · arXiv:2205.06457
-
A Generalist Agent 12 May 2022 · 3 repositories · arXiv:2205.06175Syntology 0 ran · 2 unverified (of 2 harvested samples)
-
Deep Learning for Prawn Farming: Forecasting and Anomaly Detection 12 May 2022 · 0 repositories · arXiv:2205.06359
-
Entity-aware and Motion-aware Transformers for Language-driven Action Localization in Videos 12 May 2022 · 1 repository · arXiv:2205.05854
-
Group R-CNN for Weakly Semi-supervised Object Detection with Points 12 May 2022 · 1 repository · arXiv:2205.05920
-
Multi Task Learning For Zero Shot Performance Prediction of Multilingual Models 12 May 2022 · 0 repositories · arXiv:2205.06130Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
NFLAT: Non-Flat-Lattice Transformer for Chinese Named Entity Recognition 12 May 2022 · 1 repository · arXiv:2205.05832
-
Robot Cooking with Stir-fry: Bimanual Non-prehensile Manipulation of Semi-fluid Objects 12 May 2022 · 1 repository · arXiv:2205.05960
-
Efficient and Training-Free Control of Language Generation 12 May 2022 · 0 repositories · arXiv:2205.06036
-
Simple Open-Vocabulary Object Detection with Vision Transformers 12 May 2022 · 2 repositories · arXiv:2205.06230
-
Supplementary Material: Implementation and Experiments for GAU-based Model 12 May 2022 · 0 repositories · arXiv:2205.05842
-
Vision Transformer: Vit and its Derivatives 12 May 2022 · 0 repositories · arXiv:2205.11239
-
AggPose: Deep Aggregation Vision Transformer for Infant Pose Estimation 11 May 2022 · 1 repository · arXiv:2205.05277Syntology official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
An Empirical Study Of Self-supervised Learning Approaches For Object Detection With Transformers 11 May 2022 · 2 repositories · arXiv:2205.05543
-
Clinical Prompt Learning with Frozen Language Models 11 May 2022 · 1 repository · arXiv:2205.05535
-
CV4Code: Sourcecode Understanding via Visual Code Representations 11 May 2022 · 0 repositories · arXiv:2205.08585
-
Towards the Generation of Musical Explanations with GPT-3 11 May 2022 · 1 repository · arXiv:2206.08264
-
AdMix: A Mixed Sample Data Augmentation Method for Neural Machine Translation 10 May 2022 · 0 repositories · arXiv:2205.04686
-
Ratatouille: A tool for Novel Recipe Generation 10 May 2022 · 0 repositories · arXiv:2206.08267
-
Reduce Information Loss in Transformers for Pluralistic Image Inpainting 10 May 2022 · 1 repository · arXiv:2205.05076
-
Reducing Activation Recomputation in Large Transformer Models 10 May 2022 · 4 repositories · arXiv:2205.05198Syntology official: harvested, nothing ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Spatio-Temporal Transformer for Dynamic Facial Expression Recognition in the Wild 10 May 2022 · 0 repositories · arXiv:2205.04749
-
UL2: Unifying Language Learning Paradigms 10 May 2022 · 2 repositories · arXiv:2205.05131Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples)
-
Activating More Pixels in Image Super-Resolution Transformer 9 May 2022 · 2 repositories · arXiv:2205.04437
-
Incremental-DETR: Incremental Few-Shot Object Detection via Self-Supervised Learning 9 May 2022 · 0 repositories · arXiv:2205.04042
-
Long Document Re-ranking with Modular Re-ranker 9 May 2022 · 1 repository · arXiv:2205.04275
-
Multi-segment preserving sampling for deep manifold sampler 9 May 2022 · 0 repositories · arXiv:2205.04259
-
SwinIQA: Learned Swin Distance for Compressed Image Quality Assessment 9 May 2022 · 1 repository · arXiv:2205.04264
-
Neural Architecture Search using Property Guided Synthesis 8 May 2022 · 1 repository · arXiv:2205.03960
-
Adaptive Graph Convolutional Network Framework for Multidimensional Time Series Prediction 8 May 2022 · 0 repositories · arXiv:2205.04885
-
DxFormer: A Decoupled Automatic Diagnostic System Based on Decoder-Encoder Transformer with Dense Symptom Representations 8 May 2022 · 1 repository · arXiv:2205.03755
-
Mutual Distillation Learning Network for Trajectory-User Linking 8 May 2022 · 1 repository · arXiv:2205.03773Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Vector Representations of Idioms in Conversational Systems 7 May 2022 · 0 repositories · arXiv:2205.03666
-
Explaining the Effectiveness of Multi-Task Learning for Efficient Knowledge Extraction from Spine MRI Reports 6 May 2022 · 0 repositories · arXiv:2205.02979
-
Fake News Detection with Heterogeneous Transformer 6 May 2022 · 1 repository · arXiv:2205.03100
-
RCMNet: A deep learning model assists CAR-T therapy for leukemia 6 May 2022 · 0 repositories · arXiv:2205.04230
-
The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning 6 May 2022 · 1 repository · arXiv:2205.03401
-
Transformer-Based Multi-Aspect Multi-Granularity Non-Native English Speaker Pronunciation Assessment 6 May 2022 · 1 repository · arXiv:2205.03432
-
When a sentence does not introduce a discourse entity, Transformer-based models still sometimes refer to it 6 May 2022 · 1 repository · arXiv:2205.03472
-
A Deep Learning Approach to Dst Index Prediction 5 May 2022 · 0 repositories · arXiv:2205.02447
-
Identifying Cause-and-Effect Relationships of Manufacturing Errors using Sequence-to-Sequence Learning 5 May 2022 · 0 repositories · arXiv:2205.02827
-
RaFoLa: A Rationale-Annotated Corpus for Detecting Indicators of Forced Labour 5 May 2022 · 0 repositories · arXiv:2205.02684
-
Scene Graph Expansion for Semantics-Guided Image Outpainting 5 May 2022 · 0 repositories · arXiv:2205.02958
-
Improving Multi-Document Summarization through Referenced Flexible Extraction with Credit-Awareness 4 May 2022 · 1 repository · arXiv:2205.01889
-
Knowledge Distillation of Russian Language Models with Reduction of Vocabulary 4 May 2022 · 1 repository · arXiv:2205.02340
-
Provably Confidential Language Modelling 4 May 2022 · 1 repository · arXiv:2205.01863Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Better plain ViT baselines for ImageNet-1k 3 May 2022 · 7 repositories · arXiv:2205.01580Syntology official: no sample here; runs from other or unrecorded repositories · 23 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 5 honoured, 1 violated, 14 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 24 harvested samples) · 4 pointer-only (licence)
-
Contrastive Learning for Prompt-Based Few-Shot Language Learners 3 May 2022 · 1 repository · arXiv:2205.01308Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
MTTrans: Cross-Domain Object Detection with Mean-Teacher Transformer 3 May 2022 · 1 repository · arXiv:2205.01643
-
Learn To Remember: Transformer with Recurrent Memory for Document-Level Machine Translation 3 May 2022 · 0 repositories · arXiv:2205.01546
-
Mixed-effects transformers for hierarchical adaptation 3 May 2022 · 1 repository · arXiv:2205.01749
-
Neural Language Taskonomy: Which NLP Tasks are the most Predictive of fMRI Brain Activity? 3 May 2022 · 0 repositories · arXiv:2205.01404
-
Synthesized Speech Detection Using Convolutional Transformer-Based Spectrogram Analysis 3 May 2022 · 0 repositories · arXiv:2205.01800
-
Textual Entailment for Event Argument Extraction: Zero- and Few-Shot with Multi-Source Learning 3 May 2022 · 1 repository · arXiv:2205.01376
-
Logiformer: A Two-Branch Graph Transformer Network for Interpretable Logical Reasoning 2 May 2022 · 1 repository · arXiv:2205.00731
-
Multi-Task Text Classification using Graph Convolutional Networks for Large-Scale Low Resource Language 2 May 2022 · 1 repository · arXiv:2205.01204
-
OPT: Open Pre-trained Transformer Language Models 2 May 2022 · 11 repositories · arXiv:2205.01068Syntology official (archive's flag): 9 ran · 14 ran (of which 2 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified (of 24 harvested samples) · 17 pointer-only (licence)
-
Teaching BERT to Wait: Balancing Accuracy and Latency for Streaming Disfluency Detection 2 May 2022 · 0 repositories · arXiv:2205.00620
-
Aircraft engine remaining useful life estimation via a double attention-based data-driven architecture 1 May 2022 · 0 repositories
-
Classification without (Proper) Representation: Political Heterogeneity in Social Media and Its Implications for Classification and Behavioral Analysis 1 May 2022 · 0 repositories
-
Reinforced Swin-Convs Transformer for Underwater Image Enhancement 1 May 2022 · 1 repository · arXiv:2205.00434Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
Coarse-to-Fine Video Denoising with Dual-Stage Spatial-Channel Transformer 30 Apr 2022 · 0 repositories · arXiv:2205.00214
-
HDGT: Heterogeneous Driving Graph Transformer for Multi-Agent Trajectory Prediction via Scene Encoding 30 Apr 2022 · 2 repositories · arXiv:2205.09753
-
StorSeismic: A new paradigm in deep learning for seismic processing 30 Apr 2022 · 1 repository · arXiv:2205.00222
-
Training Language Models with Language Feedback 29 Apr 2022 · 0 repositories · arXiv:2204.14146
-
Two New Datasets for Italian-Language Abstractive Text Summarization 29 Apr 2022 · 1 repository
-
Controllable Image Captioning 28 Apr 2022 · 0 repositories · arXiv:2204.13324
-
Depth Estimation with Simplified Transformer 28 Apr 2022 · 0 repositories · arXiv:2204.13791
-
HiNER: A Large Hindi Named Entity Recognition Dataset 28 Apr 2022 · 1 repository · arXiv:2204.13743