Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 130
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 130 of 190: papers 12,901 to 13,000 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Rethinking with Retrieval: Faithful Large Language Model Inference 31 Dec 2022 · 1 repository · arXiv:2301.00303
-
TransIFC: Invariant Cues-aware Feature Concentration Learning for Efficient Fine-grained Bird Image Classification 31 Dec 2022 · 0 repositories
-
Memory Augmented Lookup Dictionary based Language Modeling for Automatic Speech Recognition 30 Dec 2022 · 0 repositories · arXiv:2301.00066
-
Inconsistencies in Masked Language Models 30 Dec 2022 · 1 repository · arXiv:2301.00068
-
Targeted Phishing Campaigns using Large Scale Language Models 30 Dec 2022 · 0 repositories · arXiv:2301.00665
-
Transformer in Transformer as Backbone for Deep Reinforcement Learning 30 Dec 2022 · 1 repository · arXiv:2212.14538
-
Efficient Image Super-Resolution with Feature Interaction Weighted Hybrid Network 29 Dec 2022 · 1 repository · arXiv:2212.14181
-
Efficient Movie Scene Detection using State-Space Transformers 29 Dec 2022 · 1 repository · arXiv:2212.14427Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Error syntax aware augmentation of feedback comment generation dataset 29 Dec 2022 · 0 repositories · arXiv:2212.14293
-
Exploring Depth Information for Face Manipulation Detection 29 Dec 2022 · 0 repositories · arXiv:2212.14230
-
GPT Takes the Bar Exam 29 Dec 2022 · 5 repositories · arXiv:2212.14402
-
Maximizing Use-Case Specificity through Precision Model Tuning 29 Dec 2022 · 0 repositories · arXiv:2212.14206
-
Robust representations of oil wells' intervals via sparse attention mechanism 29 Dec 2022 · 2 repositories · arXiv:2212.14246
-
Part-guided Relational Transformers for Fine-grained Visual Recognition 28 Dec 2022 · 1 repository · arXiv:2212.13685
-
RevealED: Uncovering Pro-Eating Disorder Content on Twitter Using Deep Learning 28 Dec 2022 · 0 repositories · arXiv:2212.13949
-
Swin MAE: Masked Autoencoders for Small Datasets 28 Dec 2022 · 1 repository · arXiv:2212.13805
-
Thermal Heating in ReRAM Crossbar Arrays: Challenges and Solutions 28 Dec 2022 · 0 repositories · arXiv:2212.13707
-
1st Place Solution for YouTubeVOS Challenge 2022: Referring Video Object Segmentation 27 Dec 2022 · 1 repository · arXiv:2212.14679
-
A Generalization of ViT/MLP-Mixer to Graphs 27 Dec 2022 · 3 repositories · arXiv:2212.13350Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
BART-IT: An Efficient Sequence-to-Sequence Model for Italian Text Summarization 27 Dec 2022 · 1 repository
-
Countering Malicious Content Moderation Evasion in Online Social Networks: Simulation and Detection of Word Camouflage 27 Dec 2022 · 1 repository · arXiv:2212.14727
-
DAE-Former: Dual Attention-guided Efficient Transformer for Medical Image Segmentation 27 Dec 2022 · 1 repository · arXiv:2212.13504
-
DeepCuts: Single-Shot Interpretability based Pruning for BERT 27 Dec 2022 · 1 repository · arXiv:2212.13392
-
Exploring Transformer Backbones for Image Diffusion Models 27 Dec 2022 · 0 repositories · arXiv:2212.14678
-
TegFormer: Topic-to-Essay Generation with Good Topic Coverage and High Text Coherence 27 Dec 2022 · 0 repositories · arXiv:2212.13456
-
Using Large Language Models to Generate Engaging Captions for Data Visualizations 27 Dec 2022 · 0 repositories · arXiv:2212.14047
-
Biologically Inspired Design Concept Generation Using Generative Pre-Trained Transformers 26 Dec 2022 · 0 repositories · arXiv:2212.13196
-
Transformer and GAN Based Super-Resolution Reconstruction Network for Medical Images 26 Dec 2022 · 0 repositories · arXiv:2212.13068
-
TypeFormer: Transformers for Mobile Keystroke Biometrics 26 Dec 2022 · 1 repository · arXiv:2212.13075
-
Boosting Urban Traffic Speed Prediction via Integrating Implicit Spatial Correlations 25 Dec 2022 · 0 repositories · arXiv:2212.12932
-
Hybrid Representation Learning for Cognitive Diagnosis in Late-Life Depression Over 5 Years with Structural MRI 24 Dec 2022 · 1 repository · arXiv:2212.12810
-
On Realization of Intelligent Decision-Making in the Real World: A Foundation Decision Model Perspective 24 Dec 2022 · 1 repository · arXiv:2212.12669
-
Optimizing Deep Transformers for Chinese-Thai Low-Resource Translation 24 Dec 2022 · 0 repositories · arXiv:2212.12662
-
A Close Look at Spatial Modeling: From Attention to Convolution 23 Dec 2022 · 1 repository · arXiv:2212.12552
-
AMDET: Attention based Multiple Dimensions EEG Transformer for Emotion Recognition 23 Dec 2022 · 0 repositories · arXiv:2212.12134
-
Benchmark for Uncertainty & Robustness in Self-Supervised Learning 23 Dec 2022 · 1 repository · arXiv:2212.12411
-
Detecting Objects with Context-Likelihood Graphs and Graph Refinement 23 Dec 2022 · 0 repositories · arXiv:2212.12395
-
Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times? 23 Dec 2022 · 0 repositories · arXiv:2212.12131
-
Text Generation with Diffusion Language Models: A Pre-training Approach with Continuous Paragraph Denoise 22 Dec 2022 · 1 repository · arXiv:2212.11685Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
When are Lemons Purple? The Concept Association Bias of Vision-Language Models 22 Dec 2022 · 0 repositories · arXiv:2212.12043
-
Analyzing Semantic Faithfulness of Language Models via Input Intervention on Question Answering 21 Dec 2022 · 1 repository · arXiv:2212.10696
-
Automatic Emotion Modelling in Written Stories 21 Dec 2022 · 1 repository · arXiv:2212.11382
-
Beyond Contrastive Learning: A Variational Generative Model for Multilingual Retrieval 21 Dec 2022 · 1 repository · arXiv:2212.10726
-
DuAT: Dual-Aggregation Transformer Network for Medical Image Segmentation 21 Dec 2022 · 1 repository · arXiv:2212.11677
-
Entropy- and Distance-Based Predictors From GPT-2 Attention Patterns Predict Reading Times Over and Above GPT-2 Surprisal 21 Dec 2022 · 1 repository · arXiv:2212.11185Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 8 pointer-only (licence)
-
Investigation of Network Architecture for Multimodal Head-and-Neck Tumor Segmentation 21 Dec 2022 · 0 repositories · arXiv:2212.10724
-
JASMINE: Arabic GPT Models for Few-Shot Learning 21 Dec 2022 · 0 repositories · arXiv:2212.10755
-
KL Regularized Normalization Framework for Low Resource Tasks 21 Dec 2022 · 0 repositories · arXiv:2212.11275
-
SLGTformer: An Attention-Based Approach to Sign Language Recognition 21 Dec 2022 · 1 repository · arXiv:2212.10746
-
Spoken Language Understanding for Conversational AI: Recent Advances and Future Direction 21 Dec 2022 · 0 repositories · arXiv:2212.10728
-
Text classification in shipping industry using unsupervised models and Transformer based supervised models 21 Dec 2022 · 0 repositories · arXiv:2212.12407
-
Uncontrolled Lexical Exposure Leads to Overestimation of Compositional Generalization in Pretrained Models 21 Dec 2022 · 1 repository · arXiv:2212.10769
-
A Length-Extrapolatable Transformer 20 Dec 2022 · 5 repositories · arXiv:2212.10554Syntology official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 2 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language Models 20 Dec 2022 · 1 repository · arXiv:2212.10474Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Controllable Text Generation with Language Constraints 20 Dec 2022 · 0 repositories · arXiv:2212.10466
-
Diffusion Glancing Transformer for Parallel Sequence to Sequence Learning 20 Dec 2022 · 0 repositories · arXiv:2212.10240
-
Do language models have coherent mental models of everyday things? 20 Dec 2022 · 1 repository · arXiv:2212.10029Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
DocAsRef: An Empirical Study on Repurposing Reference-Based Summary Quality Metrics Reference-Freely 20 Dec 2022 · 1 repository · arXiv:2212.10013
-
EIT: Enhanced Interactive Transformer 20 Dec 2022 · 2 repositories · arXiv:2212.10197
-
Future Sight: Dynamic Story Generation with Large Pretrained Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.09947
-
Generic Temporal Reasoning with Differential Analysis and Explanation 20 Dec 2022 · 0 repositories · arXiv:2212.10467
-
Go-tuning: Improving Zero-shot Learning Abilities of Smaller Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10461
-
Is GPT-3 a Good Data Annotator? 20 Dec 2022 · 1 repository · arXiv:2212.10450
-
Evaluating Psychological Safety of Large Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10529
-
KronA: Parameter Efficient Tuning with Kronecker Adapter 20 Dec 2022 · 0 repositories · arXiv:2212.10650
-
Large Language Models Are Reasoning Teachers 20 Dec 2022 · 1 repository · arXiv:2212.10071Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
METEOR Guided Divergence for Video Captioning 20 Dec 2022 · 1 repository · arXiv:2212.10690
-
PairReranker: Pairwise Reranking for Natural Language Generation 20 Dec 2022 · 0 repositories · arXiv:2212.10555
-
Pay Attention to Your Tone: Introducing a New Dataset for Polite Language Rewrite 20 Dec 2022 · 1 repository · arXiv:2212.10190
-
Dissecting Transformer Length Extrapolation via the Lens of Receptive Field Analysis 20 Dec 2022 · 0 repositories · arXiv:2212.10356
-
T-Projection: High Quality Annotation Projection for Sequence Labeling Tasks 20 Dec 2022 · 2 repositories · arXiv:2212.10548
-
True Detective: A Deep Abductive Reasoning Benchmark Undoable for GPT-3 and Challenging for GPT-4 20 Dec 2022 · 0 repositories · arXiv:2212.10114
-
Why Can GPT Learn In-Context? Language Models Implicitly Perform Gradient Descent as Meta-Optimizers 20 Dec 2022 · 1 repository · arXiv:2212.10559
-
Empowering Diffusion Models on the Embedding Space for Text Generation 19 Dec 2022 · 1 repository · arXiv:2212.09412Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 3 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
Do CoNLL-2003 Named Entity Taggers Still Work Well in 2023? 19 Dec 2022 · 1 repository · arXiv:2212.09747
-
Emergent Analogical Reasoning in Large Language Models 19 Dec 2022 · 2 repositories · arXiv:2212.09196Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Evaluating Human-Language Model Interaction 19 Dec 2022 · 1 repository · arXiv:2212.09746
-
Large Language Models are Better Reasoners with Self-Verification 19 Dec 2022 · 1 repository · arXiv:2212.09561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
LENS: A Learnable Evaluation Metric for Text Simplification 19 Dec 2022 · 1 repository · arXiv:2212.09739
-
MIGA: A Unified Multi-task Generation Framework for Conversational Text-to-SQL 19 Dec 2022 · 0 repositories · arXiv:2212.09278
-
MIST: Multi-modal Iterative Spatial-Temporal Transformer for Long-form Video Question Answering 19 Dec 2022 · 1 repository · arXiv:2212.09522
-
Mu²SLAM: Multitask, Multilingual Speech and Language Models 19 Dec 2022 · 0 repositories · arXiv:2212.09553
-
Multilingual Sequence-to-Sequence Models for Hebrew NLP 19 Dec 2022 · 0 repositories · arXiv:2212.09682
-
Reasoning with Language Model Prompting: A Survey 19 Dec 2022 · 2 repositories · arXiv:2212.09597
-
SrTR: Self-reasoning Transformer with Visual-linguistic Knowledge for Scene Graph Generation 19 Dec 2022 · 0 repositories · arXiv:2212.09329
-
The case for 4-bit precision: k-bit Inference Scaling Laws 19 Dec 2022 · 1 repository · arXiv:2212.09720
-
Tokenization Consistency Matters for Generative Models on Extractive NLP Tasks 19 Dec 2022 · 1 repository · arXiv:2212.09912
-
Can Retriever-Augmented Language Models Reason? The Blame Game Between the Retriever and the Language Model 18 Dec 2022 · 1 repository · arXiv:2212.09146
-
Style-Hallucinated Dual Consistency Learning: A Unified Framework for Visual Domain Generalization 18 Dec 2022 · 1 repository · arXiv:2212.09068
-
Claim Optimization in Computational Argumentation 17 Dec 2022 · 1 repository · arXiv:2212.08913
-
Leveraging Wastewater Monitoring for COVID-19 Forecasting in the US: a Deep Learning study 17 Dec 2022 · 1 repository · arXiv:2212.08798
-
Autoencoders as Cross-Modal Teachers: Can Pretrained 2D Image Transformers Help 3D Representation Learning? 16 Dec 2022 · 4 repositories · arXiv:2212.08320
-
Convolution-enhanced Evolving Attention Networks 16 Dec 2022 · 1 repository · arXiv:2212.08330
-
Homonymy Information for English WordNet 16 Dec 2022 · 1 repository · arXiv:2212.08388
-
LOANet: A Lightweight Network Using Object Attention for Extracting Buildings and Roads from UAV Aerial Remote Sensing Images 16 Dec 2022 · 1 repository · arXiv:2212.08490
-
LegalRelectra: Mixed-domain Language Modeling for Long-range Legal Text Comprehension 16 Dec 2022 · 0 repositories · arXiv:2212.08204
-
MURMUR: Modular Multi-Step Reasoning for Semi-Structured Data-to-Text Generation 16 Dec 2022 · 0 repositories · arXiv:2212.08607
-
Assessing the Impact of Sequence Length Learning on Classification Tasks for Transformer Encoder Models 16 Dec 2022 · 0 repositories · arXiv:2212.08399
-
Rethinking Cooking State Recognition with Vision Transformers 16 Dec 2022 · 1 repository · arXiv:2212.08586
-
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA 16 Dec 2022 · 1 repository · arXiv:2212.08635