Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 157
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 157 of 190: papers 15,601 to 15,700 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
End-to-end Task-oriented Dialog Policy Learning based on Pre-trained Language Model 16 Nov 2021 · 0 repositories
-
Enhancing the Nonlinear Mutual Dependencies in Transformers with Mutual Information 16 Nov 2021 · 0 repositories
-
ERNIE-SPARSE: Learning Hierarchical Efficient Transformer Through Regularized Self-Attention 16 Nov 2021 · 0 repositories
-
Exploring and Adapting Chinese GPT to Pinyin Input Method 16 Nov 2021 · 0 repositories
-
Extreme Multi-label Text Classification with Multi-layer Experts 16 Nov 2021 · 0 repositories
-
Fast and Accurate Transformer-based Translation with Character-Level Encoding and Subword-Level Decoding 16 Nov 2021 · 0 repositories
-
Fix Bugs with Transformer through a Neural-Symbolic Edit Grammar 16 Nov 2021 · 0 repositories
-
FRSUM: Towards Faithful Abstractive Summarization via Enhancing Factual Robustness 16 Nov 2021 · 0 repositories
-
Generating Diverse and High-Quality Abstractive Summaries with Variational Transformers 16 Nov 2021 · 0 repositories
-
Generation of News Articles from Tweets : An Experiment 16 Nov 2021 · 0 repositories
-
Generative Pre-Trained Transformer for Design Concept Generation: An Exploration 16 Nov 2021 · 0 repositories · arXiv:2111.08489
-
GLM: General Language Model Pretraining with Autoregressive Blank Infilling 16 Nov 2021 · 0 repositories
-
Impact of Tokenization on Language Models: An Analysis for Turkish 16 Nov 2021 · 0 repositories
-
Improving Compositional Generalization with Self-Training for Data-to-Text Generation 16 Nov 2021 · 0 repositories
-
Improving GPT-3 after deployment with a dynamic memory of feedback 16 Nov 2021 · 0 repositories
-
Input-specific Attention Subnetworks for Adversarial Detection 16 Nov 2021 · 0 repositories
-
Interpreting the Robustness of Neural NLP Models to Textual Perturbations 16 Nov 2021 · 0 repositories
-
Knowledge Graph is in Rescue: Task Oriented Dialogue System for Response Generation without NLU and DM 16 Nov 2021 · 0 repositories
-
Knowledge-guided Transformer for Joint Theme and Emotion Classification of Chinese Classical Poetry 16 Nov 2021 · 0 repositories
-
Learn More from Less: Improving Conversational Recommender Systems via Contextual and Time-Aware Modeling 16 Nov 2021 · 0 repositories
-
Learning Methods for Solving Astronomy Course Problems 16 Nov 2021 · 0 repositories
-
Learning Non-Autoregressive Models from Search for Unsupervised Sentence Summarization 16 Nov 2021 · 0 repositories
-
Life after BERT: What do Other Muppets Understand about Language? 16 Nov 2021 · 0 repositories
-
Listen to Both Sides and be Enlightened! -- Hierarchical Modality Fusion Network for Entity and Relation Extraction 16 Nov 2021 · 0 repositories
-
Making Transformers Solve Compositional Tasks 16 Nov 2021 · 0 repositories
-
MarCQAp: Effective Context Modeling for Conversational Question Answering 16 Nov 2021 · 0 repositories
-
Meeting Summarization with Pre-training and Clustering Methods 16 Nov 2021 · 1 repository · arXiv:2111.08210
-
Modeling Hierarchical Syntax Structure with Triplet Position for Source Code Summarization 16 Nov 2021 · 1 repository
-
Moving the Eiffel Tower to ROME: Tracing and Editing Facts in GPT 16 Nov 2021 · 0 repositories
-
Multi-head or Single-head? An Empirical Comparison for Transformer Training 16 Nov 2021 · 0 repositories
-
N-grammer: Augmenting Transformers with latent n-grams 16 Nov 2021 · 0 repositories
-
Neural Keyphrase Generation: Analysis and Evaluation 16 Nov 2021 · 0 repositories
-
On the Multilingual Capabilities of Very Large-Scale English Language Models 16 Nov 2021 · 0 repositories
-
On Vision Features in Multimodal Machine Translation 16 Nov 2021 · 0 repositories
-
Online Advertising Revenue Forecasting: An Interpretable Deep Learning Approach 16 Nov 2021 · 0 repositories · arXiv:2111.08840
-
PESTO: A Post-User Fusion Network for Rumour Detection on Social Media 16 Nov 2021 · 0 repositories
-
PoliSe: Reinforcing Politeness using User Sentiment for Customer Care Response Generation 16 Nov 2021 · 0 repositories
-
Reasoning Like Program Executors 16 Nov 2021 · 0 repositories
-
Representation of Ambiguity in Pre-Trained Sentence Embeddings 16 Nov 2021 · 0 repositories
-
Retrieval-based Layer-wise Adaptive Transformer for Source Code Summarization 16 Nov 2021 · 0 repositories
-
Sequence-to-sequence AMR Parsing with Ancestor Information 16 Nov 2021 · 0 repositories
-
Sequence-to-Sequence Knowledge Graph Completion and Question Answering 16 Nov 2021 · 0 repositories
-
SHCT: A Successively Hierarchical Conditional Transformer for Controllable Paraphrase Generation 16 Nov 2021 · 0 repositories
-
ShrinkNAS : Single-Path One-Shot Operator Exploratory Training for Transformer with Dynamic Space Shrinking 16 Nov 2021 · 0 repositories
-
Softmax Bottleneck Makes Language Models Unable to Represent Multi-mode Word Distributions 16 Nov 2021 · 0 repositories
-
Solving Probability and Statistics Problems by Program Synthesis 16 Nov 2021 · 0 repositories · arXiv:2111.08267
-
Solving Probability and Statistics Problems by Program Synthesis 16 Nov 2021 · 0 repositories
-
Sparsifying Transformer Models with Trainable Representation Pooling 16 Nov 2021 · 0 repositories
-
Speaker Profiling in Multi-party Conversations 16 Nov 2021 · 0 repositories
-
Structured Pruning Learns Compact and Accurate Models 16 Nov 2021 · 0 repositories
-
SuperShaper: Task-Agnostic Super Pre-training of BERT Models with Variable Hidden Dimensions 16 Nov 2021 · 0 repositories
-
TACO: Pre-training of Deep Transformers with Attention Convolution using Disentangled Positional Representation 16 Nov 2021 · 0 repositories
-
Teaching BERT to Wait: Balancing Accuracy and Latency for Streaming Disfluency Detection 16 Nov 2021 · 0 repositories
-
Tell me who you are and i'll tell you what to do: A Persona Grounded Task Oriented Dialogue Generation System 16 Nov 2021 · 0 repositories
-
The Power of Prompt Tuning for Low-Resource Semantic Parsing 16 Nov 2021 · 0 repositories
-
Towards Coding Social Science Datasets with Language Models 16 Nov 2021 · 0 repositories
-
WeTS: A Benchmark for Translation Suggestion 16 Nov 2021 · 0 repositories
-
What Works and Doesn't Work, A Deep Decoder for Neural Machine Translation 16 Nov 2021 · 0 repositories
-
When classifying grammatical role, BERT doesn't care about word order... except when it matters 16 Nov 2021 · 0 repositories
-
AI in Human-computer Gaming: Techniques, Challenges and Opportunities 15 Nov 2021 · 0 repositories · arXiv:2111.07631
-
AUTOMATED AUDIO CAPTIONING BY FINE-TUNING BART WITH AUDIOSET TAGS 15 Nov 2021 · 1 repository
-
Calculating Question Similarity is Enough: A New Method for KBQA Tasks 15 Nov 2021 · 0 repositories · arXiv:2111.07658
-
Exploring Story Generation with Multi-task Objectives in Variational Autoencoders 15 Nov 2021 · 0 repositories · arXiv:2111.08133
-
FakeTransformer: Exposing Face Forgery From Spatial-Temporal Representation Modeled By Facial Pixel Variations 15 Nov 2021 · 0 repositories · arXiv:2111.07601
-
Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction 15 Nov 2021 · 4 repositories · arXiv:2111.07910Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Say What? Collaborative Pop Lyric Generation Using Multitask Transfer Learning 15 Nov 2021 · 0 repositories · arXiv:2111.07592
-
Scaling Law for Recommendation Models: Towards General-purpose User Representations 15 Nov 2021 · 0 repositories · arXiv:2111.11294
-
Local Multi-Head Channel Self-Attention for Facial Expression Recognition 14 Nov 2021 · 1 repository · arXiv:2111.07224
-
Transformer-based Image Compression 12 Nov 2021 · 0 repositories · arXiv:2111.06707
-
A Survey of Visual Transformers 11 Nov 2021 · 1 repository · arXiv:2111.06091
-
Automated question generation and question answering from Turkish texts 11 Nov 2021 · 1 repository · arXiv:2111.06476
-
Graph Relation Transformer: Incorporating pairwise object features into the Transformer architecture 11 Nov 2021 · 0 repositories · arXiv:2111.06075
-
Improving Large-scale Language Models and Resources for Filipino 11 Nov 2021 · 0 repositories · arXiv:2111.06053
-
A Novel Corpus of Discourse Structure in Humans and Computers 10 Nov 2021 · 1 repository · arXiv:2111.05940
-
Amazon SageMaker Model Parallelism: A General and Flexible Framework for Large Model Training 10 Nov 2021 · 0 repositories · arXiv:2111.05972
-
Attention Approximates Sparse Distributed Memory 10 Nov 2021 · 1 repository · arXiv:2111.05498Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 19 harvested samples) · 7 pointer-only (licence)
-
Multimodal Transformer with Variable-length Memory for Vision-and-Language Navigation 10 Nov 2021 · 1 repository · arXiv:2111.05759
-
Prune Once for All: Sparse Pre-Trained Language Models 10 Nov 2021 · 2 repositories · arXiv:2111.05754
-
Soft Sensing Transformer: Hundreds of Sensors are Worth a Single Word 10 Nov 2021 · 1 repository · arXiv:2111.05973
-
DistIR: An Intermediate Representation and Simulator for Efficient Neural Network Distribution 9 Nov 2021 · 0 repositories · arXiv:2111.05426
-
FPM: A Collection of Large-scale Foundation Pre-trained Language Models 9 Nov 2021 · 0 repositories · arXiv:2111.04909
-
Multi-Task Prediction of Clinical Outcomes in the Intensive Care Unit using Flexible Multimodal Transformers 9 Nov 2021 · 0 repositories · arXiv:2111.05431
-
Sliced Recursive Transformer 9 Nov 2021 · 1 repository · arXiv:2111.05297
-
Guiding Multi-Step Rearrangement Tasks with Natural Language Instructions 8 Nov 2021 · 2 repositories
-
Mixed Transformer U-Net For Medical Image Segmentation 8 Nov 2021 · 1 repository · arXiv:2111.04734Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Are we ready for a new paradigm shift? A Survey on Visual Deep MLP 7 Nov 2021 · 1 repository · arXiv:2111.04060
-
Analyzing Architectures for Neural Machine Translation Using Low Computational Resources 6 Nov 2021 · 0 repositories · arXiv:2111.03813
-
Benchmarking Data-driven Surrogate Simulators for Artificial Electromagnetic Materials 6 Nov 2021 · 1 repository
-
Convolutional Gated MLP: Combining Convolutions & gMLP 6 Nov 2021 · 0 repositories · arXiv:2111.03940
-
IBERT: Idiom Cloze-style reading comprehension with Attention 5 Nov 2021 · 0 repositories · arXiv:2112.02994
-
Improving Visual Quality of Image Synthesis by A Token-based Generator with Transformers 5 Nov 2021 · 0 repositories · arXiv:2111.03481
-
Oracle Teacher: Leveraging Target Information for Better Knowledge Distillation of CTC Models 5 Nov 2021 · 0 repositories · arXiv:2111.03664
-
An Empirical Study of the Effectiveness of an Ensemble of Stand-alone Sentiment Detection Tools for Software Engineering Datasets 4 Nov 2021 · 1 repository · arXiv:2111.03196
-
Benchmarking Multimodal AutoML for Tabular Data with Text Fields 4 Nov 2021 · 2 repositories · arXiv:2111.02705Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
MT3: Multi-Task Multitrack Music Transcription 4 Nov 2021 · 3 repositories · arXiv:2111.03017Syntology official: harvested, nothing ran · 0 ran · 6 unverified (of 6 harvested samples)
-
Multi-Airport Delay Prediction with Transformers 4 Nov 2021 · 0 repositories · arXiv:2111.04494
-
An Explanation of In-context Learning as Implicit Bayesian Inference 3 Nov 2021 · 1 repository · arXiv:2111.02080Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
ProSTformer: Pre-trained Progressive Space-Time Self-attention Model for Traffic Flow Forecasting 3 Nov 2021 · 0 repositories · arXiv:2111.03459
-
TheEyeCorpus: Experiments in Reducing NLP Bias and Identifiability for Large LMs 3 Nov 2021 · 0 repositories
-
VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts 3 Nov 2021 · 2 repositories · arXiv:2111.02358