Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 173
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 173 of 190: papers 17,201 to 17,300 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Code Summarization with Structure-induced Transformer 29 Dec 2020 · 1 repository · arXiv:2012.14710
-
Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models 28 Dec 2020 · 2 repositories · arXiv:2012.14252
-
Red Dragon AI at TextGraphs 2020 Shared Task: LIT : LSTM-Interleaved Transformer for Multi-Hop Explanation Ranking 28 Dec 2020 · 1 repository · arXiv:2012.14164
-
Syntax-Enhanced Pre-trained Model 28 Dec 2020 · 1 repository · arXiv:2012.14116
-
TransPose: Keypoint Localization via Transformer 28 Dec 2020 · 1 repository · arXiv:2012.14214
-
Learning Light-Weight Translation Models from Deep Transformer 27 Dec 2020 · 1 repository · arXiv:2012.13866Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Portfolio Optimization with 2D Relative-Attentional Gated Transformer 27 Dec 2020 · 0 repositories · arXiv:2101.03138
-
SG-Net: Syntax Guided Transformer for Language Representation 27 Dec 2020 · 0 repositories · arXiv:2012.13915
-
Detecting Hateful Memes Using a Multimodal Deep Ensemble 24 Dec 2020 · 1 repository · arXiv:2012.13235
-
I like fish, especially dolphins: Addressing Contradictions in Dialogue Modeling 24 Dec 2020 · 0 repositories · arXiv:2012.13391
-
Future-Guided Incremental Transformer for Simultaneous Translation 23 Dec 2020 · 0 repositories · arXiv:2012.12465
-
Domain Adaptation of NMT models for English-Hindi Machine Translation Task at AdapMT ICON 2020 22 Dec 2020 · 0 repositories · arXiv:2012.12112
-
Molecular CT: Unifying Geometry and Representation Learning for Molecules at Different Scales 22 Dec 2020 · 0 repositories · arXiv:2012.11816
-
Multi-Head Self-Attention with Role-Guided Masks 22 Dec 2020 · 1 repository · arXiv:2012.12366
-
Uncertainty and Surprisal Jointly Deliver the Punchline: Exploiting Incongruity-Based Features for Humor Recognition 22 Dec 2020 · 0 repositories · arXiv:2012.12007
-
3D Object Detection with Pointformer 21 Dec 2020 · 1 repository · arXiv:2012.11409
-
Encoding Syntactic Knowledge in Transformer Encoder for Intent Detection and Slot Filling 21 Dec 2020 · 0 repositories · arXiv:2012.11689
-
RealFormer: Transformer Likes Residual Attention 21 Dec 2020 · 5 repositories · arXiv:2012.11747
-
Leveraging ParsBERT and Pretrained mT5 for Persian Abstractive Text Summarization 21 Dec 2020 · 1 repository · arXiv:2012.11204
-
Sub-Linear Memory: How to Make Performers SLiM 21 Dec 2020 · 2 repositories · arXiv:2012.11346
-
Adaptive Bi-directional Attention: Exploring Multi-Granularity Representations for Machine Reading Comprehension 20 Dec 2020 · 0 repositories · arXiv:2012.10877
-
Breaking Writer's Block: Low-cost Fine-tuning of Natural Language Generation Models 19 Dec 2020 · 0 repositories · arXiv:2101.03216
-
NeurST: Neural Speech Translation Toolkit 18 Dec 2020 · 1 repository · arXiv:2012.10018
-
Toward Streaming ASR with Non-Autoregressive Insertion-based Model 18 Dec 2020 · 0 repositories · arXiv:2012.10128
-
A Generalization of Transformer Networks to Graphs 17 Dec 2020 · 3 repositories · arXiv:2012.09699Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
End-to-end Deep Object Tracking with Circular Loss Function for Rotated Bounding Box 17 Dec 2020 · 0 repositories · arXiv:2012.09771
-
PCT: Point cloud transformer 17 Dec 2020 · 11 repositories · arXiv:2012.09688Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Taming Transformers for High-Resolution Image Synthesis 17 Dec 2020 · 13 repositories · arXiv:2012.09841Syntology official (archive's flag): 2 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 4 pointer-only (licence)
-
Toward Transformer-Based Object Detection 17 Dec 2020 · 0 repositories · arXiv:2012.09958
-
Transformer Interpretability Beyond Attention Visualization 17 Dec 2020 · 3 repositories · arXiv:2012.09838Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
DialogXL: All-in-One XLNet for Multi-Party Conversation Emotion Recognition 16 Dec 2020 · 4 repositories · arXiv:2012.08695
-
Learning from Mistakes: Using Mis-predictions as Harm Alerts in Language Pre-Training 16 Dec 2020 · 0 repositories · arXiv:2012.08789
-
Point Transformer 16 Dec 2020 · 24 repositories · arXiv:2012.09164
-
Query expansion with artificially generated texts 16 Dec 2020 · 0 repositories · arXiv:2012.08787
-
Revisiting Linformer with a modified self-attention with linear complexity 16 Dec 2020 · 0 repositories · arXiv:2101.10277
-
High throughput screening with machine learning 15 Dec 2020 · 0 repositories · arXiv:2012.08275
-
RecipeNLG: A Cooking Recipes Dataset for Semi-Structured Text Generation 15 Dec 2020 · 1 repository
-
Traditional IR rivals neural models on the MS MARCO Document Ranking Leaderboard 15 Dec 2020 · 2 repositories · arXiv:2012.08020
-
Contrastive Learning with Adversarial Perturbations for Conditional Text Generation 14 Dec 2020 · 1 repository · arXiv:2012.07280Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Extracting Training Data from Large Language Models 14 Dec 2020 · 3 repositories · arXiv:2012.07805Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Informer: Beyond Efficient Transformer for Long Sequence Time-Series Forecasting 14 Dec 2020 · 14 repositories · arXiv:2012.07436Syntology official: harvested, nothing ran · 62 ran (of which 53 constructed an object rather than computing a result; 62 with no instrument failure: 0 honoured, 1 violated, 61 with no contract checked; 0 where Syntology's instrument failed) · 14 unverified (of 76 harvested samples) · 13 pointer-only (licence)
-
Reasoning in Dialog: Improving Response Generation by Context Reading Comprehension 14 Dec 2020 · 1 repository · arXiv:2012.07410
-
Discriminative Pre-training for Low Resource Title Compression in Conversational Grocery 13 Dec 2020 · 0 repositories · arXiv:2012.06943
-
Improving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network 13 Dec 2020 · 1 repository · arXiv:2012.07061
-
KVL-BERT: Knowledge Enhanced Visual-and-Linguistic BERT for Visual Commonsense Reasoning 13 Dec 2020 · 0 repositories · arXiv:2012.07000
-
DETR for Crowd Pedestrian Detection 12 Dec 2020 · 1 repository · arXiv:2012.06785
-
Yelp Review Rating Prediction: Machine Learning and Deep Learning Models 12 Dec 2020 · 1 repository · arXiv:2012.06690
-
Hardware Beyond Backpropagation: a Photonic Co-Processor for Direct Feedback Alignment 11 Dec 2020 · 0 repositories · arXiv:2012.06373
-
Spatial Temporal Transformer Network for Skeleton-based Action Recognition 11 Dec 2020 · 1 repository · arXiv:2012.06399
-
TabTransformer: Tabular Data Modeling Using Contextual Embeddings 11 Dec 2020 · 12 repositories · arXiv:2012.06678Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
As Good as New. How to Successfully Recycle English GPT-2 to Make Models for Other Languages 10 Dec 2020 · 1 repository · arXiv:2012.05628
-
GDA-HIN: A Generalized Domain Adaptive Model across Heterogeneous Information Networks 10 Dec 2020 · 0 repositories · arXiv:2012.05688
-
Towards Neural Programming Interfaces 10 Dec 2020 · 1 repository · arXiv:2012.05983
-
Simple is not Easy: A Simple Strong Baseline for TextVQA and TextCaps 9 Dec 2020 · 1 repository · arXiv:2012.05153
-
Extractive Opinion Summarization in Quantized Transformer Spaces 8 Dec 2020 · 2 repositories · arXiv:2012.04443
-
Parameter Efficient Multimodal Transformers for Video Representation Learning 8 Dec 2020 · 0 repositories · arXiv:2012.04124
-
CX DB8: A queryable extractive summarizer and semantic search engine 7 Dec 2020 · 2 repositories · arXiv:2012.03942
-
Deep Policy Networks for NPC Behaviors that Adapt to Changing Design Parameters in Roguelike Games 7 Dec 2020 · 0 repositories · arXiv:2012.03532
-
Document Graph for Neural Machine Translation 7 Dec 2020 · 0 repositories · arXiv:2012.03477
-
UBAR: Towards Fully End-to-End Task-Oriented Dialog Systems with GPT-2 7 Dec 2020 · 1 repository · arXiv:2012.03539
-
[Re] Satellite Image Time Series Classification with Pixel-Set Encoders and Temporal Self-Attention 6 Dec 2020 · 1 repository
-
Data-Efficient Methods for Dialogue Systems 5 Dec 2020 · 0 repositories · arXiv:2012.02929
-
Enhanced Offensive Language Detection Through Data Augmentation 5 Dec 2020 · 0 repositories · arXiv:2012.02954
-
Pre-training Protein Language Models with Label-Agnostic Binding Pairs Enhances Performance in Downstream Tasks 5 Dec 2020 · 1 repository · arXiv:2012.03084
-
CUED_speech at TREC 2020 Podcast Summarisation Track 4 Dec 2020 · 0 repositories · arXiv:2012.02535
-
EchoBERT: A Transformer-Based Approach for Behavior Detection in Echograms 4 Dec 2020 · 1 repository
-
Fine-tuning BERT for Low-Resource Natural Language Understanding via Active Learning 4 Dec 2020 · 0 repositories · arXiv:2012.02462
-
RPT: Relational Pre-trained Transformer Is Almost All You Need towards Democratizing Data Preparation 4 Dec 2020 · 0 repositories · arXiv:2012.02469
-
DialogBERT: Discourse-Aware Response Generation via Learning to Recover and Rank Utterances 3 Dec 2020 · 1 repository · arXiv:2012.01775Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
TRACE: Early Detection of Chronic Kidney Disease Onset with Transformer-Enhanced Feature Embedding 3 Dec 2020 · 0 repositories · arXiv:2012.03729
-
Contour Transformer Network for One-shot Segmentation of Anatomical Structures 2 Dec 2020 · 1 repository · arXiv:2012.01480
-
How Can We Know When Language Models Know? On the Calibration of Language Models for Question Answering 2 Dec 2020 · 1 repository · arXiv:2012.00955
-
Two-Stage Single Image Reflection Removal with Reflection-Aware Guidance 2 Dec 2020 · 1 repository · arXiv:2012.00945
-
A Deep Generative Approach to Native Language Identification 1 Dec 2020 · 0 repositories
-
Adversarial Sparse Transformer for Time Series Forecasting 1 Dec 2020 · 1 repository
-
Analogy Models for Neural Word Inflection 1 Dec 2020 · 1 repository
-
BERT at SemEval-2020 Task 8: Using BERT to Analyse Meme Emotions 1 Dec 2020 · 0 repositories
-
Bilingual Subword Segmentation for Neural Machine Translation 1 Dec 2020 · 0 repositories
-
Comparing Probabilistic, Distributional and Transformer-Based Models on Logical Metonymy Interpretation 1 Dec 2020 · 0 repositories
-
CPM: A Large-scale Generative Chinese Pre-trained Language Model 1 Dec 2020 · 10 repositories · arXiv:2012.00413
-
Denoising Pre-Training and Data Augmentation Strategies for Enhanced RDF Verbalization with Transformers 1 Dec 2020 · 0 repositories · arXiv:2012.00571
-
Diverse dialogue generation with context dependent dynamic loss function 1 Dec 2020 · 0 repositories
-
Domain Transfer based Data Augmentation for Neural Query Translation 1 Dec 2020 · 0 repositories
-
Evaluating Unsupervised Representation Learning for Detecting Stances of Fake News 1 Dec 2020 · 0 repositories
-
Ferryman at SemEval-2020 Task 12: BERT-Based Model with Advanced Improvement Methods for Multilingual Offensive Language Identification 1 Dec 2020 · 0 repositories
-
Ferryman at SemEval-2020 Task 7: Ensemble Model for Assessing Humor in Edited News Headlines 1 Dec 2020 · 0 repositories
-
Flight of the PEGASUS? Comparing Transformers on Few-shot and Zero-shot Multi-document Abstractive Summarization 1 Dec 2020 · 1 repository
-
Formality Style Transfer with Shared Latent Space 1 Dec 2020 · 1 repository
-
Generalized Shortest-Paths Encoders for AMR-to-Text Generation 1 Dec 2020 · 0 repositories
-
Hitachi at SemEval-2020 Task 11: An Empirical Study of Pre-Trained Transformer Family for Propaganda Detection 1 Dec 2020 · 0 repositories
-
Hitachi at SemEval-2020 Task 7: Stacking at Scale with Heterogeneous Language Models for Humor Recognition 1 Dec 2020 · 0 repositories
-
Hitachi at SemEval-2020 Task 8: Simple but Effective Modality Ensemble for Meme Emotion Recognition 1 Dec 2020 · 0 repositories
-
How Far Does BERT Look At: Distance-based Clustering and Analysis of BERT's Attention 1 Dec 2020 · 0 repositories
-
Hy-NLI: a Hybrid system for Natural Language Inference 1 Dec 2020 · 2 repositories
-
I2C at SemEval-2020 Task 12: Simple but Effective Approaches to Offensive Speech Detection in Twitter 1 Dec 2020 · 0 repositories
-
Image Caption Generation for News Articles 1 Dec 2020 · 1 repository
-
Incorporating Noisy Length Constraints into Transformer with Length-aware Positional Encodings 1 Dec 2020 · 0 repositories
-
Increasing Learning Efficiency of Self-Attention Networks through Direct Position Interactions, Learnable Temperature, and Convoluted Attention 1 Dec 2020 · 1 repository
-
Incremental Neural Lexical Coherence Modeling 1 Dec 2020 · 1 repository
-
Investigating Gender Bias in Language Models Using Causal Mediation Analysis 1 Dec 2020 · 0 repositories