Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 135
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 135 of 190: papers 13,401 to 13,500 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Eeny, meeny, miny, moe. How to choose data for morphological inflection 26 Oct 2022 · 1 repository · arXiv:2210.14465
-
Exploring Robustness of Prefix Tuning in Noisy Data: A Case Study in Financial Sentiment Analysis 26 Oct 2022 · 0 repositories · arXiv:2211.05584
-
Learning a Task-specific Descriptor for Robust Matching of 3D Point Clouds 26 Oct 2022 · 0 repositories · arXiv:2210.14899
-
Leveraging Affirmative Interpretations from Negation Improves Natural Language Understanding 26 Oct 2022 · 1 repository · arXiv:2210.14486
-
Leveraging Demonstrations with Latent Space Priors 26 Oct 2022 · 1 repository · arXiv:2210.14685Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Multilevel Transformer For Multimodal Emotion Recognition 26 Oct 2022 · 0 repositories · arXiv:2211.07711
-
Pretrained audio neural networks for Speech emotion recognition in Portuguese 26 Oct 2022 · 1 repository · arXiv:2210.14716
-
SemFormer: Semantic Guided Activation Transformer for Weakly Supervised Semantic Segmentation 26 Oct 2022 · 1 repository · arXiv:2210.14618
-
Xiaoicesing 2: A High-Fidelity Singing Voice Synthesizer Based on Generative Adversarial Network 26 Oct 2022 · 1 repository · arXiv:2210.14666
-
Audio MFCC-gram Transformers for respiratory insufficiency detection in COVID-19 25 Oct 2022 · 1 repository · arXiv:2210.14085
-
Dynamic Survival Transformers for Causal Inference with Electronic Health Records 25 Oct 2022 · 0 repositories · arXiv:2210.15417
-
End-to-end Transformer for Compressed Video Quality Enhancement 25 Oct 2022 · 0 repositories · arXiv:2210.13827
-
Explicitly Increasing Input Information Density for Vision Transformers on Small Datasets 25 Oct 2022 · 1 repository · arXiv:2210.14319Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
IELM: An Open Information Extraction Benchmark for Pre-Trained Language Models 25 Oct 2022 · 0 repositories · arXiv:2210.14128
-
KnowGL: Knowledge Generation and Linking from Text 25 Oct 2022 · 0 repositories · arXiv:2210.13952
-
MEW-UNet: Multi-axis representation learning in frequency domain for medical image segmentation 25 Oct 2022 · 1 repository · arXiv:2210.14007
-
Minutiae-Guided Fingerprint Embeddings via Vision Transformers 25 Oct 2022 · 0 repositories · arXiv:2210.13994
-
MOFormer: Self-Supervised Transformer model for Metal-Organic Framework Property Prediction 25 Oct 2022 · 1 repository · arXiv:2210.14188
-
THOR-Net: End-to-end Graformer-based Realistic Two Hands and Object Reconstruction with Self-supervision 25 Oct 2022 · 1 repository · arXiv:2210.13853
-
XRICL: Cross-lingual Retrieval-Augmented In-Context Learning for Cross-lingual Text-to-SQL Semantic Parsing 25 Oct 2022 · 0 repositories · arXiv:2210.13693
-
Inferring Past Human Actions in Homes with Abductive Reasoning 24 Oct 2022 · 1 repository · arXiv:2210.13984
-
Effective Pre-Training Objectives for Transformer-based Autoencoders 24 Oct 2022 · 0 repositories · arXiv:2210.13536
-
ELMER: A Non-Autoregressive Pre-trained Language Model for Efficient and Effective Text Generation 24 Oct 2022 · 1 repository · arXiv:2210.13304
-
Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task 24 Oct 2022 · 4 repositories · arXiv:2210.13382Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Exploring Euphemism Detection in Few-Shot and Zero-Shot Settings 24 Oct 2022 · 1 repository · arXiv:2210.12926
-
Foreground Guidance and Multi-Layer Feature Fusion for Unsupervised Object Discovery with Transformers 24 Oct 2022 · 1 repository · arXiv:2210.13053
-
High Fidelity Neural Audio Compression 24 Oct 2022 · 6 repositories · arXiv:2210.13438
-
Perfectly Secure Steganography Using Minimum Entropy Coupling 24 Oct 2022 · 2 repositories · arXiv:2210.14889Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
Sequential Recommendation with Auxiliary Item Relationships via Multi-Relational Transformer 24 Oct 2022 · 1 repository · arXiv:2210.13572
-
Video based Object 6D Pose Estimation using Transformers 24 Oct 2022 · 1 repository · arXiv:2210.13540
-
VLC-BERT: Visual Question Answering with Contextualized Commonsense Knowledge 24 Oct 2022 · 1 repository · arXiv:2210.13626Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Anticipative Feature Fusion Transformer for Multi-Modal Action Anticipation 23 Oct 2022 · 1 repository · arXiv:2210.12649
-
Delving into Masked Autoencoders for Multi-Label Thorax Disease Classification 23 Oct 2022 · 1 repository · arXiv:2210.12843Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Holistic Interaction Transformer Network for Action Detection 23 Oct 2022 · 1 repository · arXiv:2210.12686
-
Exploring the Value of Pre-trained Language Models for Clinical Named Entity Recognition 23 Oct 2022 · 2 repositories · arXiv:2210.12770
-
UIA-ViT: Unsupervised Inconsistency-Aware Method based on Vision Transformer for Face Forgery Detection 23 Oct 2022 · 0 repositories · arXiv:2210.12752
-
A Comprehensive Comparison of Neural Networks as Cognitive Models of Inflection 22 Oct 2022 · 0 repositories · arXiv:2210.12321
-
Learning Point-Language Hierarchical Alignment for 3D Visual Grounding 22 Oct 2022 · 1 repository · arXiv:2210.12513Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Leveraging Large Language Models for Multiple Choice Question Answering 22 Oct 2022 · 1 repository · arXiv:2210.12353Syntology official (archive's flag): 6 ran · 6 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Meta-learning Pathologies from Radiology Reports using Variance Aware Prototypical Networks 22 Oct 2022 · 0 repositories · arXiv:2210.13979
-
MS-DCANet: A Novel Segmentation Network For Multi-Modality COVID-19 Medical Images 22 Oct 2022 · 0 repositories · arXiv:2210.12361
-
Recurrence Boosts Diversity! Revisiting Recurrent Latent Variable in Transformer-Based Variational AutoEncoder for Diverse Text Generation 22 Oct 2022 · 0 repositories · arXiv:2210.12409
-
S2WAT: Image Style Transfer via Hierarchical Vision Transformer using Strips Window Attention 22 Oct 2022 · 1 repository · arXiv:2210.12381
-
Speech Emotion Recognition via an Attentive Time-Frequency Neural Network 22 Oct 2022 · 0 repositories · arXiv:2210.12430
-
SynGEC: Syntax-Enhanced Grammatical Error Correction with a Tailored GEC-Oriented Parser 22 Oct 2022 · 1 repository · arXiv:2210.12484
-
Transformer-Based Conditioned Variational Autoencoder for Dialogue Generation 22 Oct 2022 · 0 repositories · arXiv:2210.12326
-
A Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language Models 21 Oct 2022 · 1 repository · arXiv:2210.12023Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Amos: An Adam-style Optimizer with Adaptive Weight Decay towards Model-Oriented Scale 21 Oct 2022 · 1 repository · arXiv:2210.11693Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples)
-
Context-Enhanced Stereo Transformer 21 Oct 2022 · 1 repository · arXiv:2210.11719Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Decoding a Neural Retriever's Latent Space for Query Suggestion 21 Oct 2022 · 1 repository · arXiv:2210.12084
-
Diffuser: Efficient Transformers with Multi-hop Attention Diffusion for Long Sequences 21 Oct 2022 · 1 repository · arXiv:2210.11794
-
Do Vision-and-Language Transformers Learn Grounded Predicate-Noun Dependencies? 21 Oct 2022 · 1 repository · arXiv:2210.12079
-
Face Pyramid Vision Transformer 21 Oct 2022 · 1 repository · arXiv:2210.11974
-
Is Encoder-Decoder Redundant for Neural Machine Translation? 21 Oct 2022 · 0 repositories · arXiv:2210.11807
-
Shift-Reduce Task-Oriented Semantic Parsing with Stack-Transformers 21 Oct 2022 · 1 repository · arXiv:2210.11984
-
SLING: Sino Linguistic Evaluation of Large Language Models 21 Oct 2022 · 1 repository · arXiv:2210.11689Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
Syntax-guided Localized Self-attention by Constituency Syntactic Distance 21 Oct 2022 · 1 repository · arXiv:2210.11759Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
TransLIST: A Transformer-Based Linguistically Informed Sanskrit Tokenizer 21 Oct 2022 · 1 repository · arXiv:2210.11753
-
University of Cape Town's WMT22 System: Multilingual Machine Translation for Southern African Languages 21 Oct 2022 · 0 repositories · arXiv:2210.11757
-
WikiWhy: Answering and Explaining Cause-and-Effect Questions 21 Oct 2022 · 0 repositories · arXiv:2210.12152
-
3DALL-E: Integrating Text-to-Image AI in 3D Design Workflows 20 Oct 2022 · 0 repositories · arXiv:2210.11603
-
Composing Ensembles of Pre-trained Models via Iterative Consensus 20 Oct 2022 · 0 repositories · arXiv:2210.11522
-
General Image Descriptors for Open World Image Retrieval using ViT CLIP 20 Oct 2022 · 1 repository · arXiv:2210.11141Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Identifying Human Strategies for Generating Word-Level Adversarial Examples 20 Oct 2022 · 0 repositories · arXiv:2210.11598
-
Scaling Instruction-Finetuned Language Models 20 Oct 2022 · 9 repositories · arXiv:2210.11416Syntology 8 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 7 where Syntology's instrument failed) · 9 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
Self-Supervised Learning with Masked Image Modeling for Teeth Numbering, Detection of Dental Restorations, and Instance Segmentation in Dental Panoramic Radiographs 20 Oct 2022 · 1 repository · arXiv:2210.11404
-
SimpleClick: Interactive Image Segmentation with Simple Vision Transformers 20 Oct 2022 · 2 repositories · arXiv:2210.11006Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Single Image Super-Resolution Using Lightweight Networks Based on Swin Transformer 20 Oct 2022 · 0 repositories · arXiv:2210.11019
-
Solving Reasoning Tasks with a Slot Transformer 20 Oct 2022 · 0 repositories · arXiv:2210.11394
-
SSiT: Saliency-guided Self-supervised Image Transformer for Diabetic Retinopathy Grading 20 Oct 2022 · 1 repository · arXiv:2210.10969
-
A Unified View of Masked Image Modeling 19 Oct 2022 · 1 repository · arXiv:2210.10615
-
BioGPT: Generative Pre-trained Transformer for Biomedical Text Generation and Mining 19 Oct 2022 · 4 repositories · arXiv:2210.10341
-
Grounded Video Situation Recognition 19 Oct 2022 · 0 repositories · arXiv:2210.10828
-
Language Detoxification with Attribute-Discriminative Latent Space 19 Oct 2022 · 1 repository · arXiv:2210.10329
-
Multi-view Gait Recognition based on Siamese Vision Transformer 19 Oct 2022 · 0 repositories · arXiv:2210.10421
-
Museformer: Transformer with Fine- and Coarse-Grained Attention for Music Generation 19 Oct 2022 · 1 repository · arXiv:2210.10349Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
PoseGPT: Quantization-based 3D Human Motion Generation and Forecasting 19 Oct 2022 · 1 repository · arXiv:2210.10542
-
Revision Transformers: Instructing Language Models to Change their Values 19 Oct 2022 · 1 repository · arXiv:2210.10332
-
Self-supervised Graph Masking Pre-training for Graph-to-Text Generation 19 Oct 2022 · 1 repository · arXiv:2210.10599
-
Towards a neural architecture of language: Deep learning versus logistics of access in neural architectures for compositional processing 19 Oct 2022 · 0 repositories · arXiv:2210.10543
-
Transformers Learn Shortcuts to Automata 19 Oct 2022 · 0 repositories · arXiv:2210.10749
-
A Hybrid System of Sound Event Detection Transformer and Frame-wise Model for DCASE 2022 Task 4 18 Oct 2022 · 1 repository · arXiv:2210.09529
-
Cross-Domain Aspect Extraction using Transformers Augmented with Knowledge Graphs 18 Oct 2022 · 1 repository · arXiv:2210.10144
-
CTGAN : Cloud Transformer Generative Adversarial Network 18 Oct 2022 · 1 repository
-
From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data 18 Oct 2022 · 0 repositories · arXiv:2210.10047
-
Multimodal Image Fusion based on Hybrid CNN-Transformer and Non-local Cross-modal Attention 18 Oct 2022 · 1 repository · arXiv:2210.09847
-
Swinv2-Imagen: Hierarchical Vision Transformer Diffusion Models for Text-to-Image Generation 18 Oct 2022 · 0 repositories · arXiv:2210.09549
-
Systematicity in GPT-3's Interpretation of Novel English Noun Compounds 18 Oct 2022 · 0 repositories · arXiv:2210.09492
-
Team Flow at DRC2022: Pipeline System for Travel Destination Recommendation Task in Spoken Dialogue 18 Oct 2022 · 0 repositories · arXiv:2210.09518
-
Tiny-Attention Adapter: Contexts Are More Important Than the Number of Parameters 18 Oct 2022 · 0 repositories · arXiv:2211.01979
-
Transfer-learning for video classification: Video Swin Transformer on multiple domains 18 Oct 2022 · 0 repositories · arXiv:2210.09969
-
ViTCoD: Vision Transformer Acceleration via Dedicated Algorithm and Accelerator Co-Design 18 Oct 2022 · 1 repository · arXiv:2210.09573
-
A Generative User Simulator with GPT-based Architecture and Goal State Tracking for Reinforced Multi-Domain Dialog Systems 17 Oct 2022 · 1 repository · arXiv:2210.08692Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
A Mixing Time Lower Bound for a Simplified Version of BART 17 Oct 2022 · 0 repositories · arXiv:2210.09352
-
Deep Bidirectional Language-Knowledge Graph Pretraining 17 Oct 2022 · 2 repositories · arXiv:2210.09338Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 18 harvested samples) · 1 pointer-only (licence)
-
Histopathological Image Classification based on Self-Supervised Vision Transformer and Weak Labels 17 Oct 2022 · 1 repository · arXiv:2210.09021
-
Prompting GPT-3 To Be Reliable 17 Oct 2022 · 1 repository · arXiv:2210.09150
-
SGRAM: Improving Scene Graph Parsing via Abstract Meaning Representation 17 Oct 2022 · 0 repositories · arXiv:2210.08675
-
Zero-Shot Ranking Socio-Political Texts with Transformer Language Models to Reduce Close Reading Time 17 Oct 2022 · 0 repositories · arXiv:2210.09179
-
Accelerating Transfer Learning with Near-Data Computation on Cloud Object Stores 16 Oct 2022 · 1 repository · arXiv:2210.08650