Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 119
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 119 of 190: papers 11,801 to 11,900 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Multi-Path Transformer is Better: A Case Study on Neural Machine Translation 10 May 2023 · 0 repositories · arXiv:2305.05948
-
A Black-Box Attack on Code Models via Representation Nearest Neighbor Search 10 May 2023 · 0 repositories · arXiv:2305.05896
-
Summarizing, Simplifying, and Synthesizing Medical Evidence Using GPT-3 (with Varying Success) 10 May 2023 · 1 repository · arXiv:2305.06299
-
VTPNet for 3D deep learning on point cloud 10 May 2023 · 0 repositories · arXiv:2305.06115
-
An Exploration of Encoder-Decoder Approaches to Multi-Label Classification for Legal and Biomedical Text 9 May 2023 · 1 repository · arXiv:2305.05627
-
Tomography of Quantum States from Structured Measurements via quantum-aware transformer 9 May 2023 · 0 repositories · arXiv:2305.05433
-
AudioSlots: A slot-centric generative model for audio separation 9 May 2023 · 0 repositories · arXiv:2305.05591
-
CodeIE: Large Code Generation Models are Better Few-Shot Information Extractors 9 May 2023 · 1 repository · arXiv:2305.05711Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance 9 May 2023 · 1 repository · arXiv:2305.05176Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
GPT in Game Theory Experiments 9 May 2023 · 0 repositories · arXiv:2305.05516
-
GPT-NAS: Evolutionary Neural Architecture Search with the Generative Pre-Trained Model 9 May 2023 · 0 repositories · arXiv:2305.05351
-
Hybrid Transformer and CNN Attention Network for Stereo Image Super-resolution 9 May 2023 · 0 repositories · arXiv:2305.05177
-
InternGPT: Solving Vision-Centric Tasks by Interacting with ChatGPT Beyond Language 9 May 2023 · 2 repositories · arXiv:2305.05662Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Effects of sub-word segmentation on performance of transformer language models 9 May 2023 · 0 repositories · arXiv:2305.05480
-
Simplicial Hopfield networks 9 May 2023 · 1 repository · arXiv:2305.05179
-
Towards an Automatic Optimisation Model Generator Assisted with Generative Pre-trained Transformer 9 May 2023 · 0 repositories · arXiv:2305.05811
-
Towards Building the Federated GPT: Federated Instruction Tuning 9 May 2023 · 1 repository · arXiv:2305.05644Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Vision-Language Models in Remote Sensing: Current Progress and Future Trends 9 May 2023 · 3 repositories · arXiv:2305.05726
-
Code Execution with Pre-trained Language Models 8 May 2023 · 1 repository · arXiv:2305.05383
-
Coherent Wave Dynamics and Language Generation of a Generative Pre-trained Transformer 8 May 2023 · 0 repositories · arXiv:2305.05061
-
Do Large Language Models Show Decision Heuristics Similar to Humans? A Case Study Using GPT-3.5 8 May 2023 · 0 repositories · arXiv:2305.04400
-
Explanation-based Finetuning Makes Models More Robust to Spurious Cues 8 May 2023 · 1 repository · arXiv:2305.04990
-
Fast Conformer with Linearly Scalable Attention for Efficient Speech Recognition 8 May 2023 · 0 repositories · arXiv:2305.05084
-
GersteinLab at MEDIQA-Chat 2023: Clinical Note Summarization from Doctor-Patient Conversations through Fine-tuning and In-context Learning 8 May 2023 · 0 repositories · arXiv:2305.05001
-
Graph Masked Autoencoder for Sequential Recommendation 8 May 2023 · 2 repositories · arXiv:2305.04619
-
Multi-Task End-to-End Training Improves Conversational Recommendation 8 May 2023 · 0 repositories · arXiv:2305.06218
-
NeuroComparatives: Neuro-Symbolic Distillation of Comparative Knowledge 8 May 2023 · 1 repository · arXiv:2305.04978
-
Real-World Denoising via Diffusion Model 8 May 2023 · 0 repositories · arXiv:2305.04457
-
Revisiting Relation Extraction in the era of Large Language Models 8 May 2023 · 0 repositories · arXiv:2305.05003
-
Robust Traffic Light Detection Using Salience-Sensitive Loss: Computational Framework and Evaluations 8 May 2023 · 0 repositories · arXiv:2305.04516
-
Smart Home Device Detection Algorithm Based on FSA-YOLOv5 8 May 2023 · 0 repositories · arXiv:2305.04534
-
Unlocking Practical Applications in Legal Domain: Evaluation of GPT for Zero-Shot Semantic Annotation of Legal Texts 8 May 2023 · 0 repositories · arXiv:2305.04417
-
AdaptiveClick: Clicks-aware Transformer with Adaptive Focal Loss for Interactive Image Segmentation 7 May 2023 · 1 repository · arXiv:2305.04276Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
CIT-EmotionNet: CNN Interactive Transformer Network for EEG Emotion Recognition 7 May 2023 · 0 repositories · arXiv:2305.05548
-
Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting 7 May 2023 · 2 repositories · arXiv:2305.04388Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Model-Contrastive Federated Domain Adaptation 7 May 2023 · 0 repositories · arXiv:2305.10432
-
Poses as Queries: Image-to-LiDAR Map Localization with Transformers 7 May 2023 · 0 repositories · arXiv:2305.04298
-
Professional Certification Benchmark Dataset: The First 500 Jobs For Large Language Models 7 May 2023 · 0 repositories · arXiv:2305.05377
-
RFR-WWANet: Weighted Window Attention-Based Recovery Feature Resolution Network for Unsupervised Image Registration 7 May 2023 · 1 repository · arXiv:2305.04236
-
X-LLM: Bootstrapping Advanced Large Language Models by Treating Multi-Modalities as Foreign Languages 7 May 2023 · 2 repositories · arXiv:2305.04160Syntology 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Unlocking the Power of GANs in Non-Autoregressive Text Generation 6 May 2023 · 0 repositories · arXiv:2305.03977
-
Artificial Neuropsychology: Are Large Language Models Developing Executive Functions? 6 May 2023 · 0 repositories · arXiv:2305.04134
-
Cognition Guided Human-Object Relationship Detection 6 May 2023 · 0 repositories
-
DBAT: Dynamic Backward Attention Transformer for Material Segmentation with Cross-Resolution Patches 6 May 2023 · 1 repository · arXiv:2305.03919
-
Degradation-Noise-Aware Deep Unfolding Transformer for Hyperspectral Image Denoising 6 May 2023 · 0 repositories · arXiv:2305.04047
-
Diffusion-NAT: Self-Prompting Discrete Diffusion for Non-Autoregressive Text Generation 6 May 2023 · 0 repositories · arXiv:2305.04044
-
Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models 6 May 2023 · 3 repositories · arXiv:2305.04091Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Refining the Responses of LLMs by Themselves 6 May 2023 · 1 repository · arXiv:2305.04039
-
From Zero to Hero: Harnessing Transformers for Biomedical Named Entity Recognition in Zero- and Few-shot Contexts 5 May 2023 · 1 repository · arXiv:2305.04928
-
Adapting Transformer Language Models for Predictive Typing in Brain-Computer Interfaces 5 May 2023 · 0 repositories · arXiv:2305.03819
-
CLaC at SemEval-2023 Task 2: Comparing Span-Prediction and Sequence-Labeling approaches for NER 5 May 2023 · 0 repositories · arXiv:2305.03845
-
FM-ViT: Flexible Modal Vision Transformers for Face Anti-Spoofing 5 May 2023 · 0 repositories · arXiv:2305.03277
-
LMEye: An Interactive Perception Network for Large Language Models 5 May 2023 · 1 repository · arXiv:2305.03701
-
LOGO-Former: Local-Global Spatio-Temporal Transformer for Dynamic Facial Expression Recognition 5 May 2023 · 0 repositories · arXiv:2305.03343
-
MindGames: Targeting Theory of Mind in Large Language Models with Dynamic Epistemic Modal Logic 5 May 2023 · 2 repositories · arXiv:2305.03353
-
Neuromodulation Gated Transformer 5 May 2023 · 1 repository · arXiv:2305.03232Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Online Gesture Recognition using Transformer and Natural Language Processing 5 May 2023 · 0 repositories · arXiv:2305.03407
-
Otter: A Multi-Modal Model with In-Context Instruction Tuning 5 May 2023 · 1 repository · arXiv:2305.03726
-
Predicting COVID-19 and pneumonia complications from admission texts 5 May 2023 · 0 repositories · arXiv:2305.03661
-
Retrieval Augmented Chest X-Ray Report Generation using OpenAI GPT models 5 May 2023 · 0 repositories · arXiv:2305.03660
-
MAF-Net: Multiple attention-guided fusion network for fundus vascular image segmentation 5 May 2023 · 0 repositories · arXiv:2305.03617
-
Simulating H.P. Lovecraft horror literature with the ChatGPT large language model 5 May 2023 · 0 repositories · arXiv:2305.03429
-
Transformer Working Memory Enables Regular Language Reasoning and Natural Language Length Extrapolation 5 May 2023 · 0 repositories · arXiv:2305.03796
-
Using ChatGPT for Entity Matching 5 May 2023 · 1 repository · arXiv:2305.03423
-
Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework 5 May 2023 · 1 repository · arXiv:2305.03268
-
Masked Structural Growth for 2x Faster Language Model Pre-training 4 May 2023 · 1 repository · arXiv:2305.02869Syntology official (archive's flag): 16 ran · 16 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 18 unverified (of 34 harvested samples)
-
An automatically discovered chain-of-thought prompt generalizes to novel models and datasets 4 May 2023 · 0 repositories · arXiv:2305.02897
-
AutoML-GPT: Automatic Machine Learning with GPT 4 May 2023 · 0 repositories · arXiv:2305.02499
-
BranchNorm: Robustly Scaling Extremely Deep Transformers 4 May 2023 · 0 repositories · arXiv:2305.02790
-
Catch Missing Details: Image Reconstruction with Frequency Augmented Variational Autoencoder 4 May 2023 · 1 repository · arXiv:2305.02541Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 3 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples) · 3 pointer-only (licence)
-
Chain-of-Skills: A Configurable Model for Open-domain Question Answering 4 May 2023 · 1 repository · arXiv:2305.03130
-
Gpt-4: A Review on Advancements and Opportunities in Natural Language Processing 4 May 2023 · 0 repositories · arXiv:2305.03195
-
Hierarchical Transformer for Scalable Graph Learning 4 May 2023 · 0 repositories · arXiv:2305.02866
-
Interpretable Sentence Representation with Variational Autoencoders and Attention 4 May 2023 · 0 repositories · arXiv:2305.02810
-
Can LLMs Capture Human Preferences? 4 May 2023 · 0 repositories · arXiv:2305.02531
-
Late-Binding Scholarship in the Age of AI: Navigating Legal and Normative Challenges of a New Form of Knowledge Production 4 May 2023 · 0 repositories · arXiv:2305.11058
-
Learning Language-Specific Layers for Multilingual Machine Translation 4 May 2023 · 0 repositories · arXiv:2305.02665
-
Noise-Resistant Multimodal Transformer for Emotion Recognition 4 May 2023 · 0 repositories · arXiv:2305.02814
-
PersonaLLM: Investigating the Ability of Large Language Models to Express Personality Traits 4 May 2023 · 1 repository · arXiv:2305.02547Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Point Transformer For Coronary Artery Labeling 4 May 2023 · 0 repositories · arXiv:2305.02533
-
Self-Supervised Learning for Organs At Risk and Tumor Segmentation with Uncertainty Quantification 4 May 2023 · 0 repositories · arXiv:2305.02491
-
The Application of Affective Measures in Text-based Emotion Aware Recommender Systems 4 May 2023 · 0 repositories · arXiv:2305.04796
-
UPDExplainer: an Interpretable Transformer-based Framework for Urban Physical Disorder Detection Using Street View Imagery 4 May 2023 · 0 repositories · arXiv:2305.02911
-
What changes when you randomly choose BPE merge operations? Not much 4 May 2023 · 0 repositories · arXiv:2305.03029
-
A Lightweight CNN-Transformer Model for Learning Traveling Salesman Problems 3 May 2023 · 1 repository · arXiv:2305.01883
-
A Systematic Study of Knowledge Distillation for Natural Language Generation with Pseudo-Target Training 3 May 2023 · 1 repository · arXiv:2305.02031Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Alleviating Exposure Bias via Multi-level Contrastive Learning and Deviation Simulation in Abstractive Summarization 3 May 2023 · 1 repository
-
Cheaply Evaluating Inference Efficiency Metrics for Autoregressive Transformer APIs 3 May 2023 · 0 repositories · arXiv:2305.02440
-
WangLab at MEDIQA-Chat 2023: Clinical Note Generation from Doctor-Patient Conversations using Large Language Models 3 May 2023 · 0 repositories · arXiv:2305.02220
-
Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model Sizes 3 May 2023 · 1 repository · arXiv:2305.02301Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Entity Tracking in Language Models 3 May 2023 · 1 repository · arXiv:2305.02363
-
Glitch in the Matrix: A Large Scale Benchmark for Content Driven Audio-Visual Forgery Detection and Localization 3 May 2023 · 1 repository · arXiv:2305.01979
-
GPT-RE: In-context Learning for Relation Extraction using Large Language Models 3 May 2023 · 1 repository · arXiv:2305.02105
-
Learngene: Inheriting Condensed Knowledge from the Ancestry Model to Descendant Models 3 May 2023 · 0 repositories · arXiv:2305.02279
-
Towards Imperceptible Document Manipulations against Neural Ranking Models 3 May 2023 · 0 repositories · arXiv:2305.01860
-
A Study on the Integration of Pipeline and E2E SLU systems for Spoken Semantic Parsing toward STOP Quality Challenge 2 May 2023 · 0 repositories · arXiv:2305.01620
-
ARBEx: Attentive Feature Extraction with Reliability Balancing for Robust Facial Expression Learning 2 May 2023 · 1 repository · arXiv:2305.01486
-
AxWin Transformer: A Context-Aware Vision Transformer Backbone with Axial Windows 2 May 2023 · 0 repositories · arXiv:2305.01280
-
BrainNPT: Pre-training of Transformer networks for brain network classification 2 May 2023 · 0 repositories · arXiv:2305.01666
-
Why So Gullible? Enhancing the Robustness of Retrieval-Augmented Models against Counterfactual Noise 2 May 2023 · 1 repository · arXiv:2305.01579