Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 131
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 131 of 190: papers 13,001 to 13,100 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Teaching Small Language Models to Reason 16 Dec 2022 · 0 repositories · arXiv:2212.08410
-
Efficient Long Sequence Modeling via State Space Augmented Transformer 15 Dec 2022 · 1 repository · arXiv:2212.08136Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
FiDO: Fusion-in-Decoder optimized for stronger performance and faster inference 15 Dec 2022 · 0 repositories · arXiv:2212.08153
-
HUM3DIL: Semi-supervised Multi-modal 3D Human Pose Estimation for Autonomous Driving 15 Dec 2022 · 0 repositories · arXiv:2212.07729
-
CLIPPO: Image-and-Language Understanding from Pixels Only 15 Dec 2022 · 1 repository · arXiv:2212.08045
-
Revisiting the Gold Standard: Grounding Summarization Evaluation with Robust Human Evaluation 15 Dec 2022 · 2 repositories · arXiv:2212.07981Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Scaffold-Based Multi-Objective Drug Candidate Optimization 15 Dec 2022 · 0 repositories · arXiv:2301.07175
-
Sim-to-Real Transfer for Quadrupedal Locomotion via Terrain Transformer 15 Dec 2022 · 0 repositories · arXiv:2212.07740
-
Visually-augmented pretrained language models for NLP tasks without images 15 Dec 2022 · 1 repository · arXiv:2212.07937
-
Most Important Person-guided Dual-branch Cross-Patch Attention for Group Affect Recognition 14 Dec 2022 · 0 repositories · arXiv:2212.07055
-
Evaluating Byte and Wordpiece Level Models for Massively Multilingual Semantic Parsing 14 Dec 2022 · 0 repositories · arXiv:2212.07223
-
VTCC-NLP at NL4Opt competition subtask 1: An Ensemble Pre-trained language models for Named Entity Recognition 14 Dec 2022 · 0 repositories · arXiv:2212.07219
-
CREPE: Can Vision-Language Foundation Models Reason Compositionally? 13 Dec 2022 · 1 repository · arXiv:2212.07796Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Generative artificial intelligence-enabled dynamic detection of nicotine-related circuits 13 Dec 2022 · 0 repositories · arXiv:2212.06330
-
GPViT: A High Resolution Non-Hierarchical Vision Transformer with Group Propagation 13 Dec 2022 · 2 repositories · arXiv:2212.06795
-
Paraphrase Identification with Deep Learning: A Review of Datasets and Methods 13 Dec 2022 · 0 repositories · arXiv:2212.06933
-
RT-1: Robotics Transformer for Real-World Control at Scale 13 Dec 2022 · 1 repository · arXiv:2212.06817Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Smart Journey in Istanbul: A Mobile Application in Smart Cities for Traffic Estimation by Harnessing Time Series 13 Dec 2022 · 0 repositories · arXiv:2212.09448
-
Automated ICD Coding using Extreme Multi-label Long Text Transformer-based Models 12 Dec 2022 · 1 repository · arXiv:2212.05857
-
BeautyREC: Robust, Efficient, and Content-preserving Makeup Transfer 12 Dec 2022 · 0 repositories · arXiv:2212.05855
-
CTT-Net: A Multi-view Cross-token Transformer for Cataract Postoperative Visual Acuity Prediction 12 Dec 2022 · 1 repository · arXiv:2212.05794
-
Deep learning approaches to building rooftop thermal bridge detection from aerial images 12 Dec 2022 · 1 repository
-
NMS Strikes Back 12 Dec 2022 · 1 repository · arXiv:2212.06137
-
P-Transformer: Towards Better Document-to-Document Neural Machine Translation 12 Dec 2022 · 1 repository · arXiv:2212.05830
-
ROIFormer: Semantic-Aware Region of Interest Transformer for Efficient Self-Supervised Monocular Depth Estimation 12 Dec 2022 · 0 repositories · arXiv:2212.05729
-
T5Score: Discriminative Fine-tuning of Generative Evaluation Metrics 12 Dec 2022 · 2 repositories · arXiv:2212.05726
-
Video Prediction by Efficient Transformers 12 Dec 2022 · 1 repository · arXiv:2212.06026Syntology official: harvested, nothing ran · 0 ran · 7 unverified (of 7 harvested samples)
-
Elixir: Train a Large Language Model on a Small GPU Cluster 10 Dec 2022 · 2 repositories · arXiv:2212.05339
-
Joint Spatio-Temporal Modeling for the Semantic Change Detection in Remote Sensing Images 10 Dec 2022 · 3 repositories · arXiv:2212.05245Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Thinking Fast and Slow in Large Language Models 10 Dec 2022 · 0 repositories · arXiv:2212.05206
-
MAGVIT: Masked Generative Video Transformer 10 Dec 2022 · 1 repository · arXiv:2212.05199Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Position Embedding Needs an Independent Layer Normalization 10 Dec 2022 · 1 repository · arXiv:2212.05262
-
Punctuation Restoration for Singaporean Spoken Languages: English, Malay, and Mandarin 10 Dec 2022 · 1 repository · arXiv:2212.05356
-
SMILE: Scaling Mixture-of-Experts with Efficient Bi-level Routing 10 Dec 2022 · 0 repositories · arXiv:2212.05191
-
Structured information extraction from complex scientific text with fine-tuned large language models 10 Dec 2022 · 0 repositories · arXiv:2212.05238
-
Dynamic Test-Time Augmentation via Differentiable Functions 9 Dec 2022 · 1 repository · arXiv:2212.04681
-
Masked Lip-Sync Prediction by Audio-Visual Contextual Exploitation in Transformers 9 Dec 2022 · 0 repositories · arXiv:2212.04970
-
MIMO Is All You Need : A Strong Multi-In-Multi-Out Baseline for Video Prediction 9 Dec 2022 · 1 repository · arXiv:2212.04655
-
Cross-Domain Synthetic-to-Real In-the-Wild Depth and Normal Estimation for 3D Scene Understanding 9 Dec 2022 · 0 repositories · arXiv:2212.05040
-
RCDT: Relational Remote Sensing Change Detection with Transformer 9 Dec 2022 · 1 repository · arXiv:2212.04869
-
Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints 9 Dec 2022 · 1 repository · arXiv:2212.05055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Turing Deception 9 Dec 2022 · 0 repositories · arXiv:2212.06721
-
TRBLLmaker -- Transformer Reads Between Lyrics Lines maker 9 Dec 2022 · 0 repositories · arXiv:2212.04917
-
Explain to me like I am five -- Sentence Simplification Using Transformers 8 Dec 2022 · 1 repository · arXiv:2212.04595
-
Federated Learning for Inference at Anytime and Anywhere 8 Dec 2022 · 0 repositories · arXiv:2212.04084
-
Group Generalized Mean Pooling for Vision Transformer 8 Dec 2022 · 0 repositories · arXiv:2212.04114
-
Harnessing the Power of Multi-Task Pretraining for Ground-Truth Level Natural Language Explanations 8 Dec 2022 · 1 repository · arXiv:2212.04231Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models 8 Dec 2022 · 1 repository · arXiv:2212.04088
-
NP4G : Network Programming for Generalization 8 Dec 2022 · 1 repository · arXiv:2212.11118
-
NRTR: Neuron Reconstruction with Transformer from 3D Optical Microscopy Images 8 Dec 2022 · 0 repositories · arXiv:2212.04163
-
The Role of AI in Drug Discovery: Challenges, Opportunities, and Strategies 8 Dec 2022 · 0 repositories · arXiv:2212.08104
-
DeepSpeed Data Efficiency: Improving Deep Learning Model Quality and Training Efficiency via Efficient Data Sampling and Routing 7 Dec 2022 · 1 repository · arXiv:2212.03597
-
Gaussian Radar Transformer for Semantic Segmentation in Noisy Radar Data 7 Dec 2022 · 0 repositories · arXiv:2212.03690
-
Hierarchical multimodal transformers for Multi-Page DocVQA 7 Dec 2022 · 1 repository · arXiv:2212.05935
-
Learning-To-Embed: Adopting Transformer based models for E-commerce Products Representation Learning 7 Dec 2022 · 0 repositories · arXiv:2212.03725
-
Multimodal Vision Transformers with Forced Attention for Behavior Analysis 7 Dec 2022 · 0 repositories · arXiv:2212.03968
-
A K-variate Time Series Is Worth K Words: Evolution of the Vanilla Transformer Architecture for Long-term Multivariate Time Series Forecasting 6 Dec 2022 · 0 repositories · arXiv:2212.02789
-
AbHE: All Attention-based Homography Estimation 6 Dec 2022 · 0 repositories · arXiv:2212.03029
-
Adaptive Testing of Computer Vision Models 6 Dec 2022 · 1 repository · arXiv:2212.02774Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Controlled Text Generation using T5 based Encoder-Decoder Soft Prompt Tuning and Analysis of the Utility of Generated Text in AI 6 Dec 2022 · 0 repositories · arXiv:2212.02924
-
Counterfactual reasoning: Do language models need world knowledge for causal understanding? 6 Dec 2022 · 1 repository · arXiv:2212.03278
-
Document-Level Abstractive Summarization 6 Dec 2022 · 1 repository · arXiv:2212.03013
-
IncepFormer: Efficient Inception Transformer with Pyramid Pooling for Semantic Segmentation 6 Dec 2022 · 1 repository · arXiv:2212.03035
-
Modern French Poetry Generation with RoBERTa and GPT-2 6 Dec 2022 · 0 repositories · arXiv:2212.02911
-
Open World DETR: Transformer based Open World Object Detection 6 Dec 2022 · 0 repositories · arXiv:2212.02969
-
Pretrained Diffusion Models for Unified Human Motion Synthesis 6 Dec 2022 · 0 repositories · arXiv:2212.02837
-
Semantic-Conditional Diffusion Networks for Image Captioning 6 Dec 2022 · 2 repositories · arXiv:2212.03099
-
Simple Baseline for Weather Forecasting Using Spatiotemporal Context Aggregation Network 6 Dec 2022 · 1 repository · arXiv:2212.02952Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression 6 Dec 2022 · 2 repositories · arXiv:2212.02746Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples) · 4 pointer-only (licence)
-
Video Object of Interest Segmentation 6 Dec 2022 · 0 repositories · arXiv:2212.02871
-
3D-LatentMapper: View Agnostic Single-View Reconstruction of 3D Shapes 5 Dec 2022 · 0 repositories · arXiv:2212.02184
-
Audio-Driven Co-Speech Gesture Video Generation 5 Dec 2022 · 0 repositories · arXiv:2212.02350
-
Automatic Generation of Factual News Headlines in Finnish 5 Dec 2022 · 0 repositories · arXiv:2212.02170
-
FBLNet: FeedBack Loop Network for Driver Attention Prediction 5 Dec 2022 · 0 repositories · arXiv:2212.02096
-
Mask Matching Transformer for Few-Shot Segmentation 5 Dec 2022 · 1 repository · arXiv:2301.01208
-
Retrieval as Attention: End-to-end Learning of Retrieval and Reading within a Single Transformer 5 Dec 2022 · 1 repository · arXiv:2212.02027
-
Unifying Vision, Text, and Layout for Universal Document Processing 5 Dec 2022 · 5 repositories · arXiv:2212.02623Syntology official: no sample here; runs from other or unrecorded repositories · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 2 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Joint Self-Supervised Image-Volume Representation Learning with Intra-Inter Contrastive Clustering 4 Dec 2022 · 0 repositories · arXiv:2212.01893
-
Languages You Know Influence Those You Learn: Impact of Language Characteristics on Multi-Lingual Text-to-Text Transfer 4 Dec 2022 · 0 repositories · arXiv:2212.01757
-
Exploring the Limits of Differentially Private Deep Learning with Group-wise Clipping 3 Dec 2022 · 0 repositories · arXiv:2212.01539
-
Global memory transformer for processing long documents 3 Dec 2022 · 0 repositories · arXiv:2212.01650
-
Recognition and Prediction of Surgical Gestures and Trajectories Using Transformer Models in Robot-Assisted Surgery 3 Dec 2022 · 0 repositories · arXiv:2212.01683
-
FECAM: Frequency Enhanced Channel Attention Mechanism for Time Series Forecasting 2 Dec 2022 · 1 repository · arXiv:2212.01209
-
Relation-Aware Language-Graph Transformer for Question Answering 2 Dec 2022 · 1 repository · arXiv:2212.00975
-
Multi-scale Transformer Network with Edge-aware Pre-training for Cross-Modality MR Image Synthesis 2 Dec 2022 · 2 repositories · arXiv:2212.01108
-
SumREN: Summarizing Reported Speech about Events in News 2 Dec 2022 · 1 repository · arXiv:2212.01146
-
Tackling Low-Resourced Sign Language Translation: UPC at WMT-SLT 22 2 Dec 2022 · 1 repository · arXiv:2212.01140
-
Towards Diverse, Relevant and Coherent Open-Domain Dialogue Generation via Hybrid Latent Variables 2 Dec 2022 · 0 repositories · arXiv:2212.01145
-
a survey on GPT-3 1 Dec 2022 · 0 repositories · arXiv:2212.00857
-
CHAPTER: Exploiting Convolutional Neural Network Adapters for Self-supervised Speech Models 1 Dec 2022 · 0 repositories · arXiv:2212.01282
-
Concealed Object Detection for Passive Millimeter-Wave Security Imaging Based on Task-Aligned Detection Transformer 1 Dec 2022 · 0 repositories · arXiv:2212.00313
-
CUNI Non-Autoregressive System for the WMT 22 Efficient Translation Shared Task 1 Dec 2022 · 0 repositories · arXiv:2212.00477
-
Data-Efficient Finetuning Using Cross-Task Nearest Neighbors 1 Dec 2022 · 1 repository · arXiv:2212.00196Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Distilling Reasoning Capabilities into Smaller Language Models 1 Dec 2022 · 1 repository · arXiv:2212.00193
-
Explainable Artificial Intelligence for Improved Modeling of Processes 1 Dec 2022 · 1 repository · arXiv:2212.00695
-
Ghost-free High Dynamic Range Imaging via Hybrid CNN-Transformer and Structure Tensor 1 Dec 2022 · 1 repository · arXiv:2212.00595
-
Learning Progressive Modality-shared Transformers for Effective Visible-Infrared Person Re-identification 1 Dec 2022 · 1 repository · arXiv:2212.00226
-
DSNet: a simple yet efficient network with dual-stream attention for lesion segmentation 30 Nov 2022 · 0 repositories · arXiv:2211.16950
-
Part-based Face Recognition with Vision Transformers 30 Nov 2022 · 1 repository · arXiv:2212.00057Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Pattern Attention Transformer with Doughnut Kernel 30 Nov 2022 · 0 repositories · arXiv:2211.16961