Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 134
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 134 of 190: papers 13,301 to 13,400 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CoNMix for Source-free Single and Multi-target Domain Adaptation 7 Nov 2022 · 1 repository · arXiv:2211.03876Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining 7 Nov 2022 · 0 repositories · arXiv:2211.03594
-
How Much Does Attention Actually Attend? Questioning the Importance of Attention in Pretrained Transformers 7 Nov 2022 · 1 repository · arXiv:2211.03495
-
Sequential Transformer for End-to-End Person Search 6 Nov 2022 · 0 repositories · arXiv:2211.04323
-
Wall Street Tree Search: Risk-Aware Planning for Offline Reinforcement Learning 6 Nov 2022 · 0 repositories · arXiv:2211.04583
-
Inductive Graph Transformer for Delivery Time Estimation 5 Nov 2022 · 1 repository · arXiv:2211.02863
-
Learning to Infer from Unlabeled Data: A Semi-supervised Learning Approach for Robust Natural Language Inference 5 Nov 2022 · 1 repository · arXiv:2211.02971
-
A Transformer Architecture for Online Gesture Recognition of Mathematical Expressions 4 Nov 2022 · 0 repositories · arXiv:2211.02643
-
A Weakly-Supervised Streaming Multilingual Speech Model with Truly Zero-Shot Capability 4 Nov 2022 · 0 repositories · arXiv:2211.02499
-
CCATMos: Convolutional Context-aware Transformer Network for Non-intrusive Speech Quality Assessment 4 Nov 2022 · 0 repositories · arXiv:2211.02577
-
Deep learning for structural health monitoring: An application to heritage structures 4 Nov 2022 · 0 repositories · arXiv:2211.10351
-
Fraudulent User Detection Via Behavior Information Aggregation Network (BIAN) On Large-Scale Financial Social Network 4 Nov 2022 · 0 repositories · arXiv:2211.06315
-
Generation of Chinese classical poetry based on pre-trained model 4 Nov 2022 · 0 repositories · arXiv:2211.02541
-
OSIC: A New One-Stage Image Captioner Coined 4 Nov 2022 · 0 repositories · arXiv:2211.02321
-
Patch DCT vs LeNet 4 Nov 2022 · 0 repositories · arXiv:2211.02392
-
Real-Time Target Sound Extraction 4 Nov 2022 · 1 repository · arXiv:2211.02250Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
SPEAKER VGG CCT: Cross-corpus Speech Emotion Recognition with Speaker Embedding and Vision Transformers 4 Nov 2022 · 1 repository · arXiv:2211.02366
-
Alternative formulations for gilthead seabream diets: towards a more sustainable production 3 Nov 2022 · 0 repositories · arXiv:2211.02430
-
Channel-Aware Pretraining of Joint Encoder-Decoder Self-Supervised Model for Telephonic-Speech ASR 3 Nov 2022 · 0 repositories · arXiv:2211.01669
-
Crosslingual Generalization through Multitask Finetuning 3 Nov 2022 · 1 repository · arXiv:2211.01786Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Transformers on Multilingual Clause-Level Morphology 3 Nov 2022 · 1 repository · arXiv:2211.01736
-
FedTP: Federated Learning by Transformer Personalization 3 Nov 2022 · 1 repository · arXiv:2211.01572Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Pangu-Weather: A 3D High-Resolution Model for Fast and Accurate Global Weather Forecast 3 Nov 2022 · 6 repositories · arXiv:2211.02556
-
Phonetic-assisted Multi-Target Units Modeling for Improving Conformer-Transducer ASR system 3 Nov 2022 · 0 repositories · arXiv:2211.01571
-
PolyBuilding: Polygon Transformer for End-to-End Building Extraction 3 Nov 2022 · 0 repositories · arXiv:2211.01589
-
SAP-DETR: Bridging the Gap Between Salient Points and Queries-Based Transformer Detector for Fast Model Convergency 3 Nov 2022 · 1 repository · arXiv:2211.02006
-
Using Large Pre-Trained Language Model to Assist FDA in Premarket Medical Device 3 Nov 2022 · 0 repositories · arXiv:2212.01217
-
Attention-based Neural Cellular Automata 2 Nov 2022 · 0 repositories · arXiv:2211.01233
-
eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers 2 Nov 2022 · 2 repositories · arXiv:2211.01324Syntology 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
MAST: Multiscale Audio Spectrogram Transformers 2 Nov 2022 · 1 repository · arXiv:2211.01515
-
MPCFormer: fast, performant and private Transformer inference with MPC 2 Nov 2022 · 1 repository · arXiv:2211.01452
-
Pop2Piano : Pop Audio-based Piano Cover Generation 2 Nov 2022 · 4 repositories · arXiv:2211.00895Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
RegCLR: A Self-Supervised Framework for Tabular Representation Learning in the Wild 2 Nov 2022 · 0 repositories · arXiv:2211.01165
-
Transformer-based encoder-encoder architecture for Spoken Term Detection 2 Nov 2022 · 0 repositories · arXiv:2211.01089
-
WITT: A Wireless Image Transmission Transformer for Semantic Communications 2 Nov 2022 · 2 repositories · arXiv:2211.00937
-
FRSUM: Towards Faithful Abstractive Summarization via Enhancing Factual Robustness 1 Nov 2022 · 0 repositories · arXiv:2211.00294
-
Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small 1 Nov 2022 · 7 repositories · arXiv:2211.00593Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
An Empirical Study on Data Leakage and Generalizability of Link Prediction Models for Issues and Commits 1 Nov 2022 · 0 repositories · arXiv:2211.00381
-
Two-stage LLM Fine-tuning with Less Specialization and More Generalization 1 Nov 2022 · 0 repositories · arXiv:2211.00635
-
T5lephone: Bridging Speech and Text Self-supervised Models for Spoken Language Understanding via Phoneme level T5 1 Nov 2022 · 1 repository · arXiv:2211.00586
-
Text-Only Training for Image Captioning using Noise-Injected CLIP 1 Nov 2022 · 4 repositories · arXiv:2211.00575Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
ViT-DeiT: An Ensemble Model for Breast Cancer Histopathological Images Classification 1 Nov 2022 · 0 repositories · arXiv:2211.00749
-
AdaMix: Mixture-of-Adaptations for Parameter-efficient Model Tuning 31 Oct 2022 · 1 repository · arXiv:2210.17451Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Controllable Factuality in Document-Grounded Dialog Systems Using a Noisy Channel Model 31 Oct 2022 · 1 repository · arXiv:2210.17418
-
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers 31 Oct 2022 · 17 repositories · arXiv:2210.17323Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Joint Audio/Text Training for Transformer Rescorer of Streaming Speech Recognition 31 Oct 2022 · 0 repositories · arXiv:2211.00174
-
Leveraging Pre-trained Models for Failure Analysis Triplets Generation 31 Oct 2022 · 0 repositories · arXiv:2210.17497
-
Multi-Camera Calibration Free BEV Representation for 3D Object Detection 31 Oct 2022 · 0 repositories · arXiv:2210.17252
-
Probabilistic Decomposition Transformer for Time Series Forecasting 31 Oct 2022 · 1 repository · arXiv:2210.17393
-
QNet: A Quantum-native Sequence Encoder Architecture 31 Oct 2022 · 0 repositories · arXiv:2210.17262
-
QuaLA-MiniLM: a Quantized Length Adaptive MiniLM 31 Oct 2022 · 2 repositories · arXiv:2210.17114
-
Spatial-Temporal Synchronous Graph Transformer network (STSGT) for COVID-19 forecasting 31 Oct 2022 · 1 repository · arXiv:2211.00082
-
SSD-LM: Semi-autoregressive Simplex-based Diffusion Language Model for Text Generation and Modular Control 31 Oct 2022 · 2 repositories · arXiv:2210.17432
-
Structured State Space Decoder for Speech Recognition and Synthesis 31 Oct 2022 · 0 repositories · arXiv:2210.17098
-
Towards Zero-Shot and Few-Shot Table Question Answering using GPT-3 31 Oct 2022 · 0 repositories · arXiv:2210.17284
-
ViT-LSLA: Vision Transformer with Light Self-Limited-Attention 31 Oct 2022 · 0 repositories · arXiv:2210.17115
-
An Efficient Memory-Augmented Transformer for Knowledge-Intensive NLP Tasks 30 Oct 2022 · 1 repository · arXiv:2210.16773
-
Attention Swin U-Net: Cross-Contextual Attention Mechanism for Skin Lesion Segmentation 30 Oct 2022 · 1 repository · arXiv:2210.16898
-
Foreign Object Debris Detection for Airport Pavement Images based on Self-supervised Localization and Vision Transformer 30 Oct 2022 · 1 repository · arXiv:2210.16901
-
Learning to Decompose: Hypothetical Question Decomposition Based on Comparable Texts 30 Oct 2022 · 0 repositories · arXiv:2210.16865
-
QuEst: Graph Transformer for Quantum Circuit Reliability Estimation 30 Oct 2022 · 1 repository · arXiv:2210.16724
-
Time-rEversed diffusioN tEnsor Transformer: A new TENET of Few-Shot Object Detection 30 Oct 2022 · 0 repositories · arXiv:2210.16897
-
token2vec: A Joint Self-Supervised Pre-training Framework Using Unpaired Speech and Text 30 Oct 2022 · 0 repositories · arXiv:2210.16755
-
Unsupervised Learning of Structured Representations via Closed-Loop Transcription 30 Oct 2022 · 1 repository · arXiv:2210.16782
-
ViTASD: Robust Vision Transformer Baselines for Autism Spectrum Disorder Facial Diagnosis 30 Oct 2022 · 1 repository · arXiv:2210.16943
-
Interpretable CNN-Multilevel Attention Transformer for Rapid Recognition of Pneumonia from Chest X-Ray Images 29 Oct 2022 · 0 repositories · arXiv:2210.16584
-
Pair DETR: Contrastive Learning Speeds Up DETR Training 29 Oct 2022 · 0 repositories · arXiv:2210.16476
-
A Long-term Dependent and Trustworthy Approach to Reactor Accident Prognosis based on Temporal Fusion Transformer 28 Oct 2022 · 0 repositories · arXiv:2210.17298
-
Contextual Learning in Fourier Complex Field for VHR Remote Sensing Images 28 Oct 2022 · 3 repositories · arXiv:2210.15972
-
Dimensionality Reduced Antenna Array for Beamforming/steering 28 Oct 2022 · 0 repositories · arXiv:2210.16197
-
Efficient Speech Translation with Dynamic Latent Perceivers 28 Oct 2022 · 1 repository · arXiv:2210.16264
-
Exploring Spatial-Temporal Features for Deepfake Detection and Localization 28 Oct 2022 · 1 repository · arXiv:2210.15872
-
Grafting Vision Transformers 28 Oct 2022 · 0 repositories · arXiv:2210.15943
-
Modeling structure-building in the brain with CCG parsing and large language models 28 Oct 2022 · 0 repositories · arXiv:2210.16147
-
Parameter-efficient transfer learning of pre-trained Transformer models for speaker verification using adapters 28 Oct 2022 · 0 repositories · arXiv:2210.16032
-
Probing for targeted syntactic knowledge through grammatical error detection 28 Oct 2022 · 1 repository · arXiv:2210.16228
-
PSFormer: Point Transformer for 3D Salient Object Detection 28 Oct 2022 · 0 repositories · arXiv:2210.15933
-
SoftBart: Soft Bayesian Additive Regression Trees 28 Oct 2022 · 0 repositories · arXiv:2210.16375
-
UPainting: Unified Text-to-Image Diffusion Generation with Cross-modal Guidance 28 Oct 2022 · 0 repositories · arXiv:2210.16031
-
VLT: Vision-Language Transformer and Query Generation for Referring Segmentation 28 Oct 2022 · 1 repository · arXiv:2210.15871Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
COCO-DR: Combating Distribution Shifts in Zero-Shot Dense Retrieval with Contrastive and Distributionally Robust Learning 27 Oct 2022 · 1 repository · arXiv:2210.15212Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Fast DistilBERT on CPUs 27 Oct 2022 · 1 repository · arXiv:2211.07715
-
GaitMixer: Skeleton-based Gait Representation Learning via Wide-spectrum Multi-axial Mixer 27 Oct 2022 · 1 repository · arXiv:2210.15491
-
HYDRA-HGR: A Hybrid Transformer-based Architecture for Fusion of Macroscopic and Microscopic Neural Drive Information 27 Oct 2022 · 0 repositories · arXiv:2211.02619
-
Li3DeTr: A LiDAR based 3D Detection Transformer 27 Oct 2022 · 0 repositories · arXiv:2210.15365
-
Make More of Your Data: Minimal Effort Data Augmentation for Automatic Speech Recognition and Translation 27 Oct 2022 · 0 repositories · arXiv:2210.15398
-
Masked Transformer for image Anomaly Localization 27 Oct 2022 · 0 repositories · arXiv:2210.15540
-
MSF3DDETR: Multi-Sensor Fusion 3D Detection Transformer for Autonomous Driving 27 Oct 2022 · 0 repositories · arXiv:2210.15316
-
Multimodal Transformer Distillation for Audio-Visual Synchronization 27 Oct 2022 · 2 repositories · arXiv:2210.15563
-
Point-Voxel Adaptive Feature Abstraction for Robust Point Cloud Classification 27 Oct 2022 · 1 repository · arXiv:2210.15514
-
ProContEXT: Exploring Progressive Context Transformer for Tracking 27 Oct 2022 · 4 repositories · arXiv:2210.15511
-
Spatio-Temporal Hybrid Fusion of CAE and SWIn Transformers for Lung Cancer Malignancy Prediction 27 Oct 2022 · 0 repositories · arXiv:2210.15297
-
The 1st-place Solution for ECCV 2022 Multiple People Tracking in Group Dance Challenge 27 Oct 2022 · 3 repositories · arXiv:2210.15281
-
Transformers meet Stochastic Block Models: Attention with Data-Adaptive Sparsity and Cost 27 Oct 2022 · 1 repository · arXiv:2210.15541
-
TRScore: A Novel GPT-based Readability Scorer for ASR Segmentation and Punctuation model evaluation and selection 27 Oct 2022 · 0 repositories · arXiv:2210.15104
-
What Language Model to Train if You Have One Million GPU Hours? 27 Oct 2022 · 1 repository · arXiv:2210.15424
-
Working Alliance Transformer for Psychotherapy Dialogue Classification 27 Oct 2022 · 1 repository · arXiv:2210.15603
-
Automatic Diagnosis of Myocarditis Disease in Cardiac MRI Modality using Deep Transformers and Explainable Artificial Intelligence 26 Oct 2022 · 0 repositories · arXiv:2210.14611
-
Beyond English-Centric Bitexts for Better Multilingual Language Representation Learning 26 Oct 2022 · 0 repositories · arXiv:2210.14867
-
Disentangling Past-Future Modeling in Sequential Recommendation via Dual Networks 26 Oct 2022 · 1 repository · arXiv:2210.14577