Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 129
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 129 of 190: papers 12,801 to 12,900 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Recommending Root-Cause and Mitigation Steps for Cloud Incidents using Large Language Models 10 Jan 2023 · 0 repositories · arXiv:2301.03797
-
Scaling Laws for Generative Mixed-Modal Language Models 10 Jan 2023 · 0 repositories · arXiv:2301.03728
-
Streaming Punctuation: A Novel Punctuation Technique Leveraging Bidirectional Context for Continuous Speech Recognition 10 Jan 2023 · 0 repositories · arXiv:2301.03819
-
There is No Big Brother or Small Brother: Knowledge Infusion in Language Models for Link Prediction and Question Answering 10 Jan 2023 · 2 repositories · arXiv:2301.04013
-
Unsupervised Mandarin-Cantonese Machine Translation 10 Jan 2023 · 1 repository · arXiv:2301.03971
-
A Study on the Generality of Neural Network Structures for Monocular Depth Estimation 9 Jan 2023 · 1 repository · arXiv:2301.03169
-
Advances in Medical Image Analysis with Vision Transformers: A Comprehensive Review 9 Jan 2023 · 1 repository · arXiv:2301.03505
-
An Impartial Transformer for Story Visualization 9 Jan 2023 · 0 repositories · arXiv:2301.03563
-
DeMT: Deformable Mixer Transformer for Multi-Task Learning of Dense Prediction 9 Jan 2023 · 1 repository · arXiv:2301.03461Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Logically at Factify 2: A Multi-Modal Fact Checking System Based on Evidence Retrieval techniques and Transformer Encoder Architecture 9 Jan 2023 · 0 repositories · arXiv:2301.03127
-
Universal Multimodal Representation for Language Understanding 9 Jan 2023 · 0 repositories · arXiv:2301.03344
-
Automatic Generation of German Drama Texts Using Fine Tuned GPT-2 Models 8 Jan 2023 · 0 repositories · arXiv:2301.03119
-
DeepMatcher: A Deep Transformer-based Network for Robust and Accurate Local Feature Matching 8 Jan 2023 · 1 repository · arXiv:2301.02993
-
HRTransNet: HRFormer-Driven Two-Modality Salient Object Detection 8 Jan 2023 · 1 repository · arXiv:2301.03036
-
CodeTalker: Speech-Driven 3D Facial Animation with Discrete Motion Prior 6 Jan 2023 · 1 repository · arXiv:2301.02379Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples)
-
Generative Antibody Design for Complementary Chain Pairing Sequences through Encoder-Decoder Language Model 6 Jan 2023 · 0 repositories · arXiv:2301.02748
-
Does compressing activations help model parallel training? 6 Jan 2023 · 0 repositories · arXiv:2301.02654
-
Exploring Efficient Few-shot Adaptation for Vision Transformers 6 Jan 2023 · 1 repository · arXiv:2301.02419
-
Multi-Genre Music Transformer -- Composing Full Length Musical Piece 6 Jan 2023 · 0 repositories · arXiv:2301.02385
-
Systems for Parallel and Distributed Large-Model Deep Learning Training 6 Jan 2023 · 0 repositories · arXiv:2301.02691
-
Adaptive Pattern Extraction Multi-Task Learning for Multi-Step Conversion Estimations 6 Jan 2023 · 0 repositories · arXiv:2301.02494
-
Adaptively Clustering Neighbor Elements for Image-Text Generation 5 Jan 2023 · 1 repository · arXiv:2301.01955
-
CAT: LoCalization and IdentificAtion Cascade Detection Transformer for Open-World Object Detection 5 Jan 2023 · 0 repositories · arXiv:2301.01970
-
Critical Perspectives: A Benchmark Revealing Pitfalls in PerspectiveAPI 5 Jan 2023 · 1 repository · arXiv:2301.01874
-
Learning Feature Recovery Transformer for Occluded Person Re-identification 5 Jan 2023 · 1 repository · arXiv:2301.01879
-
Scalable Communication for Multi-Agent Reinforcement Learning via Transformer-Based Email Mechanism 5 Jan 2023 · 0 repositories · arXiv:2301.01919
-
Sequentially Controlled Text Generation 5 Jan 2023 · 0 repositories · arXiv:2301.02299
-
Towards Autoformalization of Mathematics and Code Correctness: Experiments with Elementary Proofs 5 Jan 2023 · 1 repository · arXiv:2301.02195Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Towards Long-Term Time-Series Forecasting: Feature, Pattern, and Distribution 5 Jan 2023 · 1 repository · arXiv:2301.02068
-
Extending Source Code Pre-Trained Language Models to Summarise Decompiled Binaries 4 Jan 2023 · 1 repository · arXiv:2301.01701
-
Infomaxformer: Maximum Entropy Transformer for Long Time-Series Forecasting Problem 4 Jan 2023 · 0 repositories · arXiv:2301.01772
-
InPars-v2: Large Language Models as Efficient Dataset Generators for Information Retrieval 4 Jan 2023 · 1 repository · arXiv:2301.01820
-
Multi-Aspect Explainable Inductive Relation Prediction by Sentence Transformer 4 Jan 2023 · 1 repository · arXiv:2301.01664
-
Semi-MAE: Masked Autoencoders for Semi-supervised Vision Transformers 4 Jan 2023 · 0 repositories · arXiv:2301.01431
-
SPTS v2: Single-Point Scene Text Spotting 4 Jan 2023 · 3 repositories · arXiv:2301.01635Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples) · 3 pointer-only (licence)
-
UniHD at TSAR-2022 Shared Task: Is Compute All We Need for Lexical Simplification? 4 Jan 2023 · 1 repository · arXiv:2301.01764
-
A New Perspective to Boost Vision Transformer for Medical Image Classification 3 Jan 2023 · 0 repositories · arXiv:2301.00989
-
Cross Modal Transformer: Towards Fast and Robust 3D Object Detection 3 Jan 2023 · 2 repositories · arXiv:2301.01283
-
Large Language Models as Corporate Lobbyists 3 Jan 2023 · 1 repository · arXiv:2301.01181
-
Medical Image Segmentation via Cascaded Attention Decoding 3 Jan 2023 · 1 repository
-
Modeling the Rhythm from Lyrics for Melody Generation of Pop Song 3 Jan 2023 · 0 repositories · arXiv:2301.01361
-
Rethinking Mobile Block for Efficient Attention-based Models 3 Jan 2023 · 1 repository · arXiv:2301.01146
-
TinyMIM: An Empirical Study of Distilling MIM Pre-trained Models 3 Jan 2023 · 2 repositories · arXiv:2301.01296Syntology official (archive's flag): 2 ran · 8 ran (of which 5 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Betrayed by Captions: Joint Caption Grounding and Generation for Open Vocabulary Instance Segmentation 2 Jan 2023 · 2 repositories · arXiv:2301.00805
-
MAUD: An Expert-Annotated Legal NLP Dataset for Merger Agreement Understanding 2 Jan 2023 · 2 repositories · arXiv:2301.00876
-
Multi-Stage Spatio-Temporal Aggregation Transformer for Video Person Re-identification 2 Jan 2023 · 0 repositories · arXiv:2301.00531
-
Muse: Text-To-Image Generation via Masked Generative Transformers 2 Jan 2023 · 5 repositories · arXiv:2301.00704Syntology 19 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 2 honoured, 5 violated, 1 with no contract checked; 11 where Syntology's instrument failed) · 2 unverified (of 21 harvested samples) · 11 pointer-only (licence)
-
Transformer Based Geocoding 2 Jan 2023 · 0 repositories · arXiv:2301.01170
-
3DPPE: 3D Point Positional Encoding for Transformer-based Multi-Camera 3D Object Detection 1 Jan 2023 · 1 repository
-
A Large-Scale Robustness Analysis of Video Action Recognition Models 1 Jan 2023 · 0 repositories
-
AUNet: Learning Relations Between Action Units for Face Forgery Detection 1 Jan 2023 · 0 repositories
-
Automated Knowledge Distillation via Monte Carlo Tree Search 1 Jan 2023 · 0 repositories
-
Bit-Shrinking: Limiting Instantaneous Sharpness for Improving Post-Training Quantization 1 Jan 2023 · 0 repositories
-
Boosting Whole Slide Image Classification from the Perspectives of Distribution, Correlation and Magnification 1 Jan 2023 · 0 repositories
-
Building Vision Transformers with Hierarchy Aware Feature Aggregation 1 Jan 2023 · 0 repositories
-
BUS: Efficient and Effective Vision-Language Pre-Training with Bottom-Up Patch Summarization. 1 Jan 2023 · 0 repositories
-
CLIPPING: Distilling CLIP-Based Models With a Student Base for Video-Language Retrieval 1 Jan 2023 · 0 repositories
-
CO-PILOT: Dynamic Top-Down Point Cloud with Conditional Neighborhood Aggregation for Multi-Gigapixel Histopathology Image Representation 1 Jan 2023 · 0 repositories
-
Comprehensive and Delicate: An Efficient Transformer for Image Restoration 1 Jan 2023 · 1 repository
-
DETR Does Not Need Multi-Scale or Locality Design 1 Jan 2023 · 1 repository
-
DKT: Diverse Knowledge Transfer Transformer for Class Incremental Learning 1 Jan 2023 · 0 repositories
-
DropKey for Vision Transformer 1 Jan 2023 · 0 repositories
-
Few-Shot Video Classification via Representation Fusion and Promotion Learning 1 Jan 2023 · 0 repositories
-
Focal Network for Image Restoration 1 Jan 2023 · 1 repository
-
Foreground-Background Distribution Modeling Transformer for Visual Object Tracking 1 Jan 2023 · 0 repositories
-
Fusing Pre-Trained Language Models With Multimodal Prompts Through Reinforcement Learning 1 Jan 2023 · 1 repository
-
Generating Human Motion From Textual Descriptions With Discrete Representations 1 Jan 2023 · 0 repositories
-
Geometrized Transformer for Self-Supervised Homography Estimation 1 Jan 2023 · 1 repository
-
Goal-Guided Transformer-Enabled Reinforcement Learning for Efficient Autonomous Navigation 1 Jan 2023 · 1 repository · arXiv:2301.00362
-
Heat Diffusion Based Multi-Scale and Geometric Structure-Aware Transformer for Mesh Segmentation 1 Jan 2023 · 0 repositories
-
HGNet: Learning Hierarchical Geometry From Points, Edges, and Surfaces 1 Jan 2023 · 0 repositories
-
HSR-Diff: Hyperspectral Image Super-Resolution via Conditional Diffusion Models 1 Jan 2023 · 0 repositories
-
Image To Tree with Recursive Prompting 1 Jan 2023 · 0 repositories · arXiv:2301.00447
-
Is word segmentation necessary for Vietnamese sentiment classification? 1 Jan 2023 · 0 repositories · arXiv:2301.00418
-
LaPE: Layer-adaptive Position Embedding for Vision Transformers with Independent Layer Normalization 1 Jan 2023 · 1 repository
-
Learning Long-Range Information with Dual-Scale Transformers for Indoor Scene Completion 1 Jan 2023 · 0 repositories
-
Lite DETR: An Interleaved Multi-Scale Encoder for Efficient DETR 1 Jan 2023 · 0 repositories
-
LNPL-MIL: Learning from Noisy Pseudo Labels for Promoting Multiple Instance Learning in Whole Slide Image 1 Jan 2023 · 0 repositories
-
Masked Auto-Encoders Meet Generative Adversarial Networks and Beyond 1 Jan 2023 · 1 repository
-
MSRA-SR: Image Super-resolution Transformer with Multi-scale Shared Representation Acquisition 1 Jan 2023 · 0 repositories
-
PEAL: Prior-Embedded Explicit Attention Learning for Low-Overlap Point Cloud Registration 1 Jan 2023 · 1 repository
-
PointClustering: Unsupervised Point Cloud Pre-Training Using Transformation Invariance in Clustering 1 Jan 2023 · 1 repository
-
PointListNet: Deep Learning on 3D Point Lists 1 Jan 2023 · 0 repositories
-
Polarized Color Image Denoising 1 Jan 2023 · 0 repositories
-
PromptCap: Prompt-Guided Image Captioning for VQA with GPT-3 1 Jan 2023 · 0 repositories
-
Query Refinement Transformer for 3D Instance Segmentation 1 Jan 2023 · 0 repositories
-
Rethinking Point Cloud Registration as Masking and Reconstruction 1 Jan 2023 · 1 repository
-
Segment Every Reference Object in Spatial and Temporal Spaces 1 Jan 2023 · 0 repositories
-
SemiCVT: Semi-Supervised Convolutional Vision Transformer for Semantic Segmentation 1 Jan 2023 · 0 repositories
-
Single Image Deblurring with Row-dependent Blur Magnitude 1 Jan 2023 · 0 repositories
-
SkeleTR: Towards Skeleton-based Action Recognition in the Wild 1 Jan 2023 · 0 repositories
-
SKiT: a Fast Key Information Video Transformer for Online Surgical Phase Recognition 1 Jan 2023 · 1 repository
-
Sparse Multi-Modal Graph Transformer With Shared-Context Processing for Representation Learning of Giga-Pixel Images 1 Jan 2023 · 0 repositories
-
SwinLSTM: Improving Spatiotemporal Prediction Accuracy using Swin Transformer and LSTM 1 Jan 2023 · 1 repository
-
Thinking Image Color Aesthetics Assessment: Models, Datasets and Benchmarks 1 Jan 2023 · 1 repository
-
TokenHPE: Learning Orientation Tokens for Efficient Head Pose Estimation via Transformers 1 Jan 2023 · 1 repository
-
Translating Images to Road Network: A Non-Autoregressive Sequence-to-Sequence Approach 1 Jan 2023 · 0 repositories
-
Uni-3D: A Universal Model for Panoptic 3D Scene Reconstruction 1 Jan 2023 · 1 repository
-
Weakly Supervised Referring Image Segmentation with Intra-Chunk and Inter-Chunk Consistency 1 Jan 2023 · 0 repositories
-
Rethinking Rotation Invariance with Point Cloud Registration 31 Dec 2022 · 0 repositories · arXiv:2301.00149