Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 122
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 122 of 190: papers 12,101 to 12,200 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Distilling Token-Pruned Pose Transformer for 2D Human Pose Estimation 12 Apr 2023 · 0 repositories · arXiv:2304.05548
-
DUFormer: Solving Power Line Detection Task in Aerial Images using Semantic Segmentation 12 Apr 2023 · 0 repositories · arXiv:2304.05821
-
Evaluation of ChatGPT Model for Vulnerability Detection 12 Apr 2023 · 0 repositories · arXiv:2304.07232
-
Galactic ChitChat: Using Large Language Models to Converse with Astronomy Literature 12 Apr 2023 · 0 repositories · arXiv:2304.05406
-
Localizing Model Behavior with Path Patching 12 Apr 2023 · 1 repository · arXiv:2304.05969Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
MED-VT++: Unifying Multimodal Learning with a Multiscale Encoder-Decoder Video Transformer 12 Apr 2023 · 0 repositories · arXiv:2304.05930
-
Multi-scale Geometry-aware Transformer for 3D Point Cloud Classification 12 Apr 2023 · 0 repositories · arXiv:2304.05694
-
PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting 12 Apr 2023 · 2 repositories · arXiv:2304.06107
-
Real-time Trajectory-based Social Group Detection 12 Apr 2023 · 1 repository · arXiv:2304.05678
-
Towards Evaluating Explanations of Vision Transformers for Medical Imaging 12 Apr 2023 · 1 repository · arXiv:2304.06133
-
Approximating Online Human Evaluation of Social Chatbots with Prompting 11 Apr 2023 · 0 repositories · arXiv:2304.05253
-
Bayesian Optimization of Catalysis With In-Context Learning 11 Apr 2023 · 2 repositories · arXiv:2304.05341
-
chatClimate: Grounding Conversational AI in Climate Science 11 Apr 2023 · 0 repositories · arXiv:2304.05510
-
ChemCrow: Augmenting large-language models with chemistry tools 11 Apr 2023 · 3 repositories · arXiv:2304.05376Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Data-Efficient Image Quality Assessment with Attention-Panel Decoder 11 Apr 2023 · 1 repository · arXiv:2304.04952Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis 11 Apr 2023 · 0 repositories · arXiv:2304.05534
-
Exploring the Use of Foundation Models for Named Entity Recognition and Lemmatization Tasks in Slavic Languages 11 Apr 2023 · 0 repositories · arXiv:2304.05336
-
MC-ViViT: Multi-branch Classifier-ViViT to detect Mild Cognitive Impairment in older adults using facial videos 11 Apr 2023 · 0 repositories · arXiv:2304.05292
-
Multi-Graph Convolution Network for Pose Forecasting 11 Apr 2023 · 0 repositories · arXiv:2304.04956
-
Multi-step Jailbreaking Privacy Attacks on ChatGPT 11 Apr 2023 · 1 repository · arXiv:2304.05197Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Sim-T: Simplify the Transformer Network by Multiplexing Technique for Speech Recognition 11 Apr 2023 · 0 repositories · arXiv:2304.04991
-
Training Large Language Models Efficiently with Sparsity and Dataflow 11 Apr 2023 · 0 repositories · arXiv:2304.05511
-
Video Event Restoration Based on Keyframes for Video Anomaly Detection 11 Apr 2023 · 0 repositories · arXiv:2304.05112
-
Automated Reading Passage Generation with OpenAI's Large Language Model 10 Apr 2023 · 0 repositories · arXiv:2304.04616
-
Detection Transformer with Stable Matching 10 Apr 2023 · 2 repositories · arXiv:2304.04742Syntology official (archive's flag): 1 ran · 7 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Feature Representation Learning with Adaptive Displacement Generation and Transformer Fusion for Micro-Expression Recognition 10 Apr 2023 · 0 repositories · arXiv:2304.04420
-
High Dynamic Range Imaging with Context-aware Transformer 10 Apr 2023 · 0 repositories · arXiv:2304.04416
-
HST-MRF: Heterogeneous Swin Transformer with Multi-Receptive Field for Medical Image Segmentation 10 Apr 2023 · 0 repositories · arXiv:2304.04614
-
Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis 10 Apr 2023 · 2 repositories · arXiv:2304.04675
-
On the Possibilities of AI-Generated Text Detection 10 Apr 2023 · 0 repositories · arXiv:2304.04736
-
Two Steps Forward and One Behind: Rethinking Time Series Forecasting with Deep Learning 10 Apr 2023 · 0 repositories · arXiv:2304.04553
-
Use the Detection Transformer as a Data Augmenter 10 Apr 2023 · 1 repository · arXiv:2304.04554
-
Are Large Language Models Ready for Healthcare? A Comparative Study on Clinical Language Understanding 9 Apr 2023 · 1 repository · arXiv:2304.05368
-
Slide-Transformer: Hierarchical Vision Transformer with Local Self-Attention 9 Apr 2023 · 1 repository · arXiv:2304.04237Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Sparse Dense Fusion for 3D Object Detection 9 Apr 2023 · 0 repositories · arXiv:2304.04179
-
Transformer Utilization in Medical Image Segmentation Networks 9 Apr 2023 · 0 repositories · arXiv:2304.04225
-
Factify 2: A Multimodal Fake News and Satire News Dataset 8 Apr 2023 · 1 repository · arXiv:2304.03897
-
GPT4Rec: A Generative Framework for Personalized Recommendation and User Interests Interpretation 8 Apr 2023 · 0 repositories · arXiv:2304.03879
-
Surrogate Lagrangian Relaxation: A Path To Retrain-free Deep Neural Network Pruning 8 Apr 2023 · 0 repositories · arXiv:2304.04120
-
A Cross-Scale Hierarchical Transformer with Correspondence-Augmented Attention for inferring Bird's-Eye-View Semantic Segmentation 7 Apr 2023 · 0 repositories · arXiv:2304.03650
-
Cleansing Jewel: A Neural Spelling Correction Model Built On Google OCR-ed Tibetan Manuscripts 7 Apr 2023 · 0 repositories · arXiv:2304.03427
-
Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4 7 Apr 2023 · 1 repository · arXiv:2304.03439
-
Hierarchical Catalogue Generation for Literature Review: A Benchmark 7 Apr 2023 · 1 repository · arXiv:2304.03512
-
PSLT: A Light-weight Vision Transformer with Ladder Self-Attention and Progressive Shift 7 Apr 2023 · 0 repositories · arXiv:2304.03481
-
SparseFormer: Sparse Visual Recognition via Limited Latent Tokens 7 Apr 2023 · 1 repository · arXiv:2304.03768
-
All Keypoints You Need: Detecting Arbitrary Keypoints on the Body of Triple, High, and Long Jump Athletes 6 Apr 2023 · 1 repository · arXiv:2304.02939
-
Can Large Language Models Play Text Games Well? Current State-of-the-Art and Open Questions 6 Apr 2023 · 0 repositories · arXiv:2304.02868
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
Continual Detection Transformer for Incremental Object Detection 6 Apr 2023 · 0 repositories · arXiv:2304.03110
-
Deep Learning for Opinion Mining and Topic Classification of Course Reviews 6 Apr 2023 · 0 repositories · arXiv:2304.03394
-
DeLiRa: Self-Supervised Depth, Light, and Radiance Fields 6 Apr 2023 · 0 repositories · arXiv:2304.02797
-
Efficient Audio Captioning Transformer with Patchout and Text Guidance 6 Apr 2023 · 0 repositories · arXiv:2304.02916
-
FengWu: Pushing the Skillful Global Medium-range Weather Forecast beyond 10 Days Lead 6 Apr 2023 · 2 repositories · arXiv:2304.02948
-
GPT detectors are biased against non-native English writers 6 Apr 2023 · 2 repositories · arXiv:2304.02819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Instruction Tuning with GPT-4 6 Apr 2023 · 2 repositories · arXiv:2304.03277
-
Making AI Less "Thirsty": Uncovering and Addressing the Secret Water Footprint of AI Models 6 Apr 2023 · 1 repository · arXiv:2304.03271Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
MemeFier: Dual-stage Modality Fusion for Image Meme Classification 6 Apr 2023 · 1 repository · arXiv:2304.02906
-
MULLER: Multilayer Laplacian Resizer for Vision 6 Apr 2023 · 1 repository · arXiv:2304.02859
-
Towards Interpretable Mental Health Analysis with Large Language Models 6 Apr 2023 · 2 repositories · arXiv:2304.03347
-
PointCAT: Cross-Attention Transformer for point cloud 6 Apr 2023 · 1 repository · arXiv:2304.03012
-
Towards an Effective and Efficient Transformer for Rain-by-snow Weather Removal 6 Apr 2023 · 1 repository · arXiv:2304.02860
-
Zero-Shot Next-Item Recommendation using Large Pretrained Language Models 6 Apr 2023 · 1 repository · arXiv:2304.03153Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Conceptual structure coheres in human cognition but not in large language models 5 Apr 2023 · 0 repositories · arXiv:2304.02754
-
ChartReader: A Unified Framework for Chart Derendering and Comprehension without Heuristic Rules 5 Apr 2023 · 1 repository · arXiv:2304.02173Syntology official (archive's flag): 16 ran · 16 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 10 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Context-Aware Classification of Legal Document Pages 5 Apr 2023 · 0 repositories · arXiv:2304.02787
-
Document-Level Machine Translation with Large Language Models 5 Apr 2023 · 1 repository · arXiv:2304.02210
-
Face Transformer: Towards High Fidelity and Accurate Face Swapping 5 Apr 2023 · 0 repositories · arXiv:2304.02530
-
Large Language Models as Master Key: Unlocking the Secrets of Materials Science with GPT 5 Apr 2023 · 0 repositories · arXiv:2304.02213
-
ParroT: Translating during Chat using Large Language Models tuned with Human Translation and Feedback 5 Apr 2023 · 1 repository · arXiv:2304.02426
-
Training Strategies for Vision Transformers for Object Detection 5 Apr 2023 · 0 repositories · arXiv:2304.02186
-
Blockwise Compression of Transformer-based Models without Retraining 4 Apr 2023 · 0 repositories · arXiv:2304.01483
-
Can BERT eat RuCoLA? Topological Data Analysis to Explain 4 Apr 2023 · 2 repositories · arXiv:2304.01680
-
ConvFormer: Parameter Reduction in Transformer Models for 3D Human Pose Estimation by Leveraging Dynamic Multi-Headed Convolutional Attention 4 Apr 2023 · 1 repository · arXiv:2304.02147
-
FedBot: Enhancing Privacy in Chatbots with Federated Learning 4 Apr 2023 · 0 repositories · arXiv:2304.03228
-
Generating Natural Language from Logic Expressions with Structural Representation 4 Apr 2023 · 1 repository
-
Geotechnical Parrot Tales (GPT): Harnessing Large Language Models in geotechnical engineering 4 Apr 2023 · 0 repositories · arXiv:2304.02138
-
GPT-4 to GPT-3.5: 'Hold My Scalpel' -- A Look at the Competency of OpenAI's GPT on the Plastic Surgery In-Service Training Exam 4 Apr 2023 · 0 repositories · arXiv:2304.01503
-
Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation 4 Apr 2023 · 0 repositories · arXiv:2304.01746
-
LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models 4 Apr 2023 · 2 repositories · arXiv:2304.01933Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC Era 4 Apr 2023 · 0 repositories · arXiv:2304.06488
-
Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder 4 Apr 2023 · 0 repositories · arXiv:2304.01611
-
REFINER: Reasoning Feedback on Intermediate Representations 4 Apr 2023 · 1 repository · arXiv:2304.01904Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Self-Supervised Image Denoising for Real-World Images with Context-aware Transformer 4 Apr 2023 · 0 repositories · arXiv:2304.01627
-
Strong Baselines for Parameter Efficient Few-Shot Fine-tuning 4 Apr 2023 · 0 repositories · arXiv:2304.01917
-
Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models 4 Apr 2023 · 0 repositories · arXiv:2304.01852
-
Towards Open-Vocabulary Video Instance Segmentation 4 Apr 2023 · 1 repository · arXiv:2304.01715
-
Controllable Motion Synthesis and Reconstruction with Autoregressive Diffusion Models 3 Apr 2023 · 0 repositories · arXiv:2304.04681
-
Crossword: A Semantic Approach to Data Compression via Masking 3 Apr 2023 · 0 repositories · arXiv:2304.01106
-
DoctorGLM: Fine-tuning your Chinese Doctor is not a Herculean Task 3 Apr 2023 · 1 repository · arXiv:2304.01097
-
GreekBART: The First Pretrained Greek Sequence-to-Sequence Model 3 Apr 2023 · 2 repositories · arXiv:2304.00869
-
Classification of integers based on residue classes via modern deep learning algorithms 3 Apr 2023 · 1 repository · arXiv:2304.01333
-
PEACH: Pre-Training Sequence-to-Sequence Multilingual Models for Translation with Semi-Supervised Pseudo-Parallel Document Generation 3 Apr 2023 · 1 repository · arXiv:2304.01282
-
ReMoDiffuse: Retrieval-Augmented Motion Diffusion Model 3 Apr 2023 · 1 repository · arXiv:2304.01116
-
RePAST: Relative Pose Attention Scene Representation Transformer 3 Apr 2023 · 0 repositories · arXiv:2304.00947
-
Spectral Enhanced Rectangle Transformer for Hyperspectral Image Denoising 3 Apr 2023 · 1 repository · arXiv:2304.00844
-
The StatCan Dialogue Dataset: Retrieving Data Tables through Conversations with Genuine Intents 3 Apr 2023 · 1 repository · arXiv:2304.01412
-
U-Netmer: U-Net meets Transformer for medical image segmentation 3 Apr 2023 · 0 repositories · arXiv:2304.01401
-
Does Human Collaboration Enhance the Accuracy of Identifying LLM-Generated Deepfake Texts? 3 Apr 2023 · 2 repositories · arXiv:2304.01002Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
WeakTr: Exploring Plain Vision Transformer for Weakly-supervised Semantic Segmentation 3 Apr 2023 · 1 repository · arXiv:2304.01184Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Better Language Models of Code through Self-Improvement 2 Apr 2023 · 1 repository · arXiv:2304.01228