Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 112
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 112 of 190: papers 11,101 to 11,200 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A Comparison of Time-based Models for Multimodal Emotion Recognition 22 Jun 2023 · 0 repositories · arXiv:2306.13076
-
Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs 22 Jun 2023 · 1 repository · arXiv:2306.13063Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Cross-lingual Cross-temporal Summarization: Dataset, Models, Evaluation 22 Jun 2023 · 1 repository · arXiv:2306.12916
-
Deep Metric Learning with Soft Orthogonal Proxies 22 Jun 2023 · 0 repositories · arXiv:2306.13055
-
Learning from Visual Observation via Offline Pretrained State-to-Go Transformer 22 Jun 2023 · 0 repositories · arXiv:2306.12860
-
Minimalist and High-Quality Panoramic Imaging with PSF-aware Transformers 22 Jun 2023 · 1 repository · arXiv:2306.12992
-
Prompt to GPT-3: Step-by-Step Thinking Instructions for Humor Generation 22 Jun 2023 · 1 repository · arXiv:2306.13195
-
Visual Adversarial Examples Jailbreak Aligned Large Language Models 22 Jun 2023 · 1 repository · arXiv:2306.13213
-
ARIES: A Corpus of Scientific Paper Edits Made in Response to Peer Reviews 21 Jun 2023 · 1 repository · arXiv:2306.12587Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair 21 Jun 2023 · 0 repositories · arXiv:2307.00012
-
Joint Prompt Optimization of Stacked LLMs using Variational Inference 21 Jun 2023 · 1 repository · arXiv:2306.12509Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 16 unverified (of 19 harvested samples)
-
Fast Segment Anything 21 Jun 2023 · 1 repository · arXiv:2306.12156
-
GPT-Based Models Meet Simulation: How to Efficiently Use Large-Scale Pre-Trained Language Models Across Simulation Tasks 21 Jun 2023 · 0 repositories · arXiv:2306.13679
-
HSR-Diff:Hyperspectral Image Super-Resolution via Conditional Diffusion Models 21 Jun 2023 · 0 repositories · arXiv:2306.12085
-
Inter-Instance Similarity Modeling for Contrastive Learning 21 Jun 2023 · 1 repository · arXiv:2306.12243
-
Investigating Pre-trained Language Models on Cross-Domain Datasets, a Step Closer to General AI 21 Jun 2023 · 0 repositories · arXiv:2306.12205
-
MSW-Transformer: Multi-Scale Shifted Windows Transformer Networks for 12-Lead ECG Classification 21 Jun 2023 · 0 repositories · arXiv:2306.12098
-
Neural Multigrid Memory For Computational Fluid Dynamics 21 Jun 2023 · 1 repository · arXiv:2306.12545
-
Probing the limit of hydrologic predictability with the Transformer network 21 Jun 2023 · 0 repositories · arXiv:2306.12384
-
Solving and Generating NPR Sunday Puzzles with Large Language Models 21 Jun 2023 · 1 repository · arXiv:2306.12255
-
StarVQA+: Co-training Space-Time Attention for Video Quality Assessment 21 Jun 2023 · 1 repository · arXiv:2306.12298
-
What Constitutes Good Contrastive Learning in Time-Series Forecasting? 21 Jun 2023 · 1 repository · arXiv:2306.12086Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Which Spurious Correlations Impact Reasoning in NLI Models? A Visual Interactive Diagnosis through Data-Constrained Counterfactuals 21 Jun 2023 · 0 repositories · arXiv:2306.12146
-
A Novel Counterfactual Data Augmentation Method for Aspect-Based Sentiment Analysis 20 Jun 2023 · 0 repositories · arXiv:2306.11260
-
Comparing Deep Learning Models for the Task of Volatility Prediction Using Multivariate Data 20 Jun 2023 · 0 repositories · arXiv:2306.12446
-
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models 20 Jun 2023 · 0 repositories · arXiv:2306.11698
-
Deep Fusion: Efficient Network Training via Pre-trained Initializations 20 Jun 2023 · 0 repositories · arXiv:2306.11903
-
Democratizing LLMs for Low-Resource Languages by Leveraging their English Dominant Abilities with Linguistically-Diverse Prompts 20 Jun 2023 · 0 repositories · arXiv:2306.11372
-
Event Stream GPT: A Data Pre-processing and Modeling Library for Generative, Pre-trained Transformers over Continuous-time Sequences of Complex Events 20 Jun 2023 · 1 repository · arXiv:2306.11547Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
A GPT-4 Reticular Chemist for Guiding MOF Discovery 20 Jun 2023 · 1 repository · arXiv:2306.14915
-
Harnessing the Power of Adversarial Prompting and Large Language Models for Robust Hypothesis Generation in Astronomy 20 Jun 2023 · 0 repositories · arXiv:2306.11648
-
InRank: Incremental Low-Rank Learning 20 Jun 2023 · 1 repository · arXiv:2306.11250
-
Learning to Generate Better Than Your LLM 20 Jun 2023 · 1 repository · arXiv:2306.11816
-
Multiverse Transformer: 1st Place Solution for Waymo Open Sim Agents Challenge 2023 20 Jun 2023 · 0 repositories · arXiv:2306.11868
-
Surfer: Progressive Reasoning with World Models for Robotic Manipulation 20 Jun 2023 · 0 repositories · arXiv:2306.11335
-
Textbooks Are All You Need 20 Jun 2023 · 0 repositories · arXiv:2306.11644
-
Transforming Graphs for Enhanced Attribute Clustering: An Innovative Graph Transformer-Based Method 20 Jun 2023 · 0 repositories · arXiv:2306.11307
-
Unfolding Framework with Prior of Convolution-Transformer Mixture and Uncertainty Estimation for Video Snapshot Compressive Imaging 20 Jun 2023 · 0 repositories · arXiv:2306.11316
-
A Preliminary Study of ChatGPT on News Recommendation: Personalization, Provider Fairness, Fake News 19 Jun 2023 · 1 repository · arXiv:2306.10702
-
AMRs Assemble! Learning to Ensemble with Autoregressive Models for AMR Parsing 19 Jun 2023 · 1 repository · arXiv:2306.10786
-
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language Models 19 Jun 2023 · 1 repository · arXiv:2306.10968
-
Fine-Tuning Language Models for Scientific Writing Support 19 Jun 2023 · 1 repository · arXiv:2306.10974
-
Generative Sequential Recommendation with GPTRec 19 Jun 2023 · 0 repositories · arXiv:2306.11114
-
Multi-task Learning for Radar Signal Characterisation 19 Jun 2023 · 1 repository · arXiv:2306.13105
-
Multitrack Music Transcription with a Time-Frequency Perceiver 19 Jun 2023 · 0 repositories · arXiv:2306.10785
-
NAR-Former V2: Rethinking Transformer for Universal Neural Network Representation Learning 19 Jun 2023 · 1 repository · arXiv:2306.10792Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
RaViTT: Random Vision Transformer Tokens 19 Jun 2023 · 0 repositories · arXiv:2306.10959
-
RedMotion: Motion Prediction via Redundancy Reduction 19 Jun 2023 · 3 repositories · arXiv:2306.10840Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SynerGPT: In-Context Learning for Personalized Drug Synergy Prediction and Drug Design 19 Jun 2023 · 0 repositories · arXiv:2307.11694
-
Temporal Data Meets LLM -- Explainable Financial Time Series Forecasting 19 Jun 2023 · 0 repositories · arXiv:2306.11025
-
Transformer Training Strategies for Forecasting Multiple Load Time Series 19 Jun 2023 · 1 repository · arXiv:2306.10891
-
Enhanced Masked Image Modeling for Analysis of Dental Panoramic Radiographs 18 Jun 2023 · 1 repository · arXiv:2306.10623
-
Gender Bias in Transformer Models: A comprehensive survey 18 Jun 2023 · 0 repositories · arXiv:2306.10530
-
Mixed-Curvature Transformers for Graph Representation Learning papersreview 18 Jun 2023 · 0 repositories
-
News Verifiers Showdown: A Comparative Performance Evaluation of ChatGPT 3.5, ChatGPT 4.0, Bing AI, and Bard in News Fact-Checking 18 Jun 2023 · 0 repositories · arXiv:2306.17176
-
Summarization from Leaderboards to Practice: Choosing A Representation Backbone and Ensuring Robustness 18 Jun 2023 · 0 repositories · arXiv:2306.10555
-
Enhancing social network hate detection using back translation and GPT-3 augmentations during training and test-time 17 Jun 2023 · 1 repository
-
Snowman: A Million-scale Chinese Commonsense Knowledge Graph Distilled from Foundation Model 17 Jun 2023 · 0 repositories · arXiv:2306.10241
-
AD-AutoGPT: An Autonomous GPT for Alzheimer's Disease Infodemiology 16 Jun 2023 · 0 repositories · arXiv:2306.10095
-
Is Self-Repair a Silver Bullet for Code Generation? 16 Jun 2023 · 1 repository · arXiv:2306.09896
-
End-to-End Vectorized HD-map Construction with Piecewise Bezier Curve 16 Jun 2023 · 1 repository · arXiv:2306.09700
-
Evaluating Superhuman Models with Consistency Checks 16 Jun 2023 · 2 repositories · arXiv:2306.09983
-
GPT4 is Slightly Helpful for Peer-Review Assistance: A Pilot Study 16 Jun 2023 · 2 repositories · arXiv:2307.05492
-
Investigating Masking-based Data Generation in Language Models 16 Jun 2023 · 0 repositories · arXiv:2307.00008
-
MultiWave: Multiresolution Deep Architectures through Wavelet Decomposition for Multivariate Time Series Prediction 16 Jun 2023 · 1 repository · arXiv:2306.10164
-
Robot Learning with Sensorimotor Pre-training 16 Jun 2023 · 0 repositories · arXiv:2306.10007
-
TSNet-SAC: Leveraging Transformers for Efficient Task Scheduling 16 Jun 2023 · 0 repositories · arXiv:2307.07445
-
Block-State Transformers 15 Jun 2023 · 0 repositories · arXiv:2306.09539
-
ChessGPT: Bridging Policy Learning and Language Modeling 15 Jun 2023 · 1 repository · arXiv:2306.09200Syntology official (archive's flag): 14 ran · 14 ran (of which 8 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 20 harvested samples)
-
Efficient Token-Guided Image-Text Retrieval with Consistent Multimodal Contrastive Training 15 Jun 2023 · 1 repository · arXiv:2306.08789Syntology 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples) · 2 pointer-only (licence)
-
Explaining Legal Concepts with Augmented Large Language Models (GPT-4) 15 Jun 2023 · 0 repositories · arXiv:2306.09525
-
Explore, Establish, Exploit: Red Teaming Language Models from Scratch 15 Jun 2023 · 3 repositories · arXiv:2306.09442Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring the MIT Mathematics and EECS Curriculum Using Large Language Models 15 Jun 2023 · 0 repositories · arXiv:2306.08997
-
Fast Training of Diffusion Models with Masked Transformers 15 Jun 2023 · 1 repository · arXiv:2306.09305Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
Recurrent Action Transformer with Memory 15 Jun 2023 · 1 repository · arXiv:2306.09459Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Relational Temporal Graph Reasoning for Dual-task Dialogue Language Understanding 15 Jun 2023 · 0 repositories · arXiv:2306.09114
-
Dissecting Multimodality in VideoQA Transformer Models by Impairing Modality Fusion 15 Jun 2023 · 0 repositories · arXiv:2306.08889
-
Seeing the Pose in the Pixels: Learning Pose-Aware Representations in Vision Transformers 15 Jun 2023 · 1 repository · arXiv:2306.09331
-
SLAMB: Accelerated Large Batch Training with Sparse Communication 15 Jun 2023 · 1 repository
-
The pop song generator: designing an online course to teach collaborative, creative AI 15 Jun 2023 · 0 repositories · arXiv:2306.10069
-
Thrilled by Your Progress! Large Language Models (GPT-4) No Longer Struggle to Pass Assessments in Higher Education Programming Courses 15 Jun 2023 · 0 repositories · arXiv:2306.10073
-
Ensembled Prediction Intervals for Causal Outcomes Under Hidden Confounding 15 Jun 2023 · 0 repositories · arXiv:2306.09520
-
DiffAug: A Diffuse-and-Denoise Augmentation for Training Robust Classifiers 15 Jun 2023 · 0 repositories · arXiv:2306.09192
-
Assessing the Effectiveness of GPT-3 in Detecting False Political Statements: A Case Study on the LIAR Dataset 14 Jun 2023 · 1 repository · arXiv:2306.08190
-
Language models are not naysayers: An analysis of language models on negation benchmarks 14 Jun 2023 · 1 repository · arXiv:2306.08189
-
M^2UNet: MetaFormer Multi-scale Upsampling Network for Polyp Segmentation 14 Jun 2023 · 0 repositories · arXiv:2306.08600
-
MCR-Data2vec 2.0: Improving Self-supervised Speech Pre-training via Model-level Consistency Regularization 14 Jun 2023 · 0 repositories · arXiv:2306.08463
-
MUBen: Benchmarking the Uncertainty of Molecular Representation Models 14 Jun 2023 · 2 repositories · arXiv:2306.10060Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Multimodal Optimal Transport-based Co-Attention Transformer with Global Structure Consistency for Survival Prediction 14 Jun 2023 · 3 repositories · arXiv:2306.08330Syntology official: harvested, nothing ran · 10 ran (of which 6 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Research on Named Entity Recognition in Improved transformer with R-Drop structure 14 Jun 2023 · 0 repositories · arXiv:2306.08315
-
The Expressive Leaky Memory Neuron: an Efficient and Expressive Phenomenological Neuron Model Can Solve Long-Horizon Tasks 14 Jun 2023 · 1 repository · arXiv:2306.16922Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Towards AGI in Computer Vision: Lessons Learned from GPT and Large Language Models 14 Jun 2023 · 0 repositories · arXiv:2306.08641
-
TSMixer: Lightweight MLP-Mixer Model for Multivariate Time Series Forecasting 14 Jun 2023 · 1 repository · arXiv:2306.09364Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Unraveling the ARC Puzzle: Mimicking Human Solutions with Object-Centric Decision Transformer 14 Jun 2023 · 0 repositories · arXiv:2306.08204
-
When to Use Efficient Self Attention? Profiling Text, Speech and Image Transformer Variants 14 Jun 2023 · 1 repository · arXiv:2306.08667
-
arXiVeri: Automatic table verification with GPT 13 Jun 2023 · 1 repository · arXiv:2306.07968
-
Better Generalization with Semantic IDs: A Case Study in Ranking for Recommendations 13 Jun 2023 · 0 repositories · arXiv:2306.08121
-
Can ChatGPT Enable ITS? The Case of Mixed Traffic Control via Reinforcement Learning 13 Jun 2023 · 1 repository · arXiv:2306.08094
-
Enhancing Social Network Hate Detection Using Back Translation and GPT-3 Augmentations During Training and Test-Time 13 Jun 2023 · 1 repository
-
FLamE: Few-shot Learning from Natural Language Explanations 13 Jun 2023 · 0 repositories · arXiv:2306.08042