Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 33
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 33 of 190: papers 3,201 to 3,300 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
FinDVer: Explainable Claim Verification over Long and Hybrid-Content Financial Documents 8 Nov 2024 · 1 repository · arXiv:2411.05764
-
GCI-ViTAL: Gradual Confidence Improvement with Vision Transformers for Active Learning on Label Noise 8 Nov 2024 · 0 repositories · arXiv:2411.05939
-
GPT Semantic Cache: Reducing LLM Costs and Latency via Semantic Embedding Caching 8 Nov 2024 · 0 repositories · arXiv:2411.05276
-
Image inpainting enhancement by replacing the original mask with a self-attended region from the input image 8 Nov 2024 · 0 repositories · arXiv:2411.05705
-
Improving Molecular Graph Generation with Flow Matching and Optimal Transport 8 Nov 2024 · 0 repositories · arXiv:2411.05676
-
IntellBot: Retrieval Augmented LLM Chatbot for Cyber Threat Knowledge Delivery 8 Nov 2024 · 1 repository · arXiv:2411.05442
-
LBPE: Long-token-first Tokenization to Improve Large Language Models 8 Nov 2024 · 0 repositories · arXiv:2411.05504
-
Learning the rules of peptide self-assembly through data mining with large language models 8 Nov 2024 · 1 repository · arXiv:2411.05421
-
Multi-Document Financial Question Answering using LLMs 8 Nov 2024 · 0 repositories · arXiv:2411.07264
-
NeKo: Toward Post Recognition Generative Correction Large Language Models with Task-Oriented Experts 8 Nov 2024 · 0 repositories · arXiv:2411.05945
-
Online-LoRA: Task-free Online Continual Learning via Low Rank Adaptation 8 Nov 2024 · 1 repository · arXiv:2411.05663
-
Qwen2.5-32B: Leveraging Self-Consistent Tool-Integrated Reasoning for Bengali Mathematical Olympiad Problem Solving 8 Nov 2024 · 0 repositories · arXiv:2411.05934
-
Smile upon the Face but Sadness in the Eyes: Emotion Recognition based on Facial Expressions and Eye Behaviors 8 Nov 2024 · 0 repositories · arXiv:2411.05879
-
Using Language Models to Disambiguate Lexical Choices in Translation 8 Nov 2024 · 1 repository · arXiv:2411.05781
-
ViT Enhanced Privacy-Preserving Secure Medical Data Sharing and Classification 8 Nov 2024 · 0 repositories · arXiv:2411.05901
-
Adversarial Robustness of In-Context Learning in Transformers for Linear Regression 7 Nov 2024 · 0 repositories · arXiv:2411.05189
-
DanceFusion: A Spatio-Temporal Skeleton Diffusion Transformer for Audio-Driven Dance Motion Reconstruction 7 Nov 2024 · 0 repositories · arXiv:2411.04646
-
Deploying Large Language Models With Retrieval Augmented Generation 7 Nov 2024 · 1 repository · arXiv:2411.11895
-
Enhancing classroom teaching with LLMs and RAG 7 Nov 2024 · 0 repositories · arXiv:2411.04341
-
Enhancing Low-Light Images with Kolmogorov–Arnold Networks in Transformer Attention 7 Nov 2024 · 1 repository
-
ESC-MISR: Enhancing Spatial Correlations for Multi-Image Super-Resolution in Remote Sensing 7 Nov 2024 · 0 repositories · arXiv:2411.04706
-
FineTuneBench: How well do commercial fine-tuning APIs infuse knowledge into LLMs? 7 Nov 2024 · 1 repository · arXiv:2411.05059
-
GPT-Guided Monte Carlo Tree Search for Symbolic Regression in Financial Fraud Detection 7 Nov 2024 · 0 repositories · arXiv:2411.04459
-
HourVideo: 1-Hour Video-Language Understanding 7 Nov 2024 · 1 repository · arXiv:2411.04998Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
LLM2CLIP: Powerful Language Model Unlocks Richer Visual Representation 7 Nov 2024 · 1 repository · arXiv:2411.04997
-
M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding 7 Nov 2024 · 0 repositories · arXiv:2411.04952
-
Measure-to-measure interpolation using Transformers 7 Nov 2024 · 0 repositories · arXiv:2411.04551
-
Measuring short-form factuality in large language models 7 Nov 2024 · 1 repository · arXiv:2411.04368
-
Pose2Trajectory: Using Transformers on Body Pose to Predict Tennis Player's Trajectory 7 Nov 2024 · 0 repositories · arXiv:2411.04501
-
RetrieveGPT: Merging Prompts and Mathematical Models for Enhanced Code-Mixed Information Retrieval 7 Nov 2024 · 0 repositories · arXiv:2411.04752
-
Selecting Between BERT and GPT for Text Classification in Political Science Research 7 Nov 2024 · 0 repositories · arXiv:2411.05050
-
STAND-Guard: A Small Task-Adaptive Content Moderation Model 7 Nov 2024 · 0 repositories · arXiv:2411.05214
-
A Comparative Study of Recent Large Language Models on Generating Hospital Discharge Summaries for Lung Cancer Patients 6 Nov 2024 · 0 repositories · arXiv:2411.03805
-
A Contrastive Self-Supervised Learning scheme for beat tracking amenable to few-shot learning 6 Nov 2024 · 0 repositories · arXiv:2411.04152
-
Advanced RAG Models with Graph Structures: Optimizing Complex Knowledge Reasoning and Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03572
-
Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences 6 Nov 2024 · 3 repositories · arXiv:2411.04165Syntology official (archive's flag): 27 ran · 27 ran (of which 0 constructed an object rather than computing a result; 25 with no instrument failure: 0 honoured, 1 violated, 24 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 29 harvested samples) · 1 pointer-only (licence)
-
Can Custom Models Learn In-Context? An Exploration of Hybrid Architecture Performance on In-Context Learning Tasks 6 Nov 2024 · 1 repository · arXiv:2411.03945
-
Customized Multiple Clustering via Multi-Modal Subspace Proxy Learning 6 Nov 2024 · 1 repository · arXiv:2411.03978Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Diversity Helps Jailbreak Large Language Models 6 Nov 2024 · 0 repositories · arXiv:2411.04223
-
Fine-Grained Guidance for Retrievers: Leveraging LLMs' Feedback in Retrieval-Augmented Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03957
-
From Medprompt to o1: Exploration of Run-Time Strategies for Medical Challenge Problems and Beyond 6 Nov 2024 · 0 repositories · arXiv:2411.03590
-
From Word Vectors to Multimodal Embeddings: Techniques, Applications, and Future Directions For Large Language Models 6 Nov 2024 · 0 repositories · arXiv:2411.05036
-
On-Device Emoji Classifier Trained with GPT-based Data Augmentation for a Mobile Keyboard 6 Nov 2024 · 0 repositories · arXiv:2411.05031
-
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology 6 Nov 2024 · 0 repositories · arXiv:2411.10473
-
Prion-ViT: Prions-Inspired Vision Transformers for Temperature prediction with Specklegrams 6 Nov 2024 · 0 repositories · arXiv:2411.05836
-
Prompt Engineering Using GPT for Word-Level Code-Mixed Language Identification in Low-Resource Dravidian Languages 6 Nov 2024 · 0 repositories · arXiv:2411.04025
-
RAGulator: Lightweight Out-of-Context Detectors for Grounded Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03920
-
Towards Interpreting Language Models: A Case Study in Multi-Hop Reasoning 6 Nov 2024 · 1 repository · arXiv:2411.05037Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Understanding the Effects of Human-written Paraphrases in LLM-generated Text Detection 6 Nov 2024 · 1 repository · arXiv:2411.03806
-
YouTube Comments Decoded: Leveraging LLMs for Low Resource Language Classification 6 Nov 2024 · 0 repositories · arXiv:2411.05039
-
A Mamba Foundation Model for Time Series Forecasting 5 Nov 2024 · 0 repositories · arXiv:2411.02941
-
Automatic Generation of Question Hints for Mathematics Problems using Large Language Models in Educational Technology 5 Nov 2024 · 0 repositories · arXiv:2411.03495
-
Enhanced Real-Time Threat Detection in 5G Networks: A Self-Attention RNN Autoencoder Approach for Spectral Intrusion Analysis 5 Nov 2024 · 0 repositories · arXiv:2411.03365
-
Enhancing Transformer Training Efficiency with Dynamic Dropout 5 Nov 2024 · 0 repositories · arXiv:2411.03236
-
Exploring the Benefits of Domain-Pretraining of Generative Large Language Models for Chemistry 5 Nov 2024 · 0 repositories · arXiv:2411.03542
-
Foundation AI Model for Medical Image Segmentation 5 Nov 2024 · 0 repositories · arXiv:2411.02745
-
From Pixels to Prose: Advancing Multi-Modal Language Models for Remote Sensing 5 Nov 2024 · 0 repositories · arXiv:2411.05826
-
HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems 5 Nov 2024 · 1 repository · arXiv:2411.02959
-
Kernel Approximation using Analog In-Memory Computing 5 Nov 2024 · 1 repository · arXiv:2411.03375
-
LASER: Attention with Exponential Transformation 5 Nov 2024 · 0 repositories · arXiv:2411.03493
-
Long Context RAG Performance of Large Language Models 5 Nov 2024 · 0 repositories · arXiv:2411.03538
-
Mixtures of In-Context Learners 5 Nov 2024 · 0 repositories · arXiv:2411.02830
-
Neurons for Neutrons: A Transformer Model for Computation Load Estimation on Domain-Decomposed Neutron Transport Problems 5 Nov 2024 · 0 repositories · arXiv:2411.03389
-
P-MOSS: Learned Scheduling For Indexes Over NUMA Servers Using Low-Level Hardware Statistics 5 Nov 2024 · 0 repositories · arXiv:2411.02933
-
PersianRAG: A Retrieval-Augmented Generation System for Persian Language 5 Nov 2024 · 0 repositories · arXiv:2411.02832
-
Predictor-Corrector Enhanced Transformers with Exponential Moving Average Coefficient Learning 5 Nov 2024 · 0 repositories · arXiv:2411.03042
-
Rethinking Decoders for Transformer-based Semantic Segmentation: A Compression Perspective 5 Nov 2024 · 1 repository · arXiv:2411.03033
-
TransUNext: towards a more advanced U-shaped framework for automatic vessel segmentation in the fundus image 5 Nov 2024 · 0 repositories · arXiv:2411.02724
-
Uncertainty Quantification for Clinical Outcome Predictions with (Large) Language Models 5 Nov 2024 · 0 repositories · arXiv:2411.03497
-
Receiver-Centric Generative Semantic Communications 5 Nov 2024 · 0 repositories · arXiv:2411.03127
-
VERITAS: A Unified Approach to Reliability Evaluation 5 Nov 2024 · 0 repositories · arXiv:2411.03300
-
Advancements and limitations of LLMs in replicating human color-word associations 4 Nov 2024 · 0 repositories · arXiv:2411.02116
-
Amortized Bayesian Experimental Design for Decision-Making 4 Nov 2024 · 1 repository · arXiv:2411.02064
-
Ask, and it shall be given: On the Turing completeness of prompting 4 Nov 2024 · 1 repository · arXiv:2411.01992
-
Can Language Models Enable In-Context Database? 4 Nov 2024 · 0 repositories · arXiv:2411.01807
-
Disrupting Test Development with AI Assistants 4 Nov 2024 · 0 repositories · arXiv:2411.02328
-
ElasTST: Towards Robust Varied-Horizon Forecasting with Elastic Time-Series Transformer 4 Nov 2024 · 1 repository · arXiv:2411.01842Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
Optimizing Multi-Scale Representations to Detect Effect Heterogeneity Using Earth Observation and Computer Vision: Applications to Two Anti-Poverty RCTs 4 Nov 2024 · 0 repositories · arXiv:2411.02134
-
Enhancing Risk Assessment in Transformers with Loss-at-Risk Functions 4 Nov 2024 · 0 repositories · arXiv:2411.02558
-
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast 4 Nov 2024 · 0 repositories · arXiv:2411.02318
-
Grounding Emotional Descriptions to Electrovibration Haptic Signals 4 Nov 2024 · 0 repositories · arXiv:2411.02118
-
MdEval: Massively Multilingual Code Debugging 4 Nov 2024 · 0 repositories · arXiv:2411.02310
-
RAGViz: Diagnose and Visualize Retrieval-Augmented Generation 4 Nov 2024 · 1 repository · arXiv:2411.01751
-
Regress, Don't Guess -- A Regression-like Loss on Number Tokens for Language Models 4 Nov 2024 · 1 repository · arXiv:2411.02083Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Scalable Efficient Training of Large Language Models with Low-dimensional Projected Attention 4 Nov 2024 · 1 repository · arXiv:2411.02063Syntology official (archive's flag): 5 ran · 5 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Seq-VCR: Preventing Collapse in Intermediate Transformer Representations for Enhanced Reasoning 4 Nov 2024 · 1 repository · arXiv:2411.02344
-
SIRA: Scalable Inter-frame Relation and Association for Radar Perception 4 Nov 2024 · 0 repositories · arXiv:2411.02220
-
TeleOracle: Fine-Tuned Retrieval-Augmented Generation with Long-Context Support for Network 4 Nov 2024 · 1 repository · arXiv:2411.02617
-
Towards Leveraging News Media to Support Impact Assessment of AI Technologies 4 Nov 2024 · 0 repositories · arXiv:2411.02536
-
Training Compute-Optimal Protein Language Models 4 Nov 2024 · 1 repository · arXiv:2411.02142
-
Training-free Regional Prompting for Diffusion Transformers 4 Nov 2024 · 1 repository · arXiv:2411.02395Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Wave Network: An Ultra-Small Language Model 4 Nov 2024 · 0 repositories · arXiv:2411.02674
-
xDiT: an Inference Engine for Diffusion Transformers (DiTs) with Massive Parallelism 4 Nov 2024 · 1 repository · arXiv:2411.01738
-
A Deep Dive Into Large Language Model Code Generation Mistakes: What and Why? 3 Nov 2024 · 0 repositories · arXiv:2411.01414
-
Data Extraction Attacks in Retrieval-Augmented Generation via Backdoors 3 Nov 2024 · 0 repositories · arXiv:2411.01705
-
Enhancing Glucose Level Prediction of ICU Patients through Hierarchical Modeling of Irregular Time-Series 3 Nov 2024 · 1 repository · arXiv:2411.01418
-
Enriching Tabular Data with Contextual LLM Embeddings: A Comprehensive Ablation Study for Ensemble Classifiers 3 Nov 2024 · 0 repositories · arXiv:2411.01645
-
GITSR: Graph Interaction Transformer-based Scene Representation for Multi Vehicle Collaborative Decision-making 3 Nov 2024 · 0 repositories · arXiv:2411.01608
-
GraphXForm: Graph transformer for computer-aided molecular design 3 Nov 2024 · 1 repository · arXiv:2411.01667
-
High-performance automated abstract screening with large language model ensembles 3 Nov 2024 · 0 repositories · arXiv:2411.02451