Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 12
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 12 of 190: papers 1,101 to 1,200 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
AdaViT: Adaptive Vision Transformer for Flexible Pretrain and Finetune with Variable 3D Medical Image Modalities 4 Apr 2025 · 0 repositories · arXiv:2504.03589
-
Beyond Progress Measures: Theoretical Insights into the Mechanism of Grokking 4 Apr 2025 · 1 repository · arXiv:2504.03162
-
Do LLM Evaluators Prefer Themselves for a Reason? 4 Apr 2025 · 1 repository · arXiv:2504.03846
-
DP-LET: An Efficient Spatio-Temporal Network Traffic Prediction Framework 4 Apr 2025 · 0 repositories · arXiv:2504.03792
-
Dynamic Importance in Diffusion U-Net for Enhanced Image Synthesis 4 Apr 2025 · 1 repository · arXiv:2504.03471
-
Efficient Dynamic Clustering-Based Document Compression for Retrieval-Augmented-Generation 4 Apr 2025 · 1 repository · arXiv:2504.03165
-
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT 4 Apr 2025 · 0 repositories · arXiv:2504.07982
-
Multilingual Retrieval-Augmented Generation for Knowledge-Intensive Task 4 Apr 2025 · 0 repositories · arXiv:2504.03616
-
Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models 4 Apr 2025 · 0 repositories · arXiv:2504.03624
-
Practical Poisoning Attacks against Retrieval-Augmented Generation 4 Apr 2025 · 0 repositories · arXiv:2504.03957
-
Rotation Invariance in Floor Plan Digitization using Zernike Moments 4 Apr 2025 · 0 repositories · arXiv:2504.03241
-
VISTA-OCR: Towards generative and interactive end to end OCR models 4 Apr 2025 · 0 repositories · arXiv:2504.03621
-
ZFusion: An Effective Fuser of Camera and 4D Radar for 3D Object Perception in Autonomous Driving 4 Apr 2025 · 0 repositories · arXiv:2504.03438
-
A Sensorimotor Vision Transformer 3 Apr 2025 · 0 repositories · arXiv:2504.02536
-
Adapting Large Language Models for Multi-Domain Retrieval-Augmented-Generation 3 Apr 2025 · 0 repositories · arXiv:2504.02411
-
Beyond Conventional Transformers: The Medical X-ray Attention (MXA) Block for Improved Multi-Label Diagnosis Using Knowledge Distillation 3 Apr 2025 · 1 repository · arXiv:2504.02277
-
CoLa -- Learning to Interactively Collaborate with Large LMs 3 Apr 2025 · 0 repositories · arXiv:2504.02965
-
FT-Transformer: Resilient and Reliable Transformer with End-to-End Fault Tolerant Attention 3 Apr 2025 · 0 repositories · arXiv:2504.02211
-
Graphs are everywhere -- Psst! In Music Recommendation too 3 Apr 2025 · 0 repositories · arXiv:2504.02598
-
HGFormer: Topology-Aware Vision Transformer with HyperGraph Learning 3 Apr 2025 · 0 repositories · arXiv:2504.02440
-
HQViT: Hybrid Quantum Vision Transformer for Image Classification 3 Apr 2025 · 0 repositories · arXiv:2504.02730
-
HyperRAG: Enhancing Quality-Efficiency Tradeoffs in Retrieval-Augmented Generation with Reranker KV-Cache Reuse 3 Apr 2025 · 0 repositories · arXiv:2504.02921
-
LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models 3 Apr 2025 · 0 repositories · arXiv:2504.02327
-
Localized Definitions and Distributed Reasoning: A Proof-of-Concept Mechanistic Interpretability Study via Activation Patching 3 Apr 2025 · 1 repository · arXiv:2504.02976
-
On Vanishing Variance in Transformer Length Generalization 3 Apr 2025 · 0 repositories · arXiv:2504.02827
-
Semiconductor Wafer Map Defect Classification with Tiny Vision Transformers 3 Apr 2025 · 0 repositories · arXiv:2504.02494
-
Spline-based Transformers 3 Apr 2025 · 0 repositories · arXiv:2504.02797
-
Task as Context Prompting for Accurate Medical Symptom Coding Using Large Language Models 3 Apr 2025 · 1 repository · arXiv:2504.03051
-
Towards Computation- and Communication-efficient Computational Pathology 3 Apr 2025 · 0 repositories · arXiv:2504.02628
-
A Prefixed Patch Time Series Transformer for Two-Point Boundary Value Problems in Three-Body Problems 2 Apr 2025 · 0 repositories · arXiv:2504.01464
-
A thorough benchmark of automatic text classification: From traditional approaches to large language models 2 Apr 2025 · 1 repository · arXiv:2504.01930
-
Analysis of an Idealized Stochastic Polyak Method and its Application to Black-Box Model Distillation 2 Apr 2025 · 0 repositories · arXiv:2504.01898
-
Biomedical Question Answering via Multi-Level Summarization on a Local Knowledge Graph 2 Apr 2025 · 0 repositories · arXiv:2504.01309
-
CoRAG: Collaborative Retrieval-Augmented Generation 2 Apr 2025 · 0 repositories · arXiv:2504.01883
-
Decoding Covert Speech from EEG Using a Functional Areas Spatio-Temporal Transformer 2 Apr 2025 · 1 repository · arXiv:2504.03762
-
Deep Representation Learning for Unsupervised Clustering of Myocardial Fiber Trajectories in Cardiac Diffusion Tensor Imaging 2 Apr 2025 · 0 repositories · arXiv:2504.01953
-
Dual-stream Transformer-GCN Model with Contextualized Representations Learning for Monocular 3D Human Pose Estimation 2 Apr 2025 · 1 repository · arXiv:2504.01764
-
Efficient Model Selection for Time Series Forecasting via LLMs 2 Apr 2025 · 0 repositories · arXiv:2504.02119
-
From Smør-re-brød to Subwords: Training LLMs on Danish, One Morpheme at a Time 2 Apr 2025 · 1 repository · arXiv:2504.01540
-
Geometric Reasoning in the Embedding Space 2 Apr 2025 · 0 repositories · arXiv:2504.02018
-
GeoRAG: A Question-Answering Approach from a Geographical Perspective 2 Apr 2025 · 0 repositories · arXiv:2504.01458
-
GPT Adoption and the Impact of Disclosure Policies 2 Apr 2025 · 0 repositories · arXiv:2504.01566
-
LARGE: Legal Retrieval Augmented Generation Evaluation Tool 2 Apr 2025 · 1 repository · arXiv:2504.01840
-
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System 2 Apr 2025 · 0 repositories · arXiv:2504.02894
-
PiCo: Jailbreaking Multimodal Large Language Models via Pictorial Code Contextualization 2 Apr 2025 · 0 repositories · arXiv:2504.01444
-
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval 2 Apr 2025 · 0 repositories · arXiv:2504.01348
-
Prompting Medical Vision-Language Models to Mitigate Diagnosis Bias by Generating Realistic Dermoscopic Images 2 Apr 2025 · 1 repository · arXiv:2504.01838
-
Quattro: Transformer-Accelerated Iterative Linear Quadratic Regulator Framework for Fast Trajectory Optimization 2 Apr 2025 · 1 repository · arXiv:2504.01806
-
Revisiting Funnel Transformers for Modern LLM Architectures with Comprehensive Ablations in Training and Inference Configurations 2 Apr 2025 · 0 repositories · arXiv:2504.02877
-
Scaling Test-Time Inference with Policy-Optimized, Dynamic Retrieval-Augmented Generation via KV Caching and Decoding 2 Apr 2025 · 0 repositories · arXiv:2504.01281
-
Strategize Globally, Adapt Locally: A Multi-Turn Red Teaming Agent with Dual-Level Learning 2 Apr 2025 · 0 repositories · arXiv:2504.01278
-
Towards Interpretable Soft Prompts 2 Apr 2025 · 1 repository · arXiv:2504.02144
-
UniViTAR: Unified Vision Transformer with Native Resolution 2 Apr 2025 · 0 repositories · arXiv:2504.01792
-
A Unified Virtual Mixture-of-Experts Framework:Enhanced Inference and Hallucination Mitigation in Single-Model System 1 Apr 2025 · 0 repositories · arXiv:2504.03739
-
Accelerating Causal Network Discovery of Alzheimer Disease Biomarkers via Scientific Literature-based Retrieval Augmented Generation 1 Apr 2025 · 0 repositories · arXiv:2504.08768
-
Automated Factual Benchmarking for In-Car Conversational Systems using Large Language Models 1 Apr 2025 · 0 repositories · arXiv:2504.01248
-
CellVTA: Enhancing Vision Foundation Models for Accurate Cell Segmentation and Classification 1 Apr 2025 · 1 repository · arXiv:2504.00784
-
Collaborative LLM Numerical Reasoning with Local Data Protection 1 Apr 2025 · 0 repositories · arXiv:2504.00299
-
Detecting Financial Fraud with Hybrid Deep Learning: A Mix-of-Experts Approach to Sequential and Anomalous Patterns 1 Apr 2025 · 0 repositories · arXiv:2504.03750
-
Grade Guard: A Smart System for Short Answer Automated Grading 1 Apr 2025 · 0 repositories · arXiv:2504.01253
-
LLM-Assisted Proactive Threat Intelligence for Automated Reasoning 1 Apr 2025 · 0 repositories · arXiv:2504.00428
-
Multi-Token Attention 1 Apr 2025 · 0 repositories · arXiv:2504.00927
-
QSViT: A Methodology for Quantizing Spiking Vision Transformers 1 Apr 2025 · 0 repositories · arXiv:2504.00948
-
SRLCG: Self-Rectified Large-Scale Code Generation with Multidimensional Chain-of-Thought and Dynamic Backtracking 1 Apr 2025 · 0 repositories · arXiv:2504.00532
-
WikiVideo: Article Generation from Multiple Videos 1 Apr 2025 · 1 repository · arXiv:2504.00939
-
A Systematic Evaluation of LLM Strategies for Mental Health Text Analysis: Fine-tuning vs. Prompt Engineering vs. RAG 31 Mar 2025 · 0 repositories · arXiv:2503.24307
-
Accelerating High-Efficiency Organic Photovoltaic Discovery via Pretrained Graph Neural Networks and Generative Reinforcement Learning 31 Mar 2025 · 0 repositories · arXiv:2503.23766
-
Adaptive Layer-skipping in Pre-trained LLMs 31 Mar 2025 · 0 repositories · arXiv:2503.23798
-
Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement 31 Mar 2025 · 1 repository · arXiv:2503.23895Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CITRAS: Covariate-Informed Transformer for Time Series Forecasting 31 Mar 2025 · 0 repositories · arXiv:2503.24007
-
Coarse-to-Fine Learning for Multi-Pipette Localisation in Robot-Assisted In Vivo Patch-Clamp 31 Mar 2025 · 0 repositories · arXiv:2504.01044
-
Comparing representations of long clinical texts for the task of patient note-identification 31 Mar 2025 · 0 repositories · arXiv:2503.24006
-
CrossFormer: Cross-Segment Semantic Fusion for Document Segmentation 31 Mar 2025 · 0 repositories · arXiv:2503.23671
-
Does "Reasoning" with Large Language Models Improve Recognizing, Generating, and Reframing Unhelpful Thoughts? 31 Mar 2025 · 0 repositories · arXiv:2504.00163
-
Easi3R: Estimating Disentangled Motion from DUSt3R Without Training 31 Mar 2025 · 1 repository · arXiv:2503.24391
-
Enhancing Large Language Models (LLMs) for Telecommunications using Knowledge Graphs and Retrieval-Augmented Generation 31 Mar 2025 · 0 repositories · arXiv:2503.24245
-
Graph Transformer-Based Flood Susceptibility Mapping: Application to the French Riviera and Railway Infrastructure Under Climate Change 31 Mar 2025 · 0 repositories · arXiv:2504.03727
-
JudgeLRM: Large Reasoning Models as a Judge 31 Mar 2025 · 0 repositories · arXiv:2504.00050
-
Large Language Models Pass the Turing Test 31 Mar 2025 · 0 repositories · arXiv:2503.23674
-
LLM4FS: Leveraging Large Language Models for Feature Selection and How to Improve It 31 Mar 2025 · 0 repositories · arXiv:2503.24157
-
NeuRaLaTeX: A machine learning library written in pure LaTeX 31 Mar 2025 · 0 repositories · arXiv:2503.24187
-
Rubric Is All You Need: Enhancing LLM-based Code Evaluation With Question-Specific Rubrics 31 Mar 2025 · 0 repositories · arXiv:2503.23989
-
Text Chunking for Document Classification for Urban System Management using Large Language Models 31 Mar 2025 · 1 repository · arXiv:2504.00274
-
TransMamba: Flexibly Switching between Transformer and Mamba 31 Mar 2025 · 0 repositories · arXiv:2503.24067
-
UltraRAG: A Modular and Automated Toolkit for Adaptive Retrieval-Augmented Generation 31 Mar 2025 · 1 repository · arXiv:2504.08761
-
Advancing Sentiment Analysis in Tamil-English Code-Mixed Texts: Challenges and Transformer-Based Solutions 30 Mar 2025 · 0 repositories · arXiv:2503.23295
-
Beyond Detection: Designing AI-Resilient Assessments with Automated Feedback Tool to Foster Critical Thinking 30 Mar 2025 · 0 repositories · arXiv:2503.23622
-
CADFormer: Fine-Grained Cross-modal Alignment and Decoding Transformer for Referring Remote Sensing Image Segmentation 30 Mar 2025 · 0 repositories · arXiv:2503.23456
-
Exploring GPT-4 for Robotic Agent Strategy with Real-Time State Feedback and a Reactive Behaviour Framework 30 Mar 2025 · 0 repositories · arXiv:2503.23601
-
FeRG-LLM : Feature Engineering by Reason Generation Large Language Models 30 Mar 2025 · 0 repositories · arXiv:2503.23371
-
HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose Estimation 30 Mar 2025 · 0 repositories · arXiv:2503.23331
-
Hyper-RAG: Combating LLM Hallucinations using Hypergraph-Driven Retrieval-Augmented Generation 30 Mar 2025 · 0 repositories · arXiv:2504.08758
-
JavisDiT: Joint Audio-Video Diffusion Transformer with Hierarchical Spatio-Temporal Prior Synchronization 30 Mar 2025 · 0 repositories · arXiv:2503.23377
-
Large Language Models Are Better Logical Fallacy Reasoners with Counterargument, Explanation, and Goal-Aware Prompt Formulation 30 Mar 2025 · 1 repository · arXiv:2503.23363
-
LaViC: Adapting Large Vision-Language Models to Visually-Aware Conversational Recommendation 30 Mar 2025 · 1 repository · arXiv:2503.23312
-
Object Isolated Attention for Consistent Story Visualization 30 Mar 2025 · 0 repositories · arXiv:2503.23353
-
RARE: Retrieval-Augmented Reasoning Modeling 30 Mar 2025 · 1 repository · arXiv:2503.23513Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples)
-
SCORE: Story Coherence and Retrieval Enhancement for AI Narratives 30 Mar 2025 · 0 repositories · arXiv:2503.23512
-
A Training-free LLM Framework with Interaction between Contextually Related Subtasks in Solving Complex Tasks 29 Mar 2025 · 0 repositories · arXiv:2503.23053
-
Enhancing Knowledge Graph Completion with Entity Neighborhood and Relation Context 29 Mar 2025 · 0 repositories · arXiv:2503.23205