Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 91
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 91 of 190: papers 9,001 to 9,100 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
AviationGPT: A Large Language Model for the Aviation Domain 29 Nov 2023 · 0 repositories · arXiv:2311.17686
-
Betrayed by Attention: A Simple yet Effective Approach for Self-supervised Video Object Segmentation 29 Nov 2023 · 1 repository · arXiv:2311.17893Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Biomedical knowledge graph-optimized prompt generation for large language models 29 Nov 2023 · 1 repository · arXiv:2311.17330Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
M²Chat: Empowering VLM for Multimodal LLM Interleaved Text-Image Generation 29 Nov 2023 · 1 repository · arXiv:2311.17963
-
Cross-Scope Spatial-Spectral Information Aggregation for Hyperspectral Image Super-Resolution 29 Nov 2023 · 1 repository · arXiv:2311.17340
-
Focus on Query: Adversarial Mining Transformer for Few-Shot Segmentation 29 Nov 2023 · 1 repository · arXiv:2311.17626Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 7 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Generative Hierarchical Temporal Transformer for Hand Pose and Action Modeling 29 Nov 2023 · 0 repositories · arXiv:2311.17366
-
Grounding Foundation Models through Federated Transfer Learning: A General Framework 29 Nov 2023 · 0 repositories · arXiv:2311.17431
-
Improving the Robustness of Transformer-based Large Language Models with Dynamic Attention 29 Nov 2023 · 0 repositories · arXiv:2311.17400
-
Introduction to Transformers: an NLP Perspective 29 Nov 2023 · 1 repository · arXiv:2311.17633
-
LayerCollapse: Adaptive compression of neural networks 29 Nov 2023 · 0 repositories · arXiv:2311.17943
-
MM-Narrator: Narrating Long-form Videos with Multimodal In-Context Learning 29 Nov 2023 · 0 repositories · arXiv:2311.17435
-
MoMask: Generative Masked Modeling of 3D Human Motions 29 Nov 2023 · 1 repository · arXiv:2312.00063Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 5 pointer-only (licence)
-
A Graph-Based Approach for Category-Agnostic Pose Estimation 29 Nov 2023 · 2 repositories · arXiv:2311.17891Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
PViT-6D: Overclocking Vision Transformers for 6D Pose Estimation with Confidence-Level Prediction and Pose Tokens 29 Nov 2023 · 1 repository · arXiv:2311.17504
-
RACE-IT: A Reconfigurable Analog CAM-Crossbar Engine for In-Memory Transformer Acceleration 29 Nov 2023 · 0 repositories · arXiv:2312.06532
-
SigFormer: Sparse Signal-Guided Transformer for Multi-Modal Human Action Segmentation 29 Nov 2023 · 1 repository · arXiv:2311.17428
-
TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models 29 Nov 2023 · 1 repository · arXiv:2311.17667
-
TimelyGPT: Extrapolatable Transformer Pre-training for Long-term Time-Series Forecasting in Healthcare 29 Nov 2023 · 0 repositories · arXiv:2312.00817
-
Wireless Network Digital Twin for 6G: Generative AI as A Key Enabler 29 Nov 2023 · 0 repositories · arXiv:2311.17451
-
Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine 28 Nov 2023 · 2 repositories · arXiv:2311.16452Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
CharacterGLM: Customizing Chinese Conversational AI Characters with Large Language Models 28 Nov 2023 · 1 repository · arXiv:2311.16832Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ChatGPT's One-year Anniversary: Are Open-Source Large Language Models Catching up? 28 Nov 2023 · 1 repository · arXiv:2311.16989
-
COLE: A Hierarchical Generation Framework for Multi-Layered and Editable Graphic Design 28 Nov 2023 · 0 repositories · arXiv:2311.16974
-
Comparing Generative Chatbots Based on Process Requirements 28 Nov 2023 · 0 repositories · arXiv:2312.03741
-
ContextSeg: Sketch Semantic Segmentation by Querying the Context with Attention 28 Nov 2023 · 0 repositories · arXiv:2311.16682
-
DEU-Net: Dual-Encoder U-Net for Automated Skin Lesion Segmentation 28 Nov 2023 · 1 repository
-
General-Purpose vs. Domain-Adapted Large Language Models for Extraction of Structured Data from Chest Radiology Reports 28 Nov 2023 · 0 repositories · arXiv:2311.17213
-
PHG-Net: Persistent Homology Guided Medical Image Classification 28 Nov 2023 · 1 repository · arXiv:2311.17243Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Positioning Political Texts with Large Language Models by Asking and Averaging 28 Nov 2023 · 0 repositories · arXiv:2311.16639
-
SEED-Bench-2: Benchmarking Multimodal Large Language Models 28 Nov 2023 · 2 repositories · arXiv:2311.17092
-
Self-training solutions for the ICCV 2023 GeoNet Challenge 28 Nov 2023 · 1 repository · arXiv:2311.16843
-
STR-Cert: Robustness Certification for Deep Text Recognition on Deep Learning Pipelines and Vision Transformers 28 Nov 2023 · 0 repositories · arXiv:2401.05338
-
The Falcon Series of Open Language Models 28 Nov 2023 · 0 repositories · arXiv:2311.16867
-
TLControl: Trajectory and Language Control for Human Motion Synthesis 28 Nov 2023 · 0 repositories · arXiv:2311.17135
-
Aligning Non-Causal Factors for Transformer-Based Source-Free Domain Adaptation 27 Nov 2023 · 0 repositories · arXiv:2311.16294
-
BERT Goes Off-Topic: Investigating the Domain Transfer Challenge using Genre Classification 27 Nov 2023 · 1 repository · arXiv:2311.16083
-
EgoThink: Evaluating First-Person Perspective Thinking Capability of Vision-Language Models 27 Nov 2023 · 1 repository · arXiv:2311.15596
-
ChartLlama: A Multimodal LLM for Chart Understanding and Generation 27 Nov 2023 · 0 repositories · arXiv:2311.16483
-
Data Generation for Post-OCR correction of Cyrillic handwriting 27 Nov 2023 · 2 repositories · arXiv:2311.15896
-
Decoding Logic Errors: A Comparative Study on Bug Detection by Students and Large Language Models 27 Nov 2023 · 0 repositories · arXiv:2311.16017
-
EAFP-Med: An Efficient Adaptive Feature Processing Module Based on Prompts for Medical Image Detection 27 Nov 2023 · 0 repositories · arXiv:2311.15540
-
Efficient Pre-training for Localized Instruction Generation of Videos 27 Nov 2023 · 1 repository · arXiv:2311.15964
-
GPT4Vis: What Can GPT-4 Do for Zero-shot Visual Recognition? 27 Nov 2023 · 2 repositories · arXiv:2311.15732Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Instruct2Attack: Language-Guided Semantic Adversarial Attacks 27 Nov 2023 · 0 repositories · arXiv:2311.15551
-
Machine Learning-Based Jamun Leaf Disease Detection: A Comprehensive Review 27 Nov 2023 · 0 repositories · arXiv:2311.15741
-
MEDITRON-70B: Scaling Medical Pretraining for Large Language Models 27 Nov 2023 · 1 repository · arXiv:2311.16079Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples)
-
Real Customization or Just Marketing: Are Customized Versions of Chat GPT Useful? 27 Nov 2023 · 0 repositories · arXiv:2312.03728
-
SSIN: Self-Supervised Learning for Rainfall Spatial Interpolation 27 Nov 2023 · 1 repository · arXiv:2311.15530
-
Technical Report for Argoverse Challenges on 4D Occupancy Forecasting 27 Nov 2023 · 0 repositories · arXiv:2311.15660
-
Towards Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs 27 Nov 2023 · 0 repositories · arXiv:2311.15759
-
Comparative Analysis of ChatGPT, GPT-4, and Microsoft Bing Chatbots for GRE Test 26 Nov 2023 · 0 repositories · arXiv:2312.03719
-
ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy Images 26 Nov 2023 · 2 repositories · arXiv:2311.15264Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Leveraging AI-derived Data for Carbon Accounting: Information Extraction from Alternative Sources 26 Nov 2023 · 0 repositories · arXiv:2312.03722
-
Machine-Generated Text Detection using Deep Learning 26 Nov 2023 · 1 repository · arXiv:2311.15425
-
Spectro-ViT: A Vision Transformer Model for GABA-edited MRS Reconstruction Using Spectrograms 26 Nov 2023 · 0 repositories · arXiv:2311.15386
-
UHGEval: Benchmarking the Hallucination of Chinese Large Language Models via Unconstrained Generation 26 Nov 2023 · 1 repository · arXiv:2311.15296Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Ultra-Range Gesture Recognition using a Web-Camera in Human-Robot Interaction 26 Nov 2023 · 0 repositories · arXiv:2311.15361
-
AutoEval-Video: An Automatic Benchmark for Assessing Large Vision Language Models in Open-Ended Video Question Answering 25 Nov 2023 · 1 repository · arXiv:2311.14906
-
CMed-GPT: Prompt Tuning for Entity-Aware Chinese Medical Dialogue Generation 24 Nov 2023 · 0 repositories · arXiv:2311.14539
-
Data-to-Text Bilingual Generation 24 Nov 2023 · 2 repositories · arXiv:2311.14808
-
From Text to Image: Exploring GPT-4Vision's Potential in Advanced Radiological Analysis across Subspecialties 24 Nov 2023 · 0 repositories · arXiv:2311.14777
-
GPT Struct Me: Probing GPT Models on Narrative Entity Extraction 24 Nov 2023 · 1 repository · arXiv:2311.14583
-
Image Super-Resolution with Text Prompt Diffusion 24 Nov 2023 · 1 repository · arXiv:2311.14282
-
Large Language Models as Automated Aligners for benchmarking Vision-Language Models 24 Nov 2023 · 0 repositories · arXiv:2311.14580
-
LLamol: A Dynamic Multi-Conditional Generative Transformer for De Novo Molecular Design 24 Nov 2023 · 1 repository · arXiv:2311.14407
-
Machine Translation for Ge'ez Language 24 Nov 2023 · 0 repositories · arXiv:2311.14530
-
Understanding the Role of Textual Prompts in LLM for Time Series Forecasting: an Adapter View 24 Nov 2023 · 1 repository · arXiv:2311.14782
-
RSB-Pose: Robust Short-Baseline Binocular 3D Human Pose Estimation with Occlusion Handling 24 Nov 2023 · 0 repositories · arXiv:2311.14242
-
TVT: Training-Free Vision Transformer Search on Tiny Datasets 24 Nov 2023 · 0 repositories · arXiv:2311.14337
-
A Cross Attention Approach to Diagnostic Explainability using Clinical Practice Guidelines for Depression 23 Nov 2023 · 1 repository · arXiv:2311.13852
-
Cultural Bias and Cultural Alignment of Large Language Models 23 Nov 2023 · 0 repositories · arXiv:2311.14096
-
Evaluating GPT-4's Vision Capabilities on Brazilian University Admission Exams 23 Nov 2023 · 1 repository · arXiv:2311.14169
-
Deep Learning and NLP in Cryptocurrency Forecasting: Integrating Financial, Blockchain, and Social Media Data 23 Nov 2023 · 0 repositories · arXiv:2311.14759
-
FViT-Grasp: Grasping Objects With Using Fast Vision Transformers 23 Nov 2023 · 0 repositories · arXiv:2311.13986
-
Hardware Resilience Properties of Text-Guided Image Classifiers 23 Nov 2023 · 1 repository · arXiv:2311.14062Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Knowledge Distillation Based Semantic Communications For Multiple Users 23 Nov 2023 · 0 repositories · arXiv:2311.13789
-
LACFormer: Toward accurate and efficient polyp segmentation 23 Nov 2023 · 1 repository
-
Minimizing Factual Inconsistency and Hallucination in Large Language Models 23 Nov 2023 · 0 repositories · arXiv:2311.13878
-
Progressive Learning with Visual Prompt Tuning for Variable-Rate Image Compression 23 Nov 2023 · 0 repositories · arXiv:2311.13846
-
Towards Auditing Large Language Models: Improving Text-based Stereotype Detection 23 Nov 2023 · 0 repositories · arXiv:2311.14126
-
Towards Explainable Strategy Templates using NLP Transformers 23 Nov 2023 · 0 repositories · arXiv:2311.14061
-
Beat-Aligned Spectrogram-to-Sequence Generation of Rhythm-Game Charts 22 Nov 2023 · 0 repositories · arXiv:2311.13687
-
BenthIQ: a Transformer-Based Benthic Classification Model for Coral Restoration 22 Nov 2023 · 0 repositories · arXiv:2311.13661
-
Bitformer: An efficient Transformer with bitwise operation-based attention for Big Data Analytics at low-cost low-precision devices 22 Nov 2023 · 0 repositories · arXiv:2311.13502
-
Combatting Human Trafficking in the Cyberspace: A Natural Language Processing-Based Methodology to Analyze the Language in Online Advertisements 22 Nov 2023 · 0 repositories · arXiv:2311.13118
-
Comparison of pipeline, sequence-to-sequence, and GPT models for end-to-end relation extraction: experiments with the rare disease use-case 22 Nov 2023 · 1 repository · arXiv:2311.13729
-
ComPEFT: Compression for Communicating Parameter Efficient Updates via Sparsification and Quantization 22 Nov 2023 · 1 repository · arXiv:2311.13171Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Drilling Down into the Discourse Structure with LLMs for Long Document Question Answering 22 Nov 2023 · 0 repositories · arXiv:2311.13565
-
Generation of Explanations for Logic Reasoning 22 Nov 2023 · 0 repositories · arXiv:2311.13455
-
HEViTPose: High-Efficiency Vision Transformer for Human Pose Estimation 22 Nov 2023 · 1 repository · arXiv:2311.13615
-
Input Compression with Positional Consistency for Efficient Training and Inference of Transformer Neural Networks 22 Nov 2023 · 1 repository · arXiv:2312.12385
-
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning 22 Nov 2023 · 0 repositories · arXiv:2311.13721
-
PG-Video-LLaVA: Pixel Grounding Large Video-Language Models 22 Nov 2023 · 1 repository · arXiv:2311.13435Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation 22 Nov 2023 · 1 repository · arXiv:2311.13602
-
AlignedCoT: Prompting Large Language Models via Native-Speaking Demonstrations 22 Nov 2023 · 1 repository · arXiv:2311.13538
-
Surpassing GPT-4 Medical Coding with a Two-Stage Approach 22 Nov 2023 · 0 repositories · arXiv:2311.13735
-
Towards Improving Document Understanding: An Exploration on Text-Grounding via MLLMs 22 Nov 2023 · 1 repository · arXiv:2311.13194Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
@ve: A Chatbot for Latin 22 Nov 2023 · 0 repositories · arXiv:2311.14741
-
A Survey on Large Language Models for Personalized and Explainable Recommendations 21 Nov 2023 · 0 repositories · arXiv:2311.12338