Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 108
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 108 of 190: papers 10,701 to 10,800 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Performance Evaluation of Swin Vision Transformer Model using Gradient Accumulation Optimization Technique 31 Jul 2023 · 0 repositories · arXiv:2308.00197
-
SelfSeg: A Self-supervised Sub-word Segmentation Method for Neural Machine Translation 31 Jul 2023 · 0 repositories · arXiv:2307.16400
-
CLGT: A Graph Transformer for Student Performance Prediction in Collaborative Learning 30 Jul 2023 · 1 repository · arXiv:2308.02038
-
Evaluating ChatGPT and GPT-4 for Visual Programming 30 Jul 2023 · 0 repositories · arXiv:2308.02522
-
Pre-training End-to-end ASR Models with Augmented Speech Samples Queried by Text 30 Jul 2023 · 0 repositories · arXiv:2307.16332
-
ScribbleVC: Scribble-supervised Medical Image Segmentation with Vision-Class Embedding 30 Jul 2023 · 1 repository · arXiv:2307.16226
-
SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension 30 Jul 2023 · 3 repositories · arXiv:2307.16125
-
StylePrompter: All Styles Need Is Attention 30 Jul 2023 · 1 repository · arXiv:2307.16151
-
TransFusion: A Practical and Effective Transformer-based Diffusion Model for 3D Human Motion Prediction 30 Jul 2023 · 1 repository · arXiv:2307.16106
-
Video Frame Interpolation with Flow Transformer 30 Jul 2023 · 0 repositories · arXiv:2307.16144
-
HandMIM: Pose-Aware Self-Supervised Learning for 3D Hand Mesh Estimation 29 Jul 2023 · 0 repositories · arXiv:2307.16061
-
Monaural Multi-Speaker Speech Separation Using Efficient Transformer Model 29 Jul 2023 · 0 repositories · arXiv:2308.00010
-
A Critical Review of Large Language Models: Sensitivity, Bias, and the Path Toward Specialized AI 28 Jul 2023 · 0 repositories · arXiv:2307.15425
-
Beyond Reality: The Pivotal Role of Generative AI in the Metaverse 28 Jul 2023 · 0 repositories · arXiv:2308.06272
-
ChatHome: Development and Evaluation of a Domain-Specific Language Model for Home Renovation 28 Jul 2023 · 1 repository · arXiv:2307.15290Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 17 harvested samples) · 1 pointer-only (licence)
-
Differential Evolution Algorithm based Hyper-Parameters Selection of Transformer Neural Network Model for Load Forecasting 28 Jul 2023 · 1 repository · arXiv:2307.15299
-
DocDeshadower: Frequency-Aware Transformer for Document Shadow Removal 28 Jul 2023 · 0 repositories · arXiv:2307.15318
-
Med-HALT: Medical Domain Hallucination Test for Large Language Models 28 Jul 2023 · 1 repository · arXiv:2307.15343
-
MeMOTR: Long-Term Memory-Augmented Transformer for Multi-Object Tracking 28 Jul 2023 · 1 repository · arXiv:2307.15700Syntology official (archive's flag): 10 ran · 10 ran (of which 5 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
PCNN: A Lightweight Parallel Conformer Neural Network for Efficient Monaural Speech Enhancement 28 Jul 2023 · 0 repositories · arXiv:2307.15251
-
Point Clouds Are Specialized Images: A Knowledge Transfer Approach for 3D Understanding 28 Jul 2023 · 0 repositories · arXiv:2307.15569
-
Prompt Guided Transformer for Multi-Task Dense Prediction 28 Jul 2023 · 1 repository · arXiv:2307.15362
-
RSGPT: A Remote Sensing Vision Language Model and Benchmark 28 Jul 2023 · 2 repositories · arXiv:2307.15266
-
Aligned Unsupervised Pretraining of Object Detectors with Self-training 28 Jul 2023 · 0 repositories · arXiv:2307.15697
-
VeriGen: A Large Language Model for Verilog Code Generation 28 Jul 2023 · 0 repositories · arXiv:2308.00708
-
VPP: Efficient Conditional 3D Generation via Voxel-Point Progressive Representation 28 Jul 2023 · 2 repositories · arXiv:2307.16605Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 4 pointer-only (licence)
-
A Transformer-based Approach for Arabic Offline Handwritten Text Recognition 27 Jul 2023 · 0 repositories · arXiv:2307.15045
-
ARC-NLP at PAN 2023: Hierarchical Long Text Classification for Trigger Detection 27 Jul 2023 · 0 repositories · arXiv:2307.14912
-
Evaluating Generative Models for Graph-to-Text Generation 27 Jul 2023 · 1 repository · arXiv:2307.14712
-
HTNet for micro-expression recognition 27 Jul 2023 · 1 repository · arXiv:2307.14637
-
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer 27 Jul 2023 · 1 repository · arXiv:2307.14863Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Improving Aspect-Based Sentiment with End-to-End Semantic Role Labeling Model 27 Jul 2023 · 1 repository · arXiv:2307.14785
-
LLMediator: GPT-4 Assisted Online Dispute Resolution 27 Jul 2023 · 0 repositories · arXiv:2307.16732
-
MCPA: Multi-scale Cross Perceptron Attention Network for 2D Medical Image Segmentation 27 Jul 2023 · 1 repository · arXiv:2307.14588
-
Metric-Based In-context Learning: A Case Study in Text Simplification 27 Jul 2023 · 1 repository · arXiv:2307.14632
-
New Interaction Paradigm for Complex EDA Software Leveraging GPT 27 Jul 2023 · 1 repository · arXiv:2307.14740Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Scaling Session-Based Transformer Recommendations using Optimized Negative Sampling and Loss Functions 27 Jul 2023 · 1 repository · arXiv:2307.14906
-
TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer 27 Jul 2023 · 2 repositories · arXiv:2307.14995Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Self-Supervised Graph Transformer for Deepfake Detection 27 Jul 2023 · 0 repositories · arXiv:2307.15019
-
SuperCLUE: A Comprehensive Chinese Large Language Model Benchmark 27 Jul 2023 · 0 repositories · arXiv:2307.15020
-
TextManiA: Enriching Visual Feature by Text-driven Manifold Augmentation 27 Jul 2023 · 0 repositories · arXiv:2307.14611
-
VISU at WASSA 2023 Shared Task: Detecting Emotions in Reaction to News Stories Leveraging BERT and Stacked Embeddings 27 Jul 2023 · 0 repositories · arXiv:2307.15164
-
Affective Natural Language Generation of Event Descriptions through Fine-grained Appraisal Conditions 26 Jul 2023 · 0 repositories · arXiv:2307.14004
-
Are Transformers with One Layer Self-Attention Using Low-Rank Weight Matrices Universal Approximators? 26 Jul 2023 · 0 repositories · arXiv:2307.14023
-
CliniDigest: A Case Study in Large Language Model Based Large-Scale Summarization of Clinical Trial Descriptions 26 Jul 2023 · 0 repositories · arXiv:2307.14522
-
Decoding ChatGPT: A Taxonomy of Existing Research, Current Challenges, and Possible Future Directions 26 Jul 2023 · 0 repositories · arXiv:2307.14107
-
ESSAformer: Efficient Transformer for Hyperspectral Image Super-resolution 26 Jul 2023 · 1 repository · arXiv:2307.14010Syntology official (archive's flag): 7 ran · 7 ran (of which 4 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Event-based Vision for Early Prediction of Manipulation Actions 26 Jul 2023 · 1 repository · arXiv:2307.14332
-
ExeDec: Execution Decomposition for Compositional Generalization in Neural Program Synthesis 26 Jul 2023 · 0 repositories · arXiv:2307.13883Syntology 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
FinTree: Financial Dataset Pretrain Transformer Encoder for Relation Extraction 26 Jul 2023 · 0 repositories · arXiv:2307.13900
-
How User Language Affects Conflict Fatality Estimates in ChatGPT 26 Jul 2023 · 0 repositories · arXiv:2308.00072
-
Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data 26 Jul 2023 · 1 repository · arXiv:2307.14385
-
Speed Reading Tool Powered by Artificial Intelligence for Students with ADHD, Dyslexia, or Short Attention Span 26 Jul 2023 · 0 repositories · arXiv:2307.14544
-
Unveiling Security, Privacy, and Ethical Concerns of ChatGPT 26 Jul 2023 · 0 repositories · arXiv:2307.14192
-
An End-to-End Workflow using Topic Segmentation and Text Summarisation Methods for Improved Podcast Comprehension 25 Jul 2023 · 0 repositories · arXiv:2307.13394
-
ARB: Advanced Reasoning Benchmark for Large Language Models 25 Jul 2023 · 0 repositories · arXiv:2307.13692
-
CTAGE: Curvature-Based Topology-Aware Graph Embedding for Learning Molecular Representations 25 Jul 2023 · 0 repositories · arXiv:2307.13275
-
GeoTransformer: Fast and Robust Point Cloud Registration with Geometric Transformer 25 Jul 2023 · 1 repository · arXiv:2308.03768
-
GPT-3 Models are Few-Shot Financial Reasoners 25 Jul 2023 · 0 repositories · arXiv:2307.13617
-
How Can Large Language Models Help Humans in Design and Manufacturing? 25 Jul 2023 · 0 repositories · arXiv:2307.14377
-
Is GPT a Computational Model of Emotion? Detailed Analysis 25 Jul 2023 · 0 repositories · arXiv:2307.13779
-
Multi-Granularity Prediction with Learnable Fusion for Scene Text Recognition 25 Jul 2023 · 2 repositories · arXiv:2307.13244
-
Predicting Code Coverage without Execution 25 Jul 2023 · 1 repository · arXiv:2307.13383
-
Speech representation learning: Learning bidirectional encoders with single-view, multi-view, and multi-task methods 25 Jul 2023 · 0 repositories · arXiv:2308.00129
-
Towards Resolving Word Ambiguity with Word Embeddings 25 Jul 2023 · 0 repositories · arXiv:2307.13417
-
Watermarking Conditional Text Generation for AI Detection: Unveiling Challenges and a Semantic-Aware Watermark Remedy 25 Jul 2023 · 1 repository · arXiv:2307.13808
-
Word Sense Disambiguation as a Game of Neurosymbolic Darts 25 Jul 2023 · 0 repositories · arXiv:2307.16663
-
XDLM: Cross-lingual Diffusion Language Model for Machine Translation 25 Jul 2023 · 0 repositories · arXiv:2307.13560
-
A Hybrid Machine Learning Model for Classifying Gene Mutations in Cancer using LSTM, BiLSTM, CNN, GRU, and GloVe 24 Jul 2023 · 0 repositories · arXiv:2307.14361
-
An Isometric Stochastic Optimizer 24 Jul 2023 · 0 repositories · arXiv:2307.12979
-
How Does Naming Affect LLMs on Code Analysis Tasks? 24 Jul 2023 · 0 repositories · arXiv:2307.12488
-
Comparative Analysis of Drug-GPT and ChatGPT LLMs for Healthcare Insights: Evaluating Accuracy and Relevance in Patient and HCP Contexts 24 Jul 2023 · 0 repositories · arXiv:2307.16850
-
Dense Transformer based Enhanced Coding Network for Unsupervised Metal Artifact Reduction 24 Jul 2023 · 0 repositories · arXiv:2307.12717
-
Entropy Transformer Networks: A Learning Approach via Tangent Bundle Data Manifold 24 Jul 2023 · 0 repositories · arXiv:2307.12517
-
Gradient-Based Word Substitution for Obstinate Adversarial Examples Generation in Language Models 24 Jul 2023 · 0 repositories · arXiv:2307.12507
-
Is attention all you need in medical image analysis? A review 24 Jul 2023 · 0 repositories · arXiv:2307.12775
-
Performance of Large Language Models in a Computer Science Degree Program 24 Jul 2023 · 0 repositories · arXiv:2308.02432
-
The potential of LLMs for coding with low-resource and domain-specific programming languages 24 Jul 2023 · 0 repositories · arXiv:2307.13018
-
HateModerate: Testing Hate Speech Detectors against Content Moderation Policies 23 Jul 2023 · 1 repository · arXiv:2307.12418
-
Validation of a Zero-Shot Learning Natural Language Processing Tool for Data Abstraction from Unstructured Healthcare Data 23 Jul 2023 · 1 repository · arXiv:2308.00107
-
On the Effectiveness of Spectral Discriminators for Perceptual Quality Improvement 22 Jul 2023 · 1 repository · arXiv:2307.12027
-
Sparse then Prune: Toward Efficient Vision Transformers 22 Jul 2023 · 1 repository · arXiv:2307.11988
-
Two-stream Multi-level Dynamic Point Transformer for Two-person Interaction Recognition 22 Jul 2023 · 0 repositories · arXiv:2307.11973
-
AIGC Empowering Telecom Sector White Paper_chinese 21 Jul 2023 · 0 repositories · arXiv:2307.11449
-
Artificial Intelligence-Generated Terahertz Multi-Resonant Metasurfaces via Improved Transformer and CGAN Neural Networks 21 Jul 2023 · 0 repositories · arXiv:2307.11794
-
Enhancing CLIP with GPT-4: Harnessing Visual Descriptions as Prompts 21 Jul 2023 · 1 repository · arXiv:2307.11661Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Enhancing Your Trained DETRs with Box Refinement 21 Jul 2023 · 1 repository · arXiv:2307.11828Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
GPT-4 Can't Reason 21 Jul 2023 · 0 repositories · arXiv:2308.03762
-
Assessing Large Language Models' ability to predict how humans balance self-interest and the interest of others 21 Jul 2023 · 0 repositories · arXiv:2307.12776
-
Statement-based Memory for Neural Source Code Summarization 21 Jul 2023 · 1 repository · arXiv:2307.11709
-
What can a Single Attention Layer Learn? A Study Through the Random Features Lens 21 Jul 2023 · 0 repositories · arXiv:2307.11353
-
A LLM Assisted Exploitation of AI-Guardian 20 Jul 2023 · 0 repositories · arXiv:2307.15008
-
An In-Depth Evaluation of Federated Learning on Biomedical Natural Language Processing 20 Jul 2023 · 2 repositories · arXiv:2307.11254
-
AlignDet: Aligning Pre-training and Fine-tuning in Object Detection 20 Jul 2023 · 1 repository · arXiv:2307.11077Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
An Adaptive Dual-level Reinforcement Learning Approach for Optimal Trade Execution 20 Jul 2023 · 0 repositories · arXiv:2307.10649
-
Generative Language Models on Nucleotide Sequences of Human Genes 20 Jul 2023 · 1 repository · arXiv:2307.10634
-
Hybrid Feature Embedding For Automatic Building Outline Extraction 20 Jul 2023 · 0 repositories · arXiv:2307.10609
-
Instruction-following Evaluation through Verbalizer Manipulation 20 Jul 2023 · 0 repositories · arXiv:2307.10558
-
IvyGPT: InteractiVe Chinese pathwaY language model in medical domain 20 Jul 2023 · 1 repository · arXiv:2307.10512
-
L-Eval: Instituting Standardized Evaluation for Long Context Language Models 20 Jul 2023 · 3 repositories · arXiv:2307.11088Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)