Methods › Natural Language Processing › Subword Segmentation › BPE › Papers, page 102
Byte Pair Encoding
BPE
Papers archive 2025-07-28
archive papers tagged: 18,975 · with a code link: 8,675 · where Syntology ran a sample: 2,895 (2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,895 of 18,975 tagged: 2,443 with a run with no instrument failure, 452 where every run was a failure of Syntology's instrument)
Page 102 of 190: papers 10,101 to 10,200 of 18,975, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Toward Re-Identifying Any Animal 21 Sep 2023 · 0 repositories
-
Unified 3D Segmenter As Prototypical Classifiers 21 Sep 2023 · 1 repository
-
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11674Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Automatic Bat Call Classification using Transformer Networks 20 Sep 2023 · 0 repositories · arXiv:2309.11218
-
Controlled Generation with Prompt Insertion for Natural Language Explanations in Grammatical Error Correction 20 Sep 2023 · 1 repository · arXiv:2309.11439
-
Design of Chain-of-Thought in Math Problem Solving 20 Sep 2023 · 1 repository · arXiv:2309.11054
-
Embed-Search-Align: DNA Sequence Alignment using Transformer Models 20 Sep 2023 · 0 repositories · arXiv:2309.11087
-
Fictional Worlds, Real Connections: Developing Community Storytelling Social Chatbots through LLMs 20 Sep 2023 · 0 repositories · arXiv:2309.11478
-
Generalized Face Forgery Detection via Adaptive Learning for Pre-trained Vision Transformer 20 Sep 2023 · 1 repository · arXiv:2309.11092
-
Generative AI in Mafia-like Game Simulation 20 Sep 2023 · 0 repositories · arXiv:2309.11672
-
Generative Pre-Training of Time-Series Data for Unsupervised Fault Detection in Semiconductor Manufacturing 20 Sep 2023 · 0 repositories · arXiv:2309.11427
-
GPT-MolBERTa: GPT Molecular Features Language Model for molecular property prediction 20 Sep 2023 · 0 repositories · arXiv:2310.03030
-
Is GPT4 a Good Trader? 20 Sep 2023 · 0 repositories · arXiv:2309.10982
-
KOSMOS-2.5: A Multimodal Literate Model 20 Sep 2023 · 0 repositories · arXiv:2309.11419
-
Localize, Retrieve and Fuse: A Generalized Framework for Free-Form Question Answering over Tables 20 Sep 2023 · 0 repositories · arXiv:2309.11049
-
Multi-image Tranformer for Multi-focus Image Fusion 20 Sep 2023 · 1 repository
-
PRAT: PRofiling Adversarial aTtacks 20 Sep 2023 · 0 repositories · arXiv:2309.11111
-
Rating Prediction in Conversational Task Assistants with Behavioral and Conversational-Flow Features 20 Sep 2023 · 1 repository · arXiv:2309.11307
-
RMT: Retentive Networks Meet Vision Transformers 20 Sep 2023 · 1 repository · arXiv:2309.11523Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Safurai 001: New Qualitative Approach for Code LLM Evaluation 20 Sep 2023 · 1 repository · arXiv:2309.11385
-
Sequence-to-Sequence Spanish Pre-trained Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11259
-
SkeleTR: Towrads Skeleton-based Action Recognition in the Wild 20 Sep 2023 · 0 repositories · arXiv:2309.11445
-
The Languini Kitchen: Enabling Language Modelling Research at Different Scales of Compute 20 Sep 2023 · 1 repository · arXiv:2309.11197
-
Transformers versus LSTMs for electronic trading 20 Sep 2023 · 1 repository · arXiv:2309.11400
-
A Family of Pretrained Transformer Language Models for Russian 19 Sep 2023 · 0 repositories · arXiv:2309.10931
-
An Evaluation of GPT-4 on the ETHICS Dataset 19 Sep 2023 · 0 repositories · arXiv:2309.10492
-
Audio signal based danger detection using signal processing and deep learning 19 Sep 2023 · 1 repository
-
CFGPT: Chinese Financial Assistant with Large Language Model 19 Sep 2023 · 1 repository · arXiv:2309.10654
-
Context-Aware Neural Video Compression on Solar Dynamics Observatory 19 Sep 2023 · 0 repositories · arXiv:2309.10784
-
Exploring Iterative Enhancement for Improving Learnersourced Multiple-Choice Question Explanations with Large Language Models 19 Sep 2023 · 1 repository · arXiv:2309.10444Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
FoleyGen: Visually-Guided Audio Generation 19 Sep 2023 · 0 repositories · arXiv:2309.10537
-
Generative AI vs. AGI: The Cognitive Strengths and Weaknesses of Modern LLMs 19 Sep 2023 · 0 repositories · arXiv:2309.10371
-
Interpret Vision Transformers as ConvNets with Dynamic Convolutions 19 Sep 2023 · 0 repositories · arXiv:2309.10713
-
Language as the Medium: Multimodal Video Classification through text only 19 Sep 2023 · 0 repositories · arXiv:2309.10783
-
Learning Dynamic MRI Reconstruction with Convolutional Network Assisted Reconstruction Swin Transformer 19 Sep 2023 · 0 repositories · arXiv:2309.10227
-
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition 19 Sep 2023 · 0 repositories · arXiv:2309.10294
-
LineMarkNet: Line Landmark Detection for Valet Parking 19 Sep 2023 · 0 repositories · arXiv:2309.10475
-
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback 19 Sep 2023 · 1 repository · arXiv:2309.10691
-
PolicyGPT: Automated Analysis of Privacy Policies with Large Language Models 19 Sep 2023 · 0 repositories · arXiv:2309.10238
-
Rigorously Assessing Natural Language Explanations of Neurons 19 Sep 2023 · 0 repositories · arXiv:2309.10312
-
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing 19 Sep 2023 · 0 repositories · arXiv:2309.10356
-
Writer-Defined AI Personas for On-Demand Feedback Generation 19 Sep 2023 · 0 repositories · arXiv:2309.10433
-
Deep Prompt Tuning for Graph Transformers 18 Sep 2023 · 0 repositories · arXiv:2309.10131
-
Discovering Sounding Objects by Audio Queries for Audio Visual Segmentation 18 Sep 2023 · 0 repositories · arXiv:2309.09501
-
Distilling HuBERT with LSTMs via Decoupled Knowledge Distillation 18 Sep 2023 · 0 repositories · arXiv:2309.09920
-
EGFE: End-to-end Grouping of Fragmented Elements in UI Designs with Multimodal Learning 18 Sep 2023 · 1 repository · arXiv:2309.09867
-
Evaluation of GPT-3 for Anti-Cancer Drug Sensitivity Prediction 18 Sep 2023 · 0 repositories · arXiv:2309.10016
-
Facilitating NSFW Text Detection in Open-Domain Dialogue Systems via Knowledge Distillation 18 Sep 2023 · 1 repository · arXiv:2309.09749
-
Harnessing Collective Intelligence Under a Lack of Cultural Consensus 18 Sep 2023 · 0 repositories · arXiv:2309.09787
-
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling 18 Sep 2023 · 0 repositories · arXiv:2309.09571
-
Q-Transformer: Scalable Offline Reinforcement Learning via Autoregressive Q-Functions 18 Sep 2023 · 0 repositories · arXiv:2309.10150
-
RECAP: Retrieval-Augmented Audio Captioning 18 Sep 2023 · 1 repository · arXiv:2309.09836Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Target-aware Bi-Transformer for Few-shot Segmentation 18 Sep 2023 · 0 repositories · arXiv:2309.09492
-
Towards Ontology Construction with Language Models 18 Sep 2023 · 0 repositories · arXiv:2309.09898
-
Contrastive Decoding Improves Reasoning in Large Language Models 17 Sep 2023 · 0 repositories · arXiv:2309.09117
-
Deep Neighbor Layer Aggregation for Lightweight Self-Supervised Monocular Depth Estimation 17 Sep 2023 · 1 repository · arXiv:2309.09272
-
Do Large GPT Models Discover Moral Dimensions in Language Representations? A Topological Study Of Sentence Embeddings 17 Sep 2023 · 0 repositories · arXiv:2309.09397
-
Embrace Divergence for Richer Insights: A Multi-document Summarization Benchmark and a Case Study on Summarizing Diverse Information from News Articles 17 Sep 2023 · 1 repository · arXiv:2309.09369
-
From Cooking Recipes to Robot Task Trees -- Improving Planning Correctness and Task Efficiency by Leveraging LLMs with a Knowledge Network 17 Sep 2023 · 0 repositories · arXiv:2309.09181
-
MVP: Meta Visual Prompt Tuning for Few-Shot Remote Sensing Image Scene Classification 17 Sep 2023 · 0 repositories · arXiv:2309.09276
-
Performance of the Pre-Trained Large Language Model GPT-4 on Automated Short Answer Grading 17 Sep 2023 · 0 repositories · arXiv:2309.09338
-
Empowering In-Browser Deep Learning Inference on Edge Devices with Just-in-Time Kernel Optimizations 16 Sep 2023 · 0 repositories · arXiv:2309.08978
-
Decoder-only Architecture for Speech Recognition with CTC Prompts and Text Data Augmentation 16 Sep 2023 · 0 repositories · arXiv:2309.08876
-
Examining the Influence of Varied Levels of Domain Knowledge Base Inclusion in GPT-based Intelligent Tutors 16 Sep 2023 · 1 repository · arXiv:2309.12367
-
MMST-ViT: Climate Change-aware Crop Yield Prediction via Multi-Modal Spatial-Temporal Vision Transformer 16 Sep 2023 · 1 repository · arXiv:2309.09067Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
RingMo-lite: A Remote Sensing Multi-task Lightweight Network with CNN-Transformer Hybrid Framework 16 Sep 2023 · 0 repositories · arXiv:2309.09003
-
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 16 Sep 2023 · 1 repository · arXiv:2309.08963
-
A Modern Turkish Poet: Fine-Tuned GPT-2 15 Sep 2023 · 1 repository
-
Advancing the Evaluation of Traditional Chinese Language Models: Towards a Comprehensive Benchmark Suite 15 Sep 2023 · 1 repository · arXiv:2309.08448
-
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models 15 Sep 2023 · 1 repository · arXiv:2309.08573
-
EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers 15 Sep 2023 · 2 repositories · arXiv:2309.08532Syntology 11 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
Cross-Modal Synthesis of Structural MRI and Functional Connectivity Networks via Conditional ViT-GANs 15 Sep 2023 · 0 repositories · arXiv:2309.08160
-
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending 15 Sep 2023 · 1 repository · arXiv:2309.08646
-
Differentiable Resolution Compression and Alignment for Efficient Video Classification and Retrieval 15 Sep 2023 · 1 repository · arXiv:2309.08167
-
GPT-Lab: Next Generation Of Optimal Chemistry Discovery By GPT Driven Robotic Lab 15 Sep 2023 · 0 repositories · arXiv:2309.16721
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
InvestLM: A Large Language Model for Investment using Financial Domain Instruction Tuning 15 Sep 2023 · 1 repository · arXiv:2309.13064Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Large Language Models for Failure Mode Classification: An Investigation 15 Sep 2023 · 1 repository · arXiv:2309.08181
-
M³Net: Multilevel, Mixed and Multistage Attention Network for Salient Object Detection 15 Sep 2023 · 1 repository · arXiv:2309.08365
-
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax? 15 Sep 2023 · 0 repositories · arXiv:2309.09992
-
SculptBot: Pre-Trained Models for 3D Deformable Object Manipulation 15 Sep 2023 · 0 repositories · arXiv:2309.08728
-
Structural Self-Supervised Objectives for Transformers 15 Sep 2023 · 1 repository · arXiv:2309.08272
-
TransMUSIC: A Transformer-Aided Subspace Method for DOA Estimation with Low-Resolution ADCs 15 Sep 2023 · 1 repository · arXiv:2309.08174
-
UniST: Towards Unifying Saliency Transformer for Video Saliency Prediction and Detection 15 Sep 2023 · 0 repositories · arXiv:2309.08220
-
An Empirical Evaluation of Prompting Strategies for Large Language Models in Zero-Shot Clinical Natural Language Processing 14 Sep 2023 · 0 repositories · arXiv:2309.08008
-
Are Large Language Model-based Evaluators the Solution to Scaling Up Multilingual Evaluation? 14 Sep 2023 · 0 repositories · arXiv:2309.07462
-
Assessing the nature of large language models: A caution against anthropocentrism 14 Sep 2023 · 0 repositories · arXiv:2309.07683
-
ChatGPT MT: Competitive for High- (but not Low-) Resource Languages 14 Sep 2023 · 2 repositories · arXiv:2309.07423Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Co-Salient Object Detection with Semantic-Level Consensus Extraction and Dispersion 14 Sep 2023 · 0 repositories · arXiv:2309.07753
-
DBLPLink: An Entity Linker for the DBLP Scholarly Knowledge Graph 14 Sep 2023 · 1 repository · arXiv:2309.07545
-
Echotune: A Modular Extractor Leveraging the Variable-Length Nature of Speech in ASR Tasks 14 Sep 2023 · 0 repositories · arXiv:2309.07765
-
Folding Attention: Memory and Power Optimization for On-Device Transformer-based Streaming Speech Recognition 14 Sep 2023 · 0 repositories · arXiv:2309.07988
-
HIGT: Hierarchical Interaction Graph-Transformer for Whole Slide Image Analysis 14 Sep 2023 · 1 repository · arXiv:2309.07400Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Learning Quasi-Static 3D Models of Markerless Deformable Linear Objects for Bimanual Robotic Manipulation 14 Sep 2023 · 1 repository · arXiv:2309.07609
-
Two Timin': Repairing Smart Contracts With A Two-Layered Approach 14 Sep 2023 · 0 repositories · arXiv:2309.07841
-
Aggregating Nearest Sharp Features via Hybrid Transformers for Video Deblurring 13 Sep 2023 · 1 repository · arXiv:2309.07054
-
Benchmarking Procedural Language Understanding for Low-Resource Languages: A Case Study on Turkish 13 Sep 2023 · 1 repository · arXiv:2309.06698
-
CCSPNet-Joint: Efficient Joint Training Method for Traffic Sign Detection Under Extreme Conditions 13 Sep 2023 · 1 repository · arXiv:2309.06902
-
Enhancing Keyphrase Generation by BART Finetuning with Splitting and Shuffling 13 Sep 2023 · 0 repositories · arXiv:2309.06726
-
Generative AI 13 Sep 2023 · 0 repositories · arXiv:2309.07930