Methods › Natural Language Processing › Subword Segmentation › WordPiece › Papers, page 39
WordPiece
Papers archive 2025-07-28
archive papers tagged: 7,063 · with a code link: 2,910 · where Syntology ran a sample: 650 (529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,063 tagged: 529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument)
Page 39 of 71: papers 3,801 to 3,900 of 7,063, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A Study of the Attention Abnormality in Trojaned BERTs 16 Jan 2022 · 0 repositories
-
An Exploitation of Heterogeneous Graph Neural Network for Extractive Long Document Summarization 16 Jan 2022 · 0 repositories
-
Applying SoftTriple Loss for Supervised Language Model Fine Tuning 16 Jan 2022 · 0 repositories
-
AutoAttention: Automatic Attention Head Selection Through Differentiable Pruning 16 Jan 2022 · 0 repositories
-
Bridge the Gap Between CV and NLP! A Gradient-based Textual Adversarial Attack Framework 16 Jan 2022 · 0 repositories
-
Can BERT Conduct Logical Reasoning? On the Difficulty of Learning to Reason from Data 16 Jan 2022 · 0 repositories
-
Context-Aware Prompt: Customize A Unique Prompt For Each Input 16 Jan 2022 · 0 repositories
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 16 Jan 2022 · 0 repositories
-
Divide and Conquer: Text Semantic Matching with Disentangled Keywords and Intents 16 Jan 2022 · 0 repositories
-
Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs 16 Jan 2022 · 0 repositories
-
EiCi: A New Method of Dynamic Embedding Incorporating Contextual Information in Chinese NER 16 Jan 2022 · 0 repositories
-
Event Detection via Derangement Reading Comprehension 16 Jan 2022 · 0 repositories
-
Experiments with adversarial attacks on text genres 16 Jan 2022 · 0 repositories
-
Feasibility of BERT Embeddings For Domain-Specific Knowledge Mining 16 Jan 2022 · 0 repositories
-
FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks 16 Jan 2022 · 0 repositories
-
Global Entity Disambiguation with BERT 16 Jan 2022 · 0 repositories
-
Identifying the Source of Vulnerability in Fragile Interpretations: A Case Study in Neural Text Classification 16 Jan 2022 · 0 repositories
-
IMPLI: Investigatng NLI Models' Performance on Figurative Language 16 Jan 2022 · 0 repositories
-
Improving Contextual Representation with Gloss Regularized Pre-training 16 Jan 2022 · 0 repositories
-
Investigating and Explaining Feature and Representation Learning in Translationese Classification 16 Jan 2022 · 0 repositories
-
Investigating the saliency of sentiment expressions in aspect-based sentiment analysis 16 Jan 2022 · 0 repositories
-
Language Models for Code-switch Detection of te reo Māori and English in a Low-resource Setting 16 Jan 2022 · 0 repositories
-
Magic Pyramid: Accelerating Inference with Early Exiting and Token Pruning 16 Jan 2022 · 0 repositories
-
Measuring Word-Context Biases in Lexical Semantic Datasets 16 Jan 2022 · 0 repositories
-
Minimally-Supervised Relation Induction from Pre-trained Language Model 16 Jan 2022 · 0 repositories
-
Modeling Multi-Granularity Hierarchical Features for Relation Extraction 16 Jan 2022 · 0 repositories
-
Multi-Stage Pre-Training for Math-Understanding: μ²(AL)BERT 16 Jan 2022 · 0 repositories
-
PCEE-BERT: Accelerating BERT Inference via Patient and Confident Early Exiting 16 Jan 2022 · 0 repositories
-
Probing The Linguistic Capacity of Pre-Trained Vision-Language Models 16 Jan 2022 · 0 repositories
-
Progressive Class Semantic Matching for Semi-supervised Text Classification 16 Jan 2022 · 0 repositories
-
Re2G: Retrieve, Rerank, Generate 16 Jan 2022 · 0 repositories
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 16 Jan 2022 · 0 repositories
-
Revisiting Additive Compositionality: AND, OR, and NOT Operations with Word Embeddings 16 Jan 2022 · 0 repositories
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 16 Jan 2022 · 0 repositories
-
Roof-BERT: Divide Understanding Labour and Join in Work 16 Jan 2022 · 0 repositories
-
SemAttack: Natural Textual Attacks via Different Semantic Spaces 16 Jan 2022 · 0 repositories
-
Seq-GAN-BERT:Sequence Generative Adversarial Learning for Low-resource Name Entity Recognition 16 Jan 2022 · 0 repositories
-
Simple Local Attentions Remain Competitive for Long-Context Tasks 16 Jan 2022 · 0 repositories
-
TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning 16 Jan 2022 · 0 repositories
-
Tapping BERT for Preposition Sense Disambiguation 16 Jan 2022 · 0 repositories
-
TEMPLATE: TempRel Classification Model Trained with Embedded Temporal Relation Knowledge 16 Jan 2022 · 0 repositories
-
That is a good looking car !: Visual Aspect based Sentiment Controlled Personalized Response Generation 16 Jan 2022 · 0 repositories
-
Tree Knowledge Distillation for Compressing Transformer-Based Language Models 16 Jan 2022 · 0 repositories
-
Uncovering Surprising Event Boundaries in Narratives 16 Jan 2022 · 0 repositories
-
Understand before Answer: Improve Temporal Reading Comprehension via Precise Question Understanding 16 Jan 2022 · 0 repositories
-
VEE-BERT: Accelerating BERT Inference for Named Entity Recognition via Vote Early Exiting 16 Jan 2022 · 0 repositories
-
What do tokens know about their characters and how do they know it? 16 Jan 2022 · 1 repository
-
What Role Does BERT Play in the Neural Machine Translation Encoder? 16 Jan 2022 · 0 repositories
-
Automatic Correction of Syntactic Dependency Annotation Differences 15 Jan 2022 · 0 repositories · arXiv:2201.05891
-
Automatic Lexical Simplification for Turkish 15 Jan 2022 · 0 repositories · arXiv:2201.05878
-
Machine Learning for Food Review and Recommendation 15 Jan 2022 · 0 repositories · arXiv:2201.10978
-
Polarity and Subjectivity Detection with Multitask Learning and BERT Embedding 14 Jan 2022 · 0 repositories · arXiv:2201.05363
-
Knowledge Graph Augmented Network Towards Multiview Representation Learning for Aspect-based Sentiment Analysis 13 Jan 2022 · 1 repository · arXiv:2201.04831
-
Multi-task Pre-training Language Model for Semantic Network Completion 13 Jan 2022 · 1 repository · arXiv:2201.04843
-
Towards Automated Error Analysis: Learning to Characterize Errors 13 Jan 2022 · 0 repositories · arXiv:2201.05017
-
Diagnosing BERT with Retrieval Heuristics 12 Jan 2022 · 1 repository · arXiv:2201.04458
-
Generative Adversarial Network for Text-to-Face Synthesis and Manipulation with Pretrained BERT Model 12 Jan 2022 · 0 repositories
-
PromptBERT: Improving BERT Sentence Embeddings with Prompts 12 Jan 2022 · 1 repository · arXiv:2201.04337Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; the one sample that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
A Feature Extraction based Model for Hate Speech Identification 11 Jan 2022 · 0 repositories · arXiv:2201.04227
-
Explaining Predictive Uncertainty by Looking Back at Model Explanations 11 Jan 2022 · 0 repositories · arXiv:2201.03742
-
Quantifying Robustness to Adversarial Word Substitutions 11 Jan 2022 · 0 repositories · arXiv:2201.03829
-
BERT for Sentiment Analysis: Pre-trained and Fine-Tuned Alternatives 10 Jan 2022 · 2 repositories · arXiv:2201.03382
-
Black-Box Tuning for Language-Model-as-a-Service 10 Jan 2022 · 2 repositories · arXiv:2201.03514Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Handwriting recognition and automatic scoring for descriptive answers in Japanese language tests 10 Jan 2022 · 0 repositories · arXiv:2201.03215
-
SCROLLS: Standardized CompaRison Over Long Language Sequences 10 Jan 2022 · 2 repositories · arXiv:2201.03533Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Latency Adjustable Transformer Encoder for Language Understanding 10 Jan 2022 · 0 repositories · arXiv:2201.03327
-
Self-Training Vision Language BERTs with a Unified Conditional Model 6 Jan 2022 · 0 repositories · arXiv:2201.02010
-
Formal Analysis of Art: Proxy Learning of Visual Concepts from Style Through Language Models 5 Jan 2022 · 0 repositories · arXiv:2201.01819
-
Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction 5 Jan 2022 · 2 repositories · arXiv:2201.02184
-
Comparison of biomedical relationship extraction methods and models for knowledge graph creation 5 Jan 2022 · 0 repositories · arXiv:2201.01647
-
An Adversarial Benchmark for Fake News Detection Models 3 Jan 2022 · 1 repository · arXiv:2201.00912
-
Which Student is Best? A Comprehensive Knowledge Distillation Exam for Task-Specific BERT Models 3 Jan 2022 · 0 repositories · arXiv:2201.00558
-
On Sensitivity of Deep Learning Based Text Classification Algorithms to Practical Input Perturbations 2 Jan 2022 · 0 repositories · arXiv:2201.00318
-
Continual Stereo Matching of Continuous Driving Scenes With Growing Architecture 1 Jan 2022 · 1 repository
-
Expanding Large Pre-Trained Unimodal Models With Multimodal Information Injection for Image-Text Multimodal Classification 1 Jan 2022 · 0 repositories
-
SpaceEdit: Learning a Unified Editing Space for Open-Domain Image Color Editing 1 Jan 2022 · 0 repositories
-
Clustering Vietnamese Conversations From Facebook Page To Build Training Dataset For Chatbot 31 Dec 2021 · 1 repository · arXiv:2112.15338
-
Automatic Mixed-Precision Quantization Search of BERT 30 Dec 2021 · 0 repositories · arXiv:2112.14938
-
EvoMoE: An Evolutional Mixture-of-Experts Training Framework via Dense-To-Sparse Gate 29 Dec 2021 · 2 repositories · arXiv:2112.14397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The University of Texas at Dallas HLTRI's Participation in EPIC-QA: Searching for Entailed Questions Revealing Novel Answer Nuggets 28 Dec 2021 · 0 repositories · arXiv:2112.13946
-
Contextual Sentence Analysis for the Sentiment Prediction on Financial Data 27 Dec 2021 · 0 repositories · arXiv:2112.13790
-
Event-based clinical findings extraction from radiology reports with pre-trained language model 27 Dec 2021 · 1 repository · arXiv:2112.13512
-
Mind the Gap: Cross-Lingual Information Retrieval with Hierarchical Knowledge Enhancement 27 Dec 2021 · 0 repositories · arXiv:2112.13510
-
Multi-Image Visual Question Answering 27 Dec 2021 · 1 repository · arXiv:2112.13706
-
Secondary Use of Clinical Problem List Entries for Neural Network-Based Disease Code Assignment 27 Dec 2021 · 0 repositories · arXiv:2112.13756
-
Evaluating Contextual Embeddings and their Extraction Layers for Depression Assessment 27 Dec 2021 · 0 repositories · arXiv:2112.13795
-
An Ensemble of Pre-trained Transformer Models For Imbalanced Multiclass Malware Classification 25 Dec 2021 · 1 repository · arXiv:2112.13236
-
CABACE: Injecting Character Sequence Information and Domain Knowledge for Enhanced Acronym and Long-Form Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13237
-
Deeper Clinical Document Understanding Using Relation Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13259
-
Distilling the Knowledge of Romanian BERTs Using Multiple Teachers 23 Dec 2021 · 1 repository · arXiv:2112.12650
-
Adaptive Beam Search to Enhance On-device Abstractive Summarization 22 Dec 2021 · 0 repositories · arXiv:2201.02739
-
Consistency and Coherence from Points of Contextual Similarity 22 Dec 2021 · 0 repositories · arXiv:2112.11638
-
DB-BERT: a Database Tuning Tool that "Reads the Manual" 21 Dec 2021 · 0 repositories · arXiv:2112.10925
-
Predicting Job Titles from Job Descriptions with Multi-label Text Classification 21 Dec 2021 · 1 repository · arXiv:2112.11052
-
Training dataset and dictionary sizes matter in BERT models: the case of Baltic languages 20 Dec 2021 · 0 repositories · arXiv:2112.10553
-
Data Augmentation for Mental Health Classification on Social Media 19 Dec 2021 · 0 repositories · arXiv:2112.10064
-
Leveraging Transformers for Hate Speech Detection in Conversational Code-Mixed Tweets 18 Dec 2021 · 0 repositories · arXiv:2112.09986
-
Zero-shot and Few-shot Learning with Knowledge Graphs: A Comprehensive Survey 18 Dec 2021 · 0 repositories · arXiv:2112.10006
-
Syntactic-GCN Bert based Chinese Event Extraction 18 Dec 2021 · 0 repositories · arXiv:2112.09939
-
A High-Precision Health-relatedness Score for Phrases to Mine Cause-Effect Statements from the Web 17 Dec 2021 · 0 repositories