Methods › Natural Language Processing › Subword Segmentation › WordPiece › Papers, page 68
WordPiece
Papers archive 2025-07-28
archive papers tagged: 7,063 · with a code link: 2,910 · where Syntology ran a sample: 650 (529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,063 tagged: 529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument)
Page 68 of 71: papers 6,701 to 6,800 of 7,063, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter 2 Oct 2019 · 37 repositories · arXiv:1910.01108Syntology official (archive's flag): 1 ran · 21 ran (of which 5 constructed an object rather than computing a result; 13 with no instrument failure: 3 honoured, 1 violated, 9 with no contract checked; 8 where Syntology's instrument failed) · 6 unverified (of 27 harvested samples) · 2 pointer-only (licence)
-
Exploiting BERT for End-to-End Aspect-based Sentiment Analysis 2 Oct 2019 · 1 repository · arXiv:1910.00883
-
Linking artificial and human neural representations of language 2 Oct 2019 · 1 repository · arXiv:1910.01244Syntology 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
BERT for Question Generation 1 Oct 2019 · 0 repositories
-
Specializing Word Embeddings (for Parsing) by Information Bottleneck 1 Oct 2019 · 1 repository · arXiv:1910.00163
-
VAE-PGN based Abstractive Model in Multi-stage Architecture for Text Summarization 1 Oct 2019 · 0 repositories
-
End-to-End Resume Parsing and Finding Candidates for a Job Description using BERT 30 Sep 2019 · 0 repositories · arXiv:1910.03089
-
Fake news detection using Deep Learning 29 Sep 2019 · 2 repositories · arXiv:1910.03496
-
HateMonitors: Language Agnostic Abuse Detection in Social Media 27 Sep 2019 · 1 repository · arXiv:1909.12642
-
On the use of BERT for Neural Machine Translation 27 Sep 2019 · 0 repositories · arXiv:1909.12744
-
Reweighted Proximal Pruning for Large-Scale Language Representation 27 Sep 2019 · 0 repositories · arXiv:1909.12486
-
ALBERT: A Lite BERT for Self-supervised Learning of Language Representations 26 Sep 2019 · 48 repositories · arXiv:1909.11942Syntology official (archive's flag): 6 ran · 81 ran (of which 17 constructed an object rather than computing a result; 59 with no instrument failure: 4 honoured, 0 violated, 55 with no contract checked; 22 where Syntology's instrument failed) · 45 unverified (of 126 harvested samples) · 28 pointer-only (licence)
-
Aspect and Opinion Term Extraction for Hotel Reviews using Transfer Learning and Auxiliary Labels 26 Sep 2019 · 0 repositories · arXiv:1909.11879
-
Biomedical relation extraction with pre-trained language representations and minimal task-specific architecture 26 Sep 2019 · 0 repositories · arXiv:1909.12411
-
Improving Pre-Trained Multilingual Models with Vocabulary Expansion 26 Sep 2019 · 0 repositories · arXiv:1909.12440
-
Extremely Small BERT Models from Mixed-Vocabulary Training 25 Sep 2019 · 0 repositories · arXiv:1909.11687
-
Learning to Detect Opinion Snippet for Aspect-Based Sentiment Analysis 25 Sep 2019 · 0 repositories · arXiv:1909.11297
-
Mixout: Effective Regularization to Finetune Large-scale Pretrained Language Models 25 Sep 2019 · 2 repositories · arXiv:1909.11299
-
Technical report on Conversational Question Answering 24 Sep 2019 · 0 repositories · arXiv:1909.10772
-
Understanding Semantics from Speech Through Pre-training 24 Sep 2019 · 0 repositories · arXiv:1909.10924
-
TinyBERT: Distilling BERT for Natural Language Understanding 23 Sep 2019 · 10 repositories · arXiv:1909.10351Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Does BERT Make Any Sense? Interpretable Word Sense Disambiguation with Contextualized Embeddings 23 Sep 2019 · 1 repository · arXiv:1909.10430
-
Automatic Identification and Normalisation of Physical Measurements in Scientific Literature 23 Sep 2019 · 1 repository
-
Portuguese Named Entity Recognition using BERT-CRF 23 Sep 2019 · 1 repository · arXiv:1909.10649Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
BERT Meets Chinese Word Segmentation 20 Sep 2019 · 0 repositories · arXiv:1909.09292
-
AllenNLP Interpret: A Framework for Explaining Predictions of NLP Models 19 Sep 2019 · 1 repository · arXiv:1909.09251
-
How Additional Knowledge can Improve Natural Language Commonsense Question Answering? 19 Sep 2019 · 0 repositories · arXiv:1909.08855
-
Summary Level Training of Sentence Rewriting for Abstractive Summarization 19 Sep 2019 · 0 repositories · arXiv:1909.08752
-
Enriching BERT with Knowledge Graph Embeddings for Document Classification 18 Sep 2019 · 1 repository · arXiv:1909.08402
-
Improving Natural Language Inference with a Pretrained Parser 18 Sep 2019 · 1 repository · arXiv:1909.08217
-
Language models and Automated Essay Scoring 18 Sep 2019 · 1 repository · arXiv:1909.09482
-
Using BERT for Word Sense Disambiguation 18 Sep 2019 · 0 repositories · arXiv:1909.08358
-
Do NLP Models Know Numbers? Probing Numeracy in Embeddings 17 Sep 2019 · 1 repository · arXiv:1909.07940
-
SUPP.AI: Finding Evidence for Supplement-Drug Interactions 17 Sep 2019 · 1 repository · arXiv:1909.08135
-
K-BERT: Enabling Language Representation with Knowledge Graph 17 Sep 2019 · 2 repositories · arXiv:1909.07606
-
Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism 17 Sep 2019 · 10 repositories · arXiv:1909.08053Syntology community repositories only · 12 ran (of which 2 constructed an object rather than computing a result; 7 with no instrument failure: 4 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 35 unverified (of 47 harvested samples) · 15 pointer-only (licence)
-
Simple yet Effective Bridge Reasoning for Open-Domain Multi-Hop Question Answering 17 Sep 2019 · 0 repositories · arXiv:1909.07597
-
Span-based Joint Entity and Relation Extraction with Transformer Pre-training 17 Sep 2019 · 3 repositories · arXiv:1909.07755
-
Probing Natural Language Inference Models through Semantic Fragments 16 Sep 2019 · 3 repositories · arXiv:1909.07521
-
Cross-Lingual BERT Transformation for Zero-Shot Dependency Parsing 15 Sep 2019 · 1 repository · arXiv:1909.06775
-
Tree Transformer: Integrating Tree Structures into Self-Attention 14 Sep 2019 · 3 repositories · arXiv:1909.06639Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Addressing Semantic Drift in Question Generation for Semi-Supervised Question Answering 13 Sep 2019 · 2 repositories · arXiv:1909.06356Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 2 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 19 harvested samples) · 1 pointer-only (licence)
-
Neural Correction Model for Open-Domain Named Entity Recognition 13 Sep 2019 · 1 repository · arXiv:1909.06058
-
Measuring Domain Portability and ErrorPropagation in Biomedical QA 12 Sep 2019 · 0 repositories · arXiv:1909.09704
-
Q-BERT: Hessian Based Ultra Low Precision Quantization of BERT 12 Sep 2019 · 0 repositories · arXiv:1909.05840
-
UER: An Open-Source Toolkit for Pre-training Models 12 Sep 2019 · 2 repositories · arXiv:1909.05658Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
BERTgrid: Contextualized Embedding for 2D Document Representation and Understanding 11 Sep 2019 · 2 repositories · arXiv:1909.04948
-
From English to Code-Switching: Transfer Learning with Strong Morphological Clues 11 Sep 2019 · 1 repository · arXiv:1909.05158
-
Frustratingly Easy Natural Question Answering 11 Sep 2019 · 0 repositories · arXiv:1909.05286
-
How Does BERT Answer Questions? A Layer-Wise Analysis of Transformer Representations 11 Sep 2019 · 2 repositories · arXiv:1909.04925Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
BERT-Based Arabic Social Media Author Profiling 9 Sep 2019 · 0 repositories · arXiv:1909.04181
-
Knowledge Enhanced Contextual Word Representations 9 Sep 2019 · 1 repository · arXiv:1909.04164
-
Pretrained Language Models for Sequential Sentence Classification 9 Sep 2019 · 1 repository · arXiv:1909.04054
-
Reasoning Over Semantic-Level Graph for Fact Checking 9 Sep 2019 · 0 repositories · arXiv:1909.03745
-
Span Selection Pre-training for Question Answering 9 Sep 2019 · 1 repository · arXiv:1909.04120Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Commonsense Knowledge + BERT for Level 2 Reading Comprehension Ability Test 8 Sep 2019 · 0 repositories · arXiv:1909.03415
-
Czech Text Processing with Contextual Embeddings: POS Tagging, Lemmatization, Parsing and NER 8 Sep 2019 · 0 repositories · arXiv:1909.03544
-
Entity, Relation, and Event Extraction with Contextualized Span Representations 8 Sep 2019 · 4 repositories · arXiv:1909.03546Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Multi-Task Bidirectional Transformer Representations for Irony Detection 8 Sep 2019 · 0 repositories · arXiv:1909.03526
-
Symmetric Regularization based BERT for Pair-wise Semantic Reasoning 8 Sep 2019 · 1 repository · arXiv:1909.03405
-
Transfer Learning Robustness in Multi-Class Categorization by Fine-Tuning Pre-Trained Contextualized Language Models 8 Sep 2019 · 1 repository · arXiv:1909.03564
-
A Novel Cascade Binary Tagging Framework for Relational Triple Extraction 7 Sep 2019 · 5 repositories · arXiv:1909.03227
-
RNN Architecture Learning with Sparse Regularization 6 Sep 2019 · 1 repository · arXiv:1909.03011
-
Supervised Multimodal Bitransformers for Classifying Images and Text 6 Sep 2019 · 6 repositories · arXiv:1909.02950
-
Effective Use of Transformer Networks for Entity Tracking 5 Sep 2019 · 1 repository · arXiv:1909.02635Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
In Plain Sight: Media Bias Through the Lens of Factual Reporting 5 Sep 2019 · 1 repository · arXiv:1909.02670
-
Specializing Unsupervised Pretraining Models for Word-Level Semantic Similarity 5 Sep 2019 · 1 repository · arXiv:1909.02339
-
Investigating BERT's Knowledge of Language: Five Analysis Methods with NPIs 5 Sep 2019 · 1 repository · arXiv:1909.02597Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Semantics-aware BERT for Language Understanding 5 Sep 2019 · 1 repository · arXiv:1909.02209Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 3 pointer-only (licence)
-
Syntax-Aware Aspect Level Sentiment Classification with Graph Attention Networks 5 Sep 2019 · 0 repositories · arXiv:1909.02606
-
Encode, Tag, Realize: High-Precision Text Editing 3 Sep 2019 · 5 repositories · arXiv:1909.01187Syntology official (archive's flag): 5 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Language Models as Knowledge Bases? 3 Sep 2019 · 1 repository · arXiv:1909.01066
-
Multimodal Deep Learning for Mental Disorders Prediction from Audio Speech Samples 3 Sep 2019 · 0 repositories · arXiv:1909.01067
-
Transfer Fine-Tuning: A BERT Case Study 3 Sep 2019 · 1 repository · arXiv:1909.00931
-
Unicoder: A Universal Language Encoder by Pre-training with Multiple Cross-lingual Tasks 3 Sep 2019 · 0 repositories · arXiv:1909.00964
-
How Contextual are Contextualized Word Representations? Comparing the Geometry of BERT, ELMo, and GPT-2 Embeddings 2 Sep 2019 · 1 repository · arXiv:1909.00512
-
SumQE: a BERT-based Summary Quality Estimation Model 2 Sep 2019 · 1 repository · arXiv:1909.00578
-
Classification Approaches to Identify Informative Tweets 1 Sep 2019 · 0 repositories
-
Cross-Lingual Machine Reading Comprehension 1 Sep 2019 · 1 repository · arXiv:1909.00361
-
Evaluating the Cross-Lingual Effectiveness of Massively Multilingual Neural Machine Translation 1 Sep 2019 · 0 repositories · arXiv:1909.00437
-
Evaluation of Stacked Embeddings for Bulgarian on the Downstream Tasks POS and NERC 1 Sep 2019 · 0 repositories
-
Evaluation of vector embedding models in clustering of text documents 1 Sep 2019 · 0 repositories
-
FriendsQA: Open-Domain Question Answering on TV Show Transcripts 1 Sep 2019 · 0 repositories
-
QuASE: Question-Answer Driven Sentence Encoding 1 Sep 2019 · 1 repository · arXiv:1909.00333
-
Multilingual Language Models for Named Entity Recognition in German and English 1 Sep 2019 · 0 repositories
-
Multilingual Probing of Deep Pre-Trained Contextual Encoders 1 Sep 2019 · 0 repositories
-
Predicting Sentiment of Polish Language Short Texts 1 Sep 2019 · 0 repositories
-
Semantic Role Labeling with Pretrained Language Models for Known and Unknown Predicates 1 Sep 2019 · 0 repositories
-
Adversarial Learning with Contextual Embeddings for Zero-resource Cross-lingual Classification and NER 31 Aug 2019 · 0 repositories · arXiv:1909.00153
-
Evaluation Benchmarks and Learning Criteria for Discourse-Aware Sentence Representations 31 Aug 2019 · 2 repositories · arXiv:1909.00142Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Knowledge Enhanced Attention for Robust Natural Language Inference 31 Aug 2019 · 0 repositories · arXiv:1909.00102
-
NEZHA: Neural Contextualized Representation for Chinese Language Understanding 31 Aug 2019 · 10 repositories · arXiv:1909.00204
-
Quantity doesn't buy quality syntax with neural language models 31 Aug 2019 · 0 repositories · arXiv:1909.00111
-
Small and Practical BERT Models for Sequence Labeling 31 Aug 2019 · 0 repositories · arXiv:1909.00100
-
Adapt or Get Left Behind: Domain Adaptation through BERT Language Model Finetuning for Aspect-Target Sentiment Classification 30 Aug 2019 · 3 repositories · arXiv:1908.11860Syntology official: harvested, nothing ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Bilingual is At Least Monolingual (BALM): A Novel Translation Algorithm that Encodes Monolingual Priors 30 Aug 2019 · 1 repository · arXiv:1909.01146
-
PAWS-X: A Cross-lingual Adversarial Dataset for Paraphrase Identification 30 Aug 2019 · 3 repositories · arXiv:1908.11828
-
Adversarial Representation Learning for Text-to-Image Matching 28 Aug 2019 · 0 repositories · arXiv:1908.10534
-
FinBERT: Financial Sentiment Analysis with Pre-trained Language Models 27 Aug 2019 · 3 repositories · arXiv:1908.10063Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks 27 Aug 2019 · 64 repositories · arXiv:1908.10084Syntology community repositories only · 33 ran (of which 9 constructed an object rather than computing a result; 30 with no instrument failure: 1 honoured, 0 violated, 29 with no contract checked; 3 where Syntology's instrument failed) · 25 unverified (of 58 harvested samples) · 11 pointer-only (licence)