Methods › Natural Language Processing › Subword Segmentation › WordPiece › Papers, page 69
WordPiece
Papers archive 2025-07-28
archive papers tagged: 7,063 · with a code link: 2,910 · where Syntology ran a sample: 650 (529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,063 tagged: 529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument)
Page 69 of 71: papers 6,801 to 6,900 of 7,063, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Attentive History Selection for Conversational Question Answering 26 Aug 2019 · 2 repositories · arXiv:1908.09456
-
Detecting Toxicity in News Articles: Application to Bulgarian 26 Aug 2019 · 1 repository · arXiv:1908.09785
-
Does BERT agree? Evaluating knowledge of structure dependence through agreement relations 26 Aug 2019 · 1 repository · arXiv:1908.09892
-
Measuring Patent Claim Generation by Span Relevancy 26 Aug 2019 · 0 repositories · arXiv:1908.09591
-
Patient Knowledge Distillation for BERT Model Compression 25 Aug 2019 · 5 repositories · arXiv:1908.09355Syntology community repositories only · 19 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 2 honoured, 1 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 8 unverified (of 27 harvested samples) · 27 pointer-only (licence)
-
BERT for Coreference Resolution: Baselines and Analysis 24 Aug 2019 · 2 repositories · arXiv:1908.09091
-
Well-Read Students Learn Better: On the Importance of Pre-training Compact Models 23 Aug 2019 · 40 repositories · arXiv:1908.08962Syntology community repositories only · 21 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 1 honoured, 0 violated, 15 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 32 harvested samples) · 4 pointer-only (licence)
-
Multi-passage BERT: A Globally Normalized BERT Model for Open-domain Question Answering 22 Aug 2019 · 0 repositories · arXiv:1908.08167
-
Revisiting Semantic Representation and Tree Search for Similar Question Retrieval 22 Aug 2019 · 1 repository · arXiv:1908.08326
-
Text Summarization with Pretrained Encoders 22 Aug 2019 · 19 repositories · arXiv:1908.08345Syntology community repositories only · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 21 harvested samples) · 5 pointer-only (licence)
-
VL-BERT: Pre-training of Generic Visual-Linguistic Representations 22 Aug 2019 · 3 repositories · arXiv:1908.08530Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Fine-tuning BERT for Joint Entity and Relation Extraction in Chinese Medical Text 21 Aug 2019 · 1 repository · arXiv:1908.07721
-
Revealing the Dark Secrets of BERT 21 Aug 2019 · 0 repositories · arXiv:1908.08593
-
Evaluating Contextualized Embeddings on 54 Languages in POS Tagging, Lemmatization and Dependency Parsing 20 Aug 2019 · 0 repositories · arXiv:1908.07448
-
GlossBERT: BERT for Word Sense Disambiguation with Gloss Knowledge 20 Aug 2019 · 3 repositories · arXiv:1908.07245
-
A Study of BERT for Non-Factoid Question-Answering under Passage Length Constraints 19 Aug 2019 · 0 repositories · arXiv:1908.06780
-
Align, Mask and Select: A Simple Method for Incorporating Commonsense Knowledge into Language Representation Models 19 Aug 2019 · 0 repositories · arXiv:1908.06725
-
Neural Architectures for Nested NER through Linearization 19 Aug 2019 · 1 repository · arXiv:1908.06926
-
EmotionX-IDEA: Emotion BERT -- an Affectional Model for Conversation 17 Aug 2019 · 1 repository · arXiv:1908.06264
-
Language Features Matter: Effective Language Representations for Vision-Language Tasks 17 Aug 2019 · 0 repositories · arXiv:1908.06327
-
BERT-Based Multi-Head Selection for Joint Entity-Relation Extraction 16 Aug 2019 · 1 repository · arXiv:1908.05908
-
CFO: A Framework for Building Production NLP Systems 16 Aug 2019 · 0 repositories · arXiv:1908.06121
-
CLUTRR: A Diagnostic Benchmark for Inductive Reasoning from Text 16 Aug 2019 · 5 repositories · arXiv:1908.06177Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Integrating Multimodal Information in Large Pretrained Transformers 15 Aug 2019 · 1 repository · arXiv:1908.05787
-
Towards Making the Most of BERT in Neural Machine Translation 15 Aug 2019 · 2 repositories · arXiv:1908.05672
-
Visualizing and Understanding the Effectiveness of BERT 15 Aug 2019 · 0 repositories · arXiv:1908.05620
-
Establishing Strong Baselines for the New Decade: Sequence Tagging, Syntactic and Semantic Parsing with BERT 14 Aug 2019 · 1 repository · arXiv:1908.04943
-
On-Device Text Representations Robust To Misspellings via Projections 14 Aug 2019 · 0 repositories · arXiv:1908.05763
-
Scalable Attentive Sentence-Pair Modeling via Distilled Sentence Embedding 14 Aug 2019 · 1 repository · arXiv:1908.05161
-
SG-Net: Syntax-Guided Machine Reading Comprehension 14 Aug 2019 · 1 repository · arXiv:1908.05147
-
BioFLAIR: Pretrained Pooled Contextualized Embeddings for Biomedical Sequence Labeling Tasks 13 Aug 2019 · 1 repository · arXiv:1908.05760
-
An Effective Domain Adaptive Post-Training Method for BERT in Response Selection 13 Aug 2019 · 1 repository · arXiv:1908.04812
-
Generative Question Refinement with Deep Reinforcement Learning in Retrieval-based QA System 13 Aug 2019 · 1 repository · arXiv:1908.05604
-
StructBERT: Incorporating Language Structures into Pre-training for Deep Language Understanding 13 Aug 2019 · 0 repositories · arXiv:1908.04577
-
TAPER: Time-Aware Patient EHR Representation 11 Aug 2019 · 2 repositories · arXiv:1908.03971Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
Multi-modality Latent Interaction Network for Visual Question Answering 10 Aug 2019 · 0 repositories · arXiv:1908.04289
-
BERT-based Ranking for Biomedical Entity Normalization 9 Aug 2019 · 0 repositories · arXiv:1908.03548
-
TinySearch -- Semantics based Search Engine using Bert Embeddings 7 Aug 2019 · 0 repositories · arXiv:1908.02451
-
Clustering of Deep Contextualized Representations for Summarization of Biomedical Texts 6 Aug 2019 · 1 repository · arXiv:1908.02286
-
Predicting Prosodic Prominence from Text with Pre-trained Contextualized Word Representations 6 Aug 2019 · 1 repository · arXiv:1908.02262
-
ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks 6 Aug 2019 · 11 repositories · arXiv:1908.02265Syntology 10 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 24 unverified (of 34 harvested samples) · 34 pointer-only (licence)
-
Beyond English-Only Reading Comprehension: Experiments in Zero-Shot Multilingual Transfer for Bulgarian 5 Aug 2019 · 1 repository · arXiv:1908.01519
-
Exploring Neural Net Augmentation to BERT for Question Answering on SQUAD 2.0 4 Aug 2019 · 0 repositories · arXiv:1908.01767
-
BERT Masked Language Modeling for Co-reference Resolution 1 Aug 2019 · 0 repositories
-
BSNLP2019 Shared Task Submission: Multisource Neural NER Transfer 1 Aug 2019 · 0 repositories
-
Cross-Lingual Lemmatization and Morphology Tagging with Two-Stage Multilingual BERT Fine-Tuning 1 Aug 2019 · 1 repository
-
Fill the GAP: Exploiting BERT for Pronoun Resolution 1 Aug 2019 · 1 repository
-
Filtering Pseudo-References by Paraphrasing for Automatic Evaluation of Machine Translation 1 Aug 2019 · 0 repositories
-
Gendered Ambiguous Pronoun (GAP) Shared Task at the Gender Bias in NLP Workshop 2019 1 Aug 2019 · 0 repositories
-
hULMonA: The Universal Language Model in Arabic 1 Aug 2019 · 1 repository
-
KFU NLP Team at SMM4H 2019 Tasks: Want to Extract Adverse Drugs Reactions from Tweets? BERT to The Rescue 1 Aug 2019 · 0 repositories
-
KU_ai at MEDIQA 2019: Domain-specific Pre-training and Transfer Learning for Medical NLI 1 Aug 2019 · 0 repositories
-
MIPT System for World-Level Quality Estimation 1 Aug 2019 · 0 repositories
-
MSnet: A BERT-based Network for Gendered Pronoun Resolution 1 Aug 2019 · 1 repository · arXiv:1908.00308
-
Multi-headed Architecture Based on BERT for Grammatical Errors Correction 1 Aug 2019 · 0 repositories
-
NCUEE at MEDIQA 2019: Medical Text Inference Using Ensemble BERT-BiLSTM-Attention Model 1 Aug 2019 · 0 repositories
-
No Army, No Navy: BERT Semi-Supervised Learning of Arabic Dialects 1 Aug 2019 · 0 repositories
-
Noisy Channel for Low Resource Grammatical Error Correction 1 Aug 2019 · 0 repositories
-
On GAP Coreference Resolution Shared Task: Insights from the 3rd Place Solution 1 Aug 2019 · 0 repositories
-
PANLP at MEDIQA 2019: Pre-trained Language Models, Transfer Learning and Knowledge Distillation 1 Aug 2019 · 0 repositories
-
QE BERT: Bilingual BERT Using Multi-task Learning for Neural Quality Estimation 1 Aug 2019 · 0 repositories
-
Quality Estimation and Translation Metrics via Pre-trained Word and Sentence Embeddings 1 Aug 2019 · 0 repositories
-
Saama Research at MEDIQA 2019: Pre-trained BioBERT with Attention Visualisation for Medical Natural Language Inference 1 Aug 2019 · 0 repositories
-
The Role of Protected Class Word Lists in Bias Identification of Contextualized Word Representations 1 Aug 2019 · 0 repositories
-
TMU Transformer System Using BERT for Re-ranking at BEA 2019 Grammatical Error Correction on Restricted Track 1 Aug 2019 · 0 repositories
-
Transfer Learning from Pre-trained BERT for Pronoun Resolution 1 Aug 2019 · 0 repositories
-
Tuning Multilingual Transformers for Language-Specific Named Entity Recognition 1 Aug 2019 · 1 repository
-
UU_TAILS at MEDIQA 2019: Learning Textual Entailment in the Medical Domain 1 Aug 2019 · 0 repositories
-
What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models 31 Jul 2019 · 2 repositories · arXiv:1907.13528Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
ERNIE 2.0: A Continual Pre-training Framework for Language Understanding 29 Jul 2019 · 3 repositories · arXiv:1907.12412Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Leveraging Pre-trained Checkpoints for Sequence Generation Tasks 29 Jul 2019 · 7 repositories · arXiv:1907.12461Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Machine Translation Evaluation with BERT Regressor 29 Jul 2019 · 0 repositories · arXiv:1907.12679
-
Neural Mention Detection 29 Jul 2019 · 1 repository · arXiv:1907.12524
-
Is BERT Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and Entailment 27 Jul 2019 · 7 repositories · arXiv:1907.11932Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Investigating Self-Attention Network for Chinese Word Segmentation 26 Jul 2019 · 2 repositories · arXiv:1907.11512
-
RoBERTa: A Robustly Optimized BERT Pretraining Approach 26 Jul 2019 · 67 repositories · arXiv:1907.11692Syntology community repositories only · 37 ran (of which 11 constructed an object rather than computing a result; 36 with no instrument failure: 0 honoured, 0 violated, 36 with no contract checked; 1 where Syntology's instrument failed) · 11 unverified (of 48 harvested samples) · 24 pointer-only (licence)
-
SpanBERT: Improving Pre-training by Representing and Predicting Spans 24 Jul 2019 · 6 repositories · arXiv:1907.10529Syntology official (archive's flag): 2 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples) · 6 pointer-only (licence)
-
Unbabel's Participation in the WMT19 Translation Quality Estimation Shared Task 24 Jul 2019 · 0 repositories · arXiv:1907.10352
-
EmotionX-HSU: Adopting Pre-trained BERT for Emotion Classification 23 Jul 2019 · 0 repositories · arXiv:1907.09669
-
GEAR: Graph-based Evidence Aggregating and Reasoning for Fact Verification 22 Jul 2019 · 2 repositories · arXiv:1908.01843
-
Generating Sentiment-Preserving Fake Online Reviews Using Neural Language Models and Their Human- and Machine-based Detection 22 Jul 2019 · 0 repositories · arXiv:1907.09177
-
Fake News Detection as Natural Language Inference 17 Jul 2019 · 1 repository · arXiv:1907.07347
-
Low-Shot Classification: A Comparison of Classical and Deep Transfer Machine Learning Approaches 17 Jul 2019 · 0 repositories · arXiv:1907.07543
-
Multi-modal Sentiment Analysis using Deep Canonical Correlation Analysis 15 Jul 2019 · 0 repositories · arXiv:1907.08696
-
Myers-Briggs Personality Classification and Personality-Specific Language Generation Using Pre-trained Language Models 15 Jul 2019 · 0 repositories · arXiv:1907.06333
-
Lexical Simplification with Pretrained Encoders 14 Jul 2019 · 3 repositories · arXiv:1907.06226
-
TWEETQA: A Social Media Focused Question Answering Dataset 14 Jul 2019 · 0 repositories · arXiv:1907.06292
-
BAM! Born-Again Multi-Task Networks for Natural Language Understanding 10 Jul 2019 · 1 repository · arXiv:1907.04829
-
Can Unconditional Language Models Recover Arbitrary Sentences? 10 Jul 2019 · 0 repositories · arXiv:1907.04944
-
Let's measure run time! Extending the IR replicability infrastructure to include performance aspects 10 Jul 2019 · 0 repositories · arXiv:1907.04614
-
To Tune or Not To Tune? How About the Best of Both Worlds? 9 Jul 2019 · 2 repositories · arXiv:1907.05338
-
Incorporating Query Term Independence Assumption for Efficient Retrieval and Ranking using Deep Neural Networks 8 Jul 2019 · 0 repositories · arXiv:1907.03693
-
Short Text Conversation Based on Deep Neural Network and Analysis on Evaluation Measures 6 Jul 2019 · 1 repository · arXiv:1907.03070
-
BERT-DST: Scalable End-to-End Dialogue State Tracking with Bidirectional Encoder Representations from Transformer 5 Jul 2019 · 1 repository · arXiv:1907.03040
-
A Simple and Effective Approach to Automatic Post-Editing with Transfer Learning 1 Jul 2019 · 1 repository
-
A Surprisingly Robust Trick for the Winograd Schema Challenge 1 Jul 2019 · 0 repositories
-
BERT-based Lexical Substitution 1 Jul 2019 · 1 repository
-
Coreference Resolution with Entity Equalization 1 Jul 2019 · 1 repository
-
DisSent: Learning Sentence Representations from Explicit Discourse Relations 1 Jul 2019 · 1 repository
-
Enhancing Pre-Trained Language Representations with Rich Knowledge for Machine Reading Comprehension 1 Jul 2019 · 1 repository