Methods › Natural Language Processing › Subword Segmentation › WordPiece › Papers, page 59
WordPiece
Papers archive 2025-07-28
archive papers tagged: 7,063 · with a code link: 2,910 · where Syntology ran a sample: 650 (529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,063 tagged: 529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument)
Page 59 of 71: papers 5,801 to 5,900 of 7,063, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
NEU at WNUT-2020 Task 2: Data Augmentation To Tell BERT That Death Is Not Necessarily Informative 18 Sep 2020 · 0 repositories · arXiv:2009.08590
-
The birth of Romanian BERT 18 Sep 2020 · 1 repository · arXiv:2009.08712
-
Will it Unblend? 18 Sep 2020 · 1 repository · arXiv:2009.09123
-
A Multimodal Memes Classification: A Survey and Open Research Issues 17 Sep 2020 · 0 repositories · arXiv:2009.08395
-
Compositional and Lexical Semantics in RoBERTa, BERT and DistilBERT: A Case Study on CoQA 17 Sep 2020 · 0 repositories · arXiv:2009.08257
-
Cross-Modal Alignment with Mixture Experts Neural Network for Intral-City Retail Recommendation 17 Sep 2020 · 0 repositories · arXiv:2009.09926
-
DSC IIT-ISM at SemEval-2020 Task 6: Boosting BERT with Dependencies for Definition Extraction 17 Sep 2020 · 1 repository · arXiv:2009.08180
-
Efficient Transformer-based Large Scale Language Representations using Hardware-friendly Block Structured Pruning 17 Sep 2020 · 0 repositories · arXiv:2009.08065
-
Multi²OIE: Multilingual Open Information Extraction Based on Multi-Head Attention with BERT 17 Sep 2020 · 1 repository · arXiv:2009.08128Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Deep Learning Approaches for Extracting Adverse Events and Indications of Dietary Supplements from Clinical Text 16 Sep 2020 · 0 repositories · arXiv:2009.07780
-
Simplified TinyBERT: Knowledge Distillation for Document Retrieval 16 Sep 2020 · 4 repositories · arXiv:2009.07531
-
Solomon at SemEval-2020 Task 11: Ensemble Architecture for Fine-Tuned Propaganda Detection in News Articles 16 Sep 2020 · 0 repositories · arXiv:2009.07473
-
UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation 16 Sep 2020 · 1 repository · arXiv:2009.07602Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Real-Time Execution of Large-scale Language Models on Mobile 15 Sep 2020 · 0 repositories · arXiv:2009.06823
-
Augmented Natural Language for Generative Sequence Labeling 15 Sep 2020 · 0 repositories · arXiv:2009.13272
-
BERT-QE: Contextualized Query Expansion for Document Re-ranking 15 Sep 2020 · 1 repository · arXiv:2009.07258Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
DeNERT-KG: Named Entity and Relation Extraction Model Using DQN, Knowledge Graph, and BERT 15 Sep 2020 · 0 repositories
-
Event Presence Prediction Helps Trigger Detection Across Languages 15 Sep 2020 · 0 repositories · arXiv:2009.07188
-
Lessons Learned from Applying off-the-shelf BERT: There is no Silver Bullet 15 Sep 2020 · 0 repositories · arXiv:2009.07238
-
MLMLM: Link Prediction with Mean Likelihood Masked Language Model 15 Sep 2020 · 0 repositories · arXiv:2009.07058
-
Beyond Accuracy: ROI-driven Data Analytics of Empirical Data 14 Sep 2020 · 0 repositories · arXiv:2009.06492
-
On Robustness and Bias Analysis of BERT-based Relation Extraction 14 Sep 2020 · 1 repository · arXiv:2009.06206
-
Efficient Transformers: A Survey 14 Sep 2020 · 0 repositories · arXiv:2009.06732
-
Filling the Gap of Utterance-aware and Speaker-aware Representation for Multi-turn Dialogue 14 Sep 2020 · 1 repository · arXiv:2009.06504Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
BoostingBERT:Integrating Multi-Class Boosting into BERT for NLP Tasks 13 Sep 2020 · 0 repositories · arXiv:2009.05959
-
CIA_NITT at WNUT-2020 Task 2: Classification of COVID-19 Tweets Using Pre-trained Language Models 12 Sep 2020 · 0 repositories · arXiv:2009.05782
-
Country Image in COVID-19 Pandemic: A Case Study of China 12 Sep 2020 · 1 repository · arXiv:2009.05817
-
Fine-tuning Pre-trained Contextual Embeddings for Citation Content Analysis in Scholarly Publication 12 Sep 2020 · 0 repositories · arXiv:2009.05836
-
A Comparison of LSTM and BERT for Small Corpus 11 Sep 2020 · 0 repositories · arXiv:2009.05451
-
Compressed Deep Networks: Goodbye SVD, Hello Robust Low-Rank Approximation 11 Sep 2020 · 1 repository · arXiv:2009.05647
-
UPB at SemEval-2020 Task 11: Propaganda Detection with Domain-Specific Trained BERT 11 Sep 2020 · 0 repositories · arXiv:2009.05289
-
UPB at SemEval-2020 Task 6: Pretrained Language Models for Definition Extraction 11 Sep 2020 · 3 repositories · arXiv:2009.05603
-
Do Response Selection Models Really Know What's Next? Utterance Manipulation Strategies for Multi-turn Response Selection 10 Sep 2020 · 1 repository · arXiv:2009.04703
-
Investigating Gender Bias in BERT 10 Sep 2020 · 0 repositories · arXiv:2009.05021
-
Learning Universal Representations from Word to Sentence 10 Sep 2020 · 0 repositories · arXiv:2009.04656
-
Modern Methods for Text Generation 10 Sep 2020 · 2 repositories · arXiv:2009.04968
-
Comparative Study of Language Models on Cross-Domain Data with Model Agnostic Explainability 9 Sep 2020 · 0 repositories · arXiv:2009.04095
-
Pay Attention when Required 9 Sep 2020 · 2 repositories · arXiv:2009.04534
-
ERNIE at SemEval-2020 Task 10: Learning Word Emphasis Selection by Pre-trained Language Model 8 Sep 2020 · 0 repositories · arXiv:2009.03706
-
E-BERT: A Phrase and Product Knowledge Enhanced Language Model for E-commerce 7 Sep 2020 · 0 repositories · arXiv:2009.02835
-
EdinburghNLP at WNUT-2020 Task 2: Leveraging Transformers with Generalized Augmentation for Identifying Informativeness in COVID-19 Tweets 6 Sep 2020 · 0 repositories · arXiv:2009.06375
-
QiaoNing at SemEval-2020 Task 4: Commonsense Validation and Explanation system based on ensemble of language model 6 Sep 2020 · 0 repositories · arXiv:2009.02645
-
UPB at SemEval-2020 Task 8: Joint Textual and Visual Modeling in a Multi-Task Learning Architecture for Memotion Analysis 6 Sep 2020 · 0 repositories · arXiv:2009.02779
-
Accenture at CheckThat! 2020: If you say so: Post-hoc fact-checking of claims using transformer-based models 5 Sep 2020 · 0 repositories · arXiv:2009.02431
-
Comparative Evaluation of Pretrained Transfer Learning Models on Automatic Short Answer Grading 2 Sep 2020 · 1 repository · arXiv:2009.01303
-
Automatic Assignment of Radiology Examination Protocols Using Pre-trained Language Models with Knowledge Distillation 1 Sep 2020 · 1 repository · arXiv:2009.00694
-
Sentimental LIAR: Extended Corpus and Deep Learning Models for Fake Claim Classification 1 Sep 2020 · 2 repositories · arXiv:2009.01047
-
A Bidirectional Tree Tagging Scheme for Joint Medical Relation Extraction 31 Aug 2020 · 0 repositories · arXiv:2008.13339
-
SocCogCom at SemEval-2020 Task 11: Characterizing and Detecting Propaganda using Sentence-Level Emotional Salience Features 29 Aug 2020 · 1 repository · arXiv:2008.13012
-
Knowledge Efficient Deep Learning for Natural Language Processing 28 Aug 2020 · 0 repositories · arXiv:2008.12878
-
Rethinking the Objectives of Extractive Question Answering 28 Aug 2020 · 1 repository · arXiv:2008.12804Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Fast and Robust BERT-based Dialogue State Tracker for Schema-Guided Dialogue Dataset 27 Aug 2020 · 1 repository · arXiv:2008.12335
-
AMBERT: A Pre-trained Language Model with Multi-Grained Tokenization 27 Aug 2020 · 0 repositories · arXiv:2008.11869
-
Entity and Evidence Guided Relation Extraction for DocRED 27 Aug 2020 · 0 repositories · arXiv:2008.12283
-
GREEK-BERT: The Greeks visiting Sesame Street 27 Aug 2020 · 1 repository · arXiv:2008.12014Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
MultiGBS: A multi-layer graph approach to biomedical summarization 27 Aug 2020 · 0 repositories · arXiv:2008.11908
-
Query Focused Multi-document Summarisation of Biomedical Texts 27 Aug 2020 · 1 repository · arXiv:2008.11986
-
Query Focused Multi-document Summarisation of Biomedical Texts: Macquarie Universiy and the Australian National University at BioASQ8b 27 Aug 2020 · 1 repository
-
APMSqueeze: A Communication Efficient Adam-Preconditioned Momentum SGD Algorithm 26 Aug 2020 · 0 repositories · arXiv:2008.11343
-
Analysis and Evaluation of Language Models for Word Sense Disambiguation 26 Aug 2020 · 1 repository · arXiv:2008.11608
-
Conceptualized Representation Learning for Chinese Biomedical Text Mining 25 Aug 2020 · 0 repositories · arXiv:2008.10813
-
ETC-NLG: End-to-end Topic-Conditioned Natural Language Generation 25 Aug 2020 · 1 repository · arXiv:2008.10875
-
Knowledge-Empowered Representation Learning for Chinese Medical Reading Comprehension: Task, Model and Resources 24 Aug 2020 · 1 repository · arXiv:2008.10327
-
Prediction of ICD Codes with Clinical BERT Embeddings and Text Augmentation with Label Balancing using MIMIC-III 24 Aug 2020 · 0 repositories · arXiv:2008.10492
-
syrapropa at SemEval-2020 Task 11: BERT-based Models Design For Propagandistic Technique and Span Detection 24 Aug 2020 · 0 repositories · arXiv:2008.10163
-
Two Stages Approach for Tweet Engagement Prediction 24 Aug 2020 · 0 repositories · arXiv:2008.10419
-
YNU-HPCC at SemEval-2020 Task 11: LSTM Network for Detection of Propaganda Techniques in News Articles 24 Aug 2020 · 1 repository · arXiv:2008.10166Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Applications of BERT Based Sequence Tagging Models on Chinese Medical Text Attributes Extraction 22 Aug 2020 · 0 repositories · arXiv:2008.09740
-
CyberWallE at SemEval-2020 Task 11: An Analysis of Feature Engineering for Ensemble Models for Propaganda Detection 22 Aug 2020 · 1 repository · arXiv:2008.09859
-
DUTH at SemEval-2020 Task 11: BERT with Entity Mapping for Propaganda Classification 22 Aug 2020 · 1 repository · arXiv:2008.09894
-
FAT ALBERT: Finding Answers in Large Texts using Semantic Similarity Attention Layer based on BERT 22 Aug 2020 · 1 repository · arXiv:2009.01004
-
HinglishNLP: Fine-tuned Language Models for Hinglish Sentiment Detection 22 Aug 2020 · 2 repositories · arXiv:2008.09820
-
Abstractive Summarization of Spoken andWritten Instructions with BERT 21 Aug 2020 · 2 repositories
-
Adapting Event Extractors to Medical Data: Bridging the Covariate Shift 21 Aug 2020 · 0 repositories · arXiv:2008.09266
-
An Experimental Study of Deep Neural Network Models for Vietnamese Multiple-Choice Reading Comprehension 20 Aug 2020 · 0 repositories · arXiv:2008.08810
-
PARADE: Passage Representation Aggregation for Document Reranking 20 Aug 2020 · 1 repository · arXiv:2008.09093Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Top2Vec: Distributed Representations of Topics 19 Aug 2020 · 2 repositories · arXiv:2008.09470Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
UoB at SemEval-2020 Task 12: Boosting BERT with Corpus Level Information 19 Aug 2020 · 0 repositories · arXiv:2008.08547
-
Ranking Clarification Questions via Natural Language Inference 18 Aug 2020 · 0 repositories · arXiv:2008.07688
-
Stock Index Prediction with Multi-task Learning and Word Polarity Over Time 17 Aug 2020 · 0 repositories · arXiv:2008.07605
-
DeVLBert: Learning Deconfounded Visio-Linguistic Representations 16 Aug 2020 · 1 repository · arXiv:2008.06884
-
Finding Fast Transformers: One-Shot Neural Architecture Search by Component Composition 15 Aug 2020 · 0 repositories · arXiv:2008.06808
-
Jointly Fine-Tuning "BERT-like" Self Supervised Models to Improve Multimodal Speech Emotion Recognition 15 Aug 2020 · 1 repository · arXiv:2008.06682
-
Jointly Fine-Tuning “BERT-like” Self Supervised Models to Improve Multimodal Speech Emotion Recognition 15 Aug 2020 · 1 repository
-
Hate Speech Detection and Racial Bias Mitigation in Social Media based on BERT model 14 Aug 2020 · 0 repositories · arXiv:2008.06460
-
ANDES at SemEval-2020 Task 12: A jointly-trained BERT multilingual model for offensive language detection 13 Aug 2020 · 1 repository · arXiv:2008.06408
-
MICE: Mining Idioms with Contextual Embeddings 13 Aug 2020 · 1 repository · arXiv:2008.05759
-
Variance-reduced Language Pretraining via a Mask Proposal Network 12 Aug 2020 · 0 repositories · arXiv:2008.05333
-
Beyond Lexical: A Semantic Retrieval Framework for Textual SearchEngine 10 Aug 2020 · 0 repositories · arXiv:2008.03917
-
On Commonsense Cues in BERT for Solving Commonsense Tasks 10 Aug 2020 · 0 repositories · arXiv:2008.03945
-
FireBERT: Hardening BERT-based classifiers against adversarial attack 10 Aug 2020 · 1 repository · arXiv:2008.04203
-
GANBERT: Generative Adversarial Networks with Bidirectional Encoder Representations from Transformers for MRI to PET synthesis 10 Aug 2020 · 0 repositories · arXiv:2008.04393
-
KR-BERT: A Small-Scale Korean-Specific Language Model 10 Aug 2020 · 1 repository · arXiv:2008.03979
-
Distilling the Knowledge of BERT for Sequence-to-Sequence ASR 9 Aug 2020 · 1 repository · arXiv:2008.03822
-
Fast and Accurate Neural CRF Constituency Parsing 9 Aug 2020 · 2 repositories · arXiv:2008.03736Syntology official (archive's flag): 1 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
SemEval-2020 Task 10: Emphasis Selection for Written Text in Visual Media 7 Aug 2020 · 0 repositories · arXiv:2008.03274
-
aschern at SemEval-2020 Task 11: It Takes Three to Tango: RoBERTa, CRF, and Transfer Learning 6 Aug 2020 · 1 repository · arXiv:2008.02837
-
ConvBERT: Improving BERT with Span-based Dynamic Convolution 6 Aug 2020 · 8 repositories · arXiv:2008.02496
-
DeText: A Deep Text Ranking Framework with BERT 6 Aug 2020 · 1 repository · arXiv:2008.02460Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Ontology-driven weak supervision for clinical entity classification in electronic health records 5 Aug 2020 · 1 repository · arXiv:2008.01972Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)