Methods › Natural Language Processing › Subword Segmentation › WordPiece › Papers, page 45
WordPiece
Papers archive 2025-07-28
archive papers tagged: 7,063 · with a code link: 2,910 · where Syntology ran a sample: 650 (529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,063 tagged: 529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument)
Page 45 of 71: papers 4,401 to 4,500 of 7,063, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards Offensive Language Identification for Tamil Code-Mixed YouTube Comments and Posts 24 Aug 2021 · 1 repository · arXiv:2108.10939
-
Weakly Supervised Cross-platform Teenager Detection with Adversarial BERT 24 Aug 2021 · 0 repositories · arXiv:2108.10619
-
Deploying a BERT-based Query-Title Relevance Classifier in a Production System: a View from the Trenches 23 Aug 2021 · 0 repositories · arXiv:2108.10197
-
Query Embedding Pruning for Dense Retrieval 23 Aug 2021 · 1 repository · arXiv:2108.10341
-
Regularizing Transformers With Deep Probabilistic Layers 23 Aug 2021 · 0 repositories · arXiv:2108.10764
-
Sarcasm Detection in Twitter -- Performance Impact while using Data Augmentation: Word Embeddings 23 Aug 2021 · 1 repository · arXiv:2108.09924
-
Using Large Pre-Trained Models with Cross-Modal Attention for Multi-Modal Emotion Recognition 22 Aug 2021 · 0 repositories · arXiv:2108.09669
-
UzBERT: pretraining a BERT model for Uzbek 22 Aug 2021 · 0 repositories · arXiv:2108.09814
-
Extracting Radiological Findings With Normalized Anatomical Information Using a Span-Based BERT Relation Extraction Model 20 Aug 2021 · 0 repositories · arXiv:2108.09211
-
Frozen Pretrained Transformers for Neural Sign Language Translation 20 Aug 2021 · 1 repository
-
A Framework for Neural Topic Modeling of Text Corpora 19 Aug 2021 · 1 repository · arXiv:2108.08946
-
Detection of Illicit Drug Trafficking Events on Instagram: A Deep Multimodal Multilabel Learning Approach 19 Aug 2021 · 0 repositories · arXiv:2108.08920
-
Fast Passage Re-ranking with Contextualized Exact Term Matching and Efficient Passage Expansion 19 Aug 2021 · 1 repository · arXiv:2108.08513
-
Fine-Grained Element Identification in Complaint Text of Internet Fraud 19 Aug 2021 · 0 repositories · arXiv:2108.08676
-
How Hateful are Movies? A Study and Prediction on Movie Subtitles 19 Aug 2021 · 1 repository · arXiv:2108.10724
-
Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models 19 Aug 2021 · 2 repositories · arXiv:2108.08877
-
UNIQORN: Unified Question Answering over RDF Knowledge Graphs and Natural Language Text 19 Aug 2021 · 1 repository · arXiv:2108.08614
-
Contributions of Transformer Attention Heads in Multi- and Cross-lingual Tasks 18 Aug 2021 · 0 repositories · arXiv:2108.08375
-
Integrating Dialog History into End-to-End Spoken Language Understanding Systems 18 Aug 2021 · 0 repositories · arXiv:2108.08405
-
SIFN: A Sentiment-aware Interactive Fusion Network for Review-based Item Recommendation 18 Aug 2021 · 0 repositories · arXiv:2108.08022
-
TSI: an Ad Text Strength Indicator using Text-to-CTR and Semantic-Ad-Similarity 18 Aug 2021 · 0 repositories · arXiv:2108.08226
-
Response Ranking with Multi-types of Deep Interactive Representations in Retrieval-based Dialogues 17 Aug 2021 · 1 repository
-
Deep Natural Language Processing for LinkedIn Search 16 Aug 2021 · 0 repositories · arXiv:2108.13300
-
On the Opportunities and Risks of Foundation Models 16 Aug 2021 · 2 repositories · arXiv:2108.07258
-
Maps Search Misspelling Detection Leveraging Domain-Augmented Contextual Representations 15 Aug 2021 · 0 repositories · arXiv:2108.06842
-
Towards Structured Dynamic Sparse Pre-Training of BERT 13 Aug 2021 · 0 repositories · arXiv:2108.06277
-
AMMUS : A Survey of Transformer-based Pretrained Models in Natural Language Processing 12 Aug 2021 · 1 repository · arXiv:2108.05542
-
Modeling Relevance Ranking under the Pre-training and Fine-tuning Paradigm 12 Aug 2021 · 0 repositories · arXiv:2108.05652
-
Multimodal analysis of the predictability of hand-gesture properties 12 Aug 2021 · 0 repositories · arXiv:2108.05762
-
Overview of the HASOC track at FIRE 2020: Hate Speech and Offensive Content Identification in Indo-European Languages 12 Aug 2021 · 0 repositories · arXiv:2108.05927
-
Medical-VLBERT: Medical Visual Language BERT for COVID-19 CT Report Generation With Alternate Learning 11 Aug 2021 · 0 repositories · arXiv:2108.05067
-
NoFake at CheckThat! 2021: Fake News Detection Using BERT 11 Aug 2021 · 0 repositories · arXiv:2108.05419
-
A Study of Social and Behavioral Determinants of Health in Lung Cancer Patients Using Transformers-based Natural Language Processing Models 10 Aug 2021 · 0 repositories · arXiv:2108.04949
-
BROS: A Pre-trained Language Model Focusing on Text and Layout for Better Key Information Extraction from Documents 10 Aug 2021 · 2 repositories · arXiv:2108.04539Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples)
-
SynCoBERT: Syntax-Guided Multi-Modal Contrastive Pre-Training for Code Representation 10 Aug 2021 · 0 repositories · arXiv:2108.04556
-
Embodied BERT: A Transformer Model for Embodied, Language-guided Visual Task Completion 10 Aug 2021 · 1 repository · arXiv:2108.04927
-
TrUMAn: Trope Understanding in Movies and Animations 10 Aug 2021 · 0 repositories · arXiv:2108.04542
-
DoSSIER@COLIEE 2021: Leveraging dense retrieval and summarization-based re-ranking for case law retrieval 9 Aug 2021 · 1 repository · arXiv:2108.03937
-
FiLMing Multimodal Sarcasm Detection with Attention 9 Aug 2021 · 1 repository · arXiv:2110.00416
-
Efficacy of BERT embeddings on predicting disaster from Twitter data 8 Aug 2021 · 1 repository · arXiv:2108.10698
-
Leveraging Commonsense Knowledge on Classifying False News and Determining Checkworthiness of Claims 8 Aug 2021 · 0 repositories · arXiv:2108.03731
-
Deriving Disinformation Insights from Geolocalized Twitter Callouts 6 Aug 2021 · 1 repository · arXiv:2108.03067
-
Offensive Language and Hate Speech Detection with Deep Learning and Transfer Learning 6 Aug 2021 · 0 repositories · arXiv:2108.03305
-
Adaptive Residue-wise Profile Fusion for Low Homologous Protein SecondaryStructure Prediction Using External Knowledge 5 Aug 2021 · 0 repositories · arXiv:2108.04176
-
Decoupled Transformer for Scalable Inference in Open-domain Question Answering 5 Aug 2021 · 0 repositories · arXiv:2108.02765
-
Robust Transfer Learning with Pretrained Language Models through Adapters 5 Aug 2021 · 0 repositories · arXiv:2108.02340
-
Sentiment Analysis on the News to Improve Mental Health 5 Aug 2021 · 0 repositories · arXiv:2108.07706
-
Curriculum learning for language modeling 4 Aug 2021 · 1 repository · arXiv:2108.02170Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
ExBERT: An External Knowledge Enhanced BERT for Natural Language Inference 3 Aug 2021 · 0 repositories · arXiv:2108.01589
-
HTTP2vec: Embedding of HTTP Requests for Detection of Anomalous Traffic 3 Aug 2021 · 0 repositories · arXiv:2108.01763
-
Large-Scale Differentially Private BERT 3 Aug 2021 · 0 repositories · arXiv:2108.01624
-
Changes in European Solidarity Before and During COVID-19: Evidence from a Large Crowd- and Expert-Annotated Twitter Dataset 2 Aug 2021 · 1 repository · arXiv:2108.01042
-
LICHEE: Improving Language Model Pre-training with Multi-grained Tokenization 2 Aug 2021 · 1 repository · arXiv:2108.00801
-
Relation Aware Semi-autoregressive Semantic Parsing for NL2SQL 2 Aug 2021 · 0 repositories · arXiv:2108.00804
-
Transfer Learning for Mining Feature Requests and Bug Reports from Tweets and App Store Reviews 2 Aug 2021 · 1 repository · arXiv:2108.00663
-
1213Li at SemEval-2021 Task 6: Detection of Propaganda with Multi-modal Attention and Pre-trained Models 1 Aug 2021 · 0 repositories
-
ADEPT: An Adjective-Dependent Plausibility Task 1 Aug 2021 · 0 repositories
-
AStarTwice at SemEval-2021 Task 5: Toxic Span Detection Using RoBERTa-CRF, Domain Specific Pre-Training and Self-Training 1 Aug 2021 · 0 repositories
-
AttesTable at SemEval-2021 Task 9: Extending Statement Verification with Tables for Unknown Class, and Semantic Evidence Finding 1 Aug 2021 · 1 repository
-
BennettNLP at SemEval-2021 Task 5: Toxic Spans Detection using Stacked Embedding Powered Toxic Entity Recognizer 1 Aug 2021 · 0 repositories
-
BERTAC: Enhancing Transformer-based Language Models with Adversarially Pretrained Convolutional Neural Networks 1 Aug 2021 · 1 repository
-
Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection Incremental 1 Aug 2021 · 0 repositories
-
Can Transformer Models Measure Coherence In Text: Re-Thinking the Shuffle Test 1 Aug 2021 · 0 repositories
-
CMTA: COVID-19 Misinformation Multilingual Analysis on Twitter 1 Aug 2021 · 0 repositories
-
Cross-lingual Evidence Improves Monolingual Fake News Detection 1 Aug 2021 · 1 repository
-
CSECU-DSG at SemEval-2021 Task 1: Fusion of Transformer Models for Lexical Complexity Prediction 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 5: Leveraging Ensemble of Sequence Tagging Models for Toxic Spans Detection 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 6: Orchestrating Multimodal Neural Architectures for Identifying Persuasion Techniques in Texts and Images 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 7: Detecting and Rating Humor and Offense Employing Transformers 1 Aug 2021 · 0 repositories
-
DeepBlueAI at SemEval-2021 Task 7: Detecting and Rating Humor and Offense with Stacking Diverse Language Model-Based Methods 1 Aug 2021 · 0 repositories
-
DLJUST at SemEval-2021 Task 7: Hahackathon: Linking Humor and Offense 1 Aug 2021 · 0 repositories
-
Early Detection of Sexual Predators in Chats 1 Aug 2021 · 1 repository
-
eMLM: A New Pre-training Objective for Emotion Related Tasks 1 Aug 2021 · 1 repository
-
EndTimes at SemEval-2021 Task 7: Detecting and Rating Humor and Offense with BERT and Ensembles 1 Aug 2021 · 0 repositories
-
ES-JUST at SemEval-2021 Task 7: Detecting and Rating Humor and Offensive Text Using Deep Learning 1 Aug 2021 · 0 repositories
-
Explaining Contextualization in Language Models using Visual Analytics 1 Aug 2021 · 0 repositories
-
GHOST at SemEval-2021 Task 5: Is explanation all you need? 1 Aug 2021 · 0 repositories
-
GhostBERT: Generate More Features with Cheap Operations for BERT 1 Aug 2021 · 0 repositories
-
Grenzlinie at SemEval-2021 Task 7: Detecting and Rating Humor and Offense 1 Aug 2021 · 0 repositories
-
Gulu at SemEval-2021 Task 7: Detecting and Rating Humor and Offense 1 Aug 2021 · 0 repositories
-
GX at SemEval-2021 Task 2: BERT with Lemma Information for MCL-WiC Task 1 Aug 2021 · 1 repository
-
HamiltonDinggg at SemEval-2021 Task 5: Investigating Toxic Span Detection using RoBERTa Pre-training 1 Aug 2021 · 0 repositories
-
How effective is BERT without word ordering? Implications for language understanding and data privacy 1 Aug 2021 · 0 repositories
-
How Many Layers and Why? An Analysis of the Model Depth in Transformers 1 Aug 2021 · 0 repositories
-
hub at SemEval-2021 Task 1: Fusion of Sentence and Word Frequency to Predict Lexical Complexity 1 Aug 2021 · 0 repositories
-
hub at SemEval-2021 Task 2: Word Meaning Similarity Prediction Model Based on RoBERTa and Word Frequency 1 Aug 2021 · 0 repositories
-
hub at SemEval-2021 Task 7: Fusion of ALBERT and Word Frequency Information Detecting and Rating Humor and Offense 1 Aug 2021 · 0 repositories
-
IITK@LCP at SemEval-2021 Task 1: Classification for Lexical Complexity Regression Task 1 Aug 2021 · 0 repositories
-
Issues with Entailment-based Zero-shot Text Classification 1 Aug 2021 · 1 repository
-
JCT at SemEval-2021 Task 1: Context-aware Representation for Lexical Complexity Prediction 1 Aug 2021 · 0 repositories
-
JUST-BLUE at SemEval-2021 Task 1: Predicting Lexical Complexity using BERT and RoBERTa Pre-trained Language Models 1 Aug 2021 · 0 repositories
-
LeCun at SemEval-2021 Task 6: Detecting Persuasion Techniques in Text Using Ensembled Pretrained Transformers and Data Augmentation 1 Aug 2021 · 0 repositories
-
LeeBERT: Learned Early Exit for BERT with cross-level optimization 1 Aug 2021 · 0 repositories
-
LIORI at SemEval-2021 Task 8: Ask Transformer for measurements 1 Aug 2021 · 0 repositories
-
Lotus at SemEval-2021 Task 2: Combination of BERT and Paraphrasing for English Word Sense Disambiguation 1 Aug 2021 · 0 repositories
-
Measure and Evaluation of Semantic Divergence across Two Languages 1 Aug 2021 · 0 repositories
-
Measuring and Improving BERT's Mathematical Abilities by Predicting the Order of Reasoning. 1 Aug 2021 · 0 repositories
-
MedAI at SemEval-2021 Task 5: Start-to-end Tagging Framework for Toxic Spans Detection 1 Aug 2021 · 0 repositories
-
MinD at SemEval-2021 Task 6: Propaganda Detection using Transfer Learning and Multimodal Fusion 1 Aug 2021 · 0 repositories
-
MVP-BERT: Multi-Vocab Pre-training for Chinese BERT 1 Aug 2021 · 0 repositories