Methods › Natural Language Processing › Autoencoding Transformers › BERT › Papers, page 64
BERT
Papers archive 2025-07-28
archive papers tagged: 6,938 · with a code link: 2,862 · where Syntology ran a sample: 640 (520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (640 of 6,938 tagged: 520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument)
Page 64 of 70: papers 6,301 to 6,400 of 6,938, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Exploring BERT Parameter Efficiency on the Stanford Question Answering Dataset v2.0 25 Feb 2020 · 0 repositories · arXiv:2002.10670
-
Improving BERT Fine-Tuning via Self-Ensemble and Self-Distillation 24 Feb 2020 · 1 repository · arXiv:2002.10345Syntology 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Predicting Subjective Features of Questions of QA Websites using BERT 24 Feb 2020 · 5 repositories · arXiv:2002.10107
-
Federated pretraining and fine tuning of BERT using clinical notes from multiple silos 20 Feb 2020 · 0 repositories · arXiv:2002.08562
-
Compressing BERT: Studying the Effects of Weight Pruning on Transfer Learning 19 Feb 2020 · 1 repository · arXiv:2002.08307
-
The Microsoft Toolkit of Multi-Task Deep Neural Networks for Natural Language Understanding 19 Feb 2020 · 3 repositories · arXiv:2002.07972
-
From English To Foreign Languages: Transferring Pre-trained Language Models 18 Feb 2020 · 1 repository · arXiv:2002.07306Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A Financial Service Chatbot based on Deep Bidirectional Transformers 17 Feb 2020 · 0 repositories · arXiv:2003.04987
-
Incorporating BERT into Neural Machine Translation 17 Feb 2020 · 3 repositories · arXiv:2002.06823
-
SBERT-WK: A Sentence Embedding Method by Dissecting BERT-based Word Models 16 Feb 2020 · 3 repositories · arXiv:2002.06652
-
The Utility of General Domain Transfer Learning for Medical Language Tasks 16 Feb 2020 · 0 repositories · arXiv:2002.06670
-
Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping 15 Feb 2020 · 4 repositories · arXiv:2002.06305
-
UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation 15 Feb 2020 · 2 repositories · arXiv:2002.06353Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
FQuAD: French Question Answering Dataset 14 Feb 2020 · 0 repositories · arXiv:2002.06071
-
Stress Test Evaluation of Transformer-based Models in Natural Language Understanding Tasks 14 Feb 2020 · 0 repositories · arXiv:2002.06261
-
Transformer on a Diet 14 Feb 2020 · 1 repository · arXiv:2002.06170
-
TwinBERT: Distilling Knowledge to Twin-Structured BERT Models for Efficient Retrieval 14 Feb 2020 · 2 repositories · arXiv:2002.06275
-
Understanding patient complaint characteristics using contextual clinical BERT embeddings 14 Feb 2020 · 0 repositories · arXiv:2002.05902
-
Training Large Neural Networks with Constant Memory using a New Execution Algorithm 13 Feb 2020 · 2 repositories · arXiv:2002.05645Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Learning to Compare for Better Training and Evaluation of Open Domain Natural Language Generation Models 12 Feb 2020 · 0 repositories · arXiv:2002.05058
-
Utilizing BERT Intermediate Layers for Aspect Based Sentiment Analysis and Natural Language Inference 12 Feb 2020 · 1 repository · arXiv:2002.04815
-
Multilingual Alignment of Contextual Word Representations 10 Feb 2020 · 1 repository · arXiv:2002.03518Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Application of Pre-training Models in Named Entity Recognition 9 Feb 2020 · 0 repositories · arXiv:2002.08902
-
Momentum Improves Normalized SGD 9 Feb 2020 · 0 repositories · arXiv:2002.03305
-
BERT-of-Theseus: Compressing BERT by Progressive Module Replacing 7 Feb 2020 · 2 repositories · arXiv:2002.02925Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
K-Adapter: Infusing Knowledge into Pre-Trained Models with Adapters 5 Feb 2020 · 2 repositories · arXiv:2002.01808Syntology 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Rapid Adaptation of BERT for Information Extraction on Domain-Specific Business Documents 5 Feb 2020 · 1 repository · arXiv:2002.01861
-
Interpretable & Time-Budget-Constrained Contextualization for Re-Ranking 4 Feb 2020 · 1 repository · arXiv:2002.01854
-
Bertrand-DR: Improving Text-to-SQL using a Discriminative Re-ranker 3 Feb 2020 · 1 repository · arXiv:2002.00557
-
Beat the AI: Investigating Adversarial Human Annotation for Reading Comprehension 2 Feb 2020 · 1 repository · arXiv:2002.00293
-
Fine-Tuning BERT for Schema-Guided Zero-Shot Dialogue State Tracking 1 Feb 2020 · 0 repositories · arXiv:2002.00181
-
Pretrained Transformers for Simple Question Answering over Knowledge Graphs 31 Jan 2020 · 1 repository · arXiv:2001.11985
-
Adversarial Training for Aspect-Based Sentiment Analysis with BERT 30 Jan 2020 · 4 repositories · arXiv:2001.11316
-
On the Importance of Word Order Information in Cross-lingual Sequence Labeling 30 Jan 2020 · 0 repositories · arXiv:2001.11164
-
PEL-BERT: A Joint Model for Protocol Entity Linking 28 Jan 2020 · 0 repositories · arXiv:2002.00744
-
BERT's output layer recognizes all hidden layers? Some Intriguing Phenomena and a simple way to boost BERT 25 Jan 2020 · 0 repositories · arXiv:2001.09309
-
Generation-Distillation for Efficient Natural Language Understanding in Low-Data Settings 25 Jan 2020 · 0 repositories · arXiv:2002.00733
-
Gesticulator: A framework for semantically-aware speech-driven gesture generation 25 Jan 2020 · 1 repository · arXiv:2001.09326Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
PoWER-BERT: Accelerating BERT Inference via Progressive Word-vector Elimination 24 Jan 2020 · 1 repository · arXiv:2001.08950Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Navigation-Based Candidate Expansion and Pretrained Language Models for Citation Recommendation 23 Jan 2020 · 0 repositories · arXiv:2001.08687
-
A multimodal deep learning approach for named entity recognition from social media 19 Jan 2020 · 0 repositories · arXiv:2001.06888
-
Deep Learning for Hindi Text Classification: A Comparison 19 Jan 2020 · 0 repositories · arXiv:2001.10340
-
Capturing Evolution in Word Usage: Just Add More Clusters? 18 Jan 2020 · 0 repositories · arXiv:2001.06629
-
RobBERT: a Dutch RoBERTa-based Language Model 17 Jan 2020 · 1 repository · arXiv:2001.06286Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Schema2QA: High-Quality and Low-Cost Q&A Agents for the Structured Web 16 Jan 2020 · 3 repositories · arXiv:2001.05609
-
FGN: Fusion Glyph Network for Chinese Named Entity Recognition 15 Jan 2020 · 1 repository · arXiv:2001.05272
-
A BERT based Sentiment Analysis and Key Entity Detection Approach for Online Financial Texts 14 Jan 2020 · 0 repositories · arXiv:2001.05326
-
AdaBERT: Task-Adaptive BERT Compression with Differentiable Neural Architecture Search 13 Jan 2020 · 1 repository · arXiv:2001.04246Syntology 0 ran · 3 unverified (of 3 harvested samples)
-
Représentations lexicales pour la détection non supervisée d'événements dans un flux de tweets : étude sur des corpus français et anglais 13 Jan 2020 · 1 repository · arXiv:2001.04139
-
Exploring and Improving Robustness of Multi Task Deep Neural Networks via Domain Agnostic Defenses 11 Jan 2020 · 1 repository · arXiv:2001.05286
-
Resolving the Scope of Speculation and Negation using Transformer-Based Architectures 9 Jan 2020 · 1 repository · arXiv:2001.02885
-
To Transfer or Not to Transfer: Misclassification Attacks Against Transfer Learned Text Classifiers 8 Jan 2020 · 0 repositories · arXiv:2001.02438
-
Improving Entity Linking by Modeling Latent Entity Type Information 6 Jan 2020 · 0 repositories · arXiv:2001.01447
-
Multi-Layer Content Interaction Through Quaternion Product For Visual Question Answering 3 Jan 2020 · 0 repositories · arXiv:2001.05840
-
BERT-AL: BERT for Arbitrarily Long Document Understanding 1 Jan 2020 · 0 repositories
-
Stacked DeBERT: All Attention in Incomplete Data for Text Classification 1 Jan 2020 · 1 repository · arXiv:2001.00137Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
oLMpics -- On what Language Model Pre-training Captures 31 Dec 2019 · 2 repositories · arXiv:1912.13283
-
AutoDiscern: Rating the Quality of Online Health Information with Hierarchical Encoder Attention-based Neural Networks 30 Dec 2019 · 1 repository · arXiv:1912.12999
-
Probing the phonetic and phonological knowledge of tones in Mandarin TTS models 23 Dec 2019 · 1 repository · arXiv:1912.10915
-
Harnessing Evolution of Multi-Turn Conversations for Effective Answer Retrieval 22 Dec 2019 · 1 repository · arXiv:1912.10554
-
Learning and Evaluating Contextual Embedding of Source Code 21 Dec 2019 · 2 repositories · arXiv:2001.00059
-
Pretrained Encyclopedia: Weakly Supervised Knowledge-Pretrained Language Model 20 Dec 2019 · 0 repositories · arXiv:1912.09637
-
Shareable Representations for Search Query Understanding 20 Dec 2019 · 0 repositories · arXiv:2001.04345
-
BERTje: A Dutch BERT Model 19 Dec 2019 · 2 repositories · arXiv:1912.09582Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
CJRC: A Reliable Human-Annotated Benchmark DataSet for Chinese Judicial Reading Comprehension 19 Dec 2019 · 0 repositories · arXiv:1912.09156
-
Neural Simile Recognition with Cyclic Multitask Learning and Local Attention 19 Dec 2019 · 1 repository · arXiv:1912.09084
-
A Multi-task Learning Model for Chinese-oriented Aspect Polarity Classification and Aspect Term Extraction 17 Dec 2019 · 6 repositories · arXiv:1912.07976
-
Cross-Lingual Ability of Multilingual BERT: An Empirical Study 17 Dec 2019 · 0 repositories · arXiv:1912.07840
-
The performance evaluation of Multi-representation in the Deep Learning models for Relation Extraction Task 17 Dec 2019 · 0 repositories · arXiv:1912.08290
-
Learning Malware Representation based on Execution Sequences 16 Dec 2019 · 0 repositories · arXiv:1912.07250
-
Multilingual is not enough: BERT for Finnish 15 Dec 2019 · 1 repository · arXiv:1912.07076
-
Robust Named Entity Recognition with Truecasing Pretraining 15 Dec 2019 · 0 repositories · arXiv:1912.07095
-
BERTQA -- Attention on Steroids 14 Dec 2019 · 0 repositories · arXiv:1912.10435
-
Towards Robust Toxic Content Classification 14 Dec 2019 · 1 repository · arXiv:1912.06872
-
TopoAct: Visually Exploring the Shape of Activations in Deep Learning 13 Dec 2019 · 1 repository · arXiv:1912.06332
-
WaLDORf: Wasteless Language-model Distillation On Reading-comprehension 13 Dec 2019 · 0 repositories · arXiv:1912.06638
-
BERT has a Moral Compass: Improvements of ethical and moral values of machines 11 Dec 2019 · 0 repositories · arXiv:1912.05238
-
Unsupervised Transfer Learning via BERT Neuron Selection 10 Dec 2019 · 0 repositories · arXiv:1912.05308
-
Adversarial Analysis of Natural Language Inference Systems 7 Dec 2019 · 0 repositories · arXiv:1912.03441
-
Personalized Patent Claim Generation and Measurement 7 Dec 2019 · 0 repositories · arXiv:1912.03502
-
Semantic Mask for Transformer based End-to-End Speech Recognition 6 Dec 2019 · 1 repository · arXiv:1912.03010
-
Why are Adaptive Methods Good for Attention Models? 6 Dec 2019 · 0 repositories · arXiv:1912.03194
-
Self-Supervised Contextual Language Representation of Radiology Reports to Improve the Identification of Communication Urgency 5 Dec 2019 · 0 repositories · arXiv:1912.02703
-
Acquiring Knowledge from Pre-trained Model to Neural Machine Translation 4 Dec 2019 · 0 repositories · arXiv:1912.01774
-
Enhancing Relation Extraction Using Syntactic Indicators and Sentential Contexts 4 Dec 2019 · 1 repository · arXiv:1912.01858
-
A Comparative Study of Pretrained Language Models on Thai Social Text Categorization 3 Dec 2019 · 0 repositories · arXiv:1912.01580
-
BERT for Large-scale Video Segment Classification with Test-time Augmentation 2 Dec 2019 · 0 repositories · arXiv:1912.01127
-
Leveraging Contextual Embeddings for Detecting Diachronic Semantic Shift 2 Dec 2019 · 0 repositories · arXiv:1912.01072
-
Fast and Accurate Stochastic Gradient Estimation 1 Dec 2019 · 1 repository
-
Inducing Relational Knowledge from BERT 28 Nov 2019 · 0 repositories · arXiv:1911.12753
-
Do Attention Heads in BERT Track Syntactic Dependencies? 27 Nov 2019 · 1 repository · arXiv:1911.12246
-
Evaluating Commonsense in Pre-trained Language Models 27 Nov 2019 · 1 repository · arXiv:1911.11931
-
Taking a Stance on Fake News: Towards Automatic Disinformation Assessment via Deep Bidirectional Transformer Language Models for Stance Detection 27 Nov 2019 · 0 repositories · arXiv:1911.11951
-
Low Rank Factorization for Compact Multi-Head Self-Attention 26 Nov 2019 · 1 repository · arXiv:1912.00835
-
Who did They Respond to? Conversation Structure Modeling using Masked Hierarchical Transformer 25 Nov 2019 · 1 repository · arXiv:1911.10666
-
Automatically Neutralizing Subjective Bias in Text 21 Nov 2019 · 1 repository · arXiv:1911.09709Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
Chemical-protein Interaction Extraction via Gaussian Probability Distribution and External Biomedical Knowledge 21 Nov 2019 · 1 repository · arXiv:1911.09487
-
Joint Emotion Label Space Modelling for Affect Lexica 20 Nov 2019 · 0 repositories · arXiv:1911.08782
-
Towards Lingua Franca Named Entity Recognition with BERT 19 Nov 2019 · 0 repositories · arXiv:1912.01389
-
Towards non-toxic landscapes: Automatic toxic comment detection using DNN 19 Nov 2019 · 0 repositories · arXiv:1911.08395