Methods › Natural Language Processing › Autoencoding Transformers › BERT › Papers, page 41
BERT
Papers archive 2025-07-28
archive papers tagged: 6,938 · with a code link: 2,862 · where Syntology ran a sample: 640 (520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (640 of 6,938 tagged: 520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument)
Page 41 of 70: papers 4,001 to 4,100 of 6,938, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
On the Robustness of Reading Comprehension Models to Entity Renaming 16 Nov 2021 · 0 repositories
-
PARE: A Simple and Strong Baseline for Monolingual and Multilingual Distantly Supervised Relation Extraction 16 Nov 2021 · 0 repositories
-
Perturbations in the Wild: Leveraging Human-Written Text Perturbations for Realistic Adversarial Attack and Defense 16 Nov 2021 · 0 repositories
-
Pinyin-bert: A new solution to Chinese pinyin to character conversion task 16 Nov 2021 · 0 repositories
-
Probing BERT’s priors with serial reproduction chains 16 Nov 2021 · 0 repositories
-
PromptBERT: Improving BERT Sentence Embeddings with Prompts 16 Nov 2021 · 0 repositories
-
ReCo: Reliable Multi-hop Causal Reasoning via Structural Causal Recurrent Unit 16 Nov 2021 · 0 repositories
-
Representation of Ambiguity in Pre-Trained Sentence Embeddings 16 Nov 2021 · 0 repositories
-
SAMBERT: Improve Aspect Sentiment Triplet Extraction by Segmenting the Attention Maps of BERT 16 Nov 2021 · 0 repositories
-
Self-Supervised Contrastive Learning with Adversarial Perturbations for Robust Pretrained Language Models 16 Nov 2021 · 0 repositories
-
SHIELD: Defending Textual Neural Networks against Black-Box Adversarial Attacks with Stochastic Multi-Expert Patcher 16 Nov 2021 · 0 repositories
-
Softmax Bottleneck Makes Language Models Unable to Represent Multi-mode Word Distributions 16 Nov 2021 · 0 repositories
-
TACO: Pre-training of Deep Transformers with Attention Convolution using Disentangled Positional Representation 16 Nov 2021 · 0 repositories
-
The impact of lexical and grammatical processing on generating code from natural language 16 Nov 2021 · 0 repositories
-
Towards Fully Self-Supervised Learning of Knowledge from Unstructured Text 16 Nov 2021 · 0 repositories
-
Towards Improving Topic Models with the BERT-based Neural Topic Encoder 16 Nov 2021 · 0 repositories
-
Understanding Attention in Machine Reading Comprehension 16 Nov 2021 · 0 repositories
-
UNICON: Unsupervised Intent Discovery via Semantic-level Contrastive Learning 16 Nov 2021 · 0 repositories
-
Unsupervised multiple-choice question generation for out-of-domain Q&A fine-tuning 16 Nov 2021 · 0 repositories
-
Weight Squeezing: Reparameterization for Knowledge Transfer and Model Compression 16 Nov 2021 · 0 repositories
-
When classifying grammatical role, BERT doesn't care about word order... except when it matters 16 Nov 2021 · 0 repositories
-
Assessing gender bias in medical and scientific masked language models with StereoSet 15 Nov 2021 · 0 repositories · arXiv:2111.08088
-
Exploring Story Generation with Multi-task Objectives in Variational Autoencoders 15 Nov 2021 · 0 repositories · arXiv:2111.08133
-
IIITT@Dravidian-CodeMix-FIRE2021: Transliterate or translate? Sentiment analysis of code-mixed text in Dravidian languages 15 Nov 2021 · 1 repository · arXiv:2111.07906
-
Improving Prosody for Unseen Texts in Speech Synthesis by Utilizing Linguistic Information and Noisy Data 15 Nov 2021 · 0 repositories · arXiv:2111.07549
-
Scaling Law for Recommendation Models: Towards General-purpose User Representations 15 Nov 2021 · 0 repositories · arXiv:2111.11294
-
"Will You Find These Shortcuts?" A Protocol for Evaluating the Faithfulness of Input Salience Methods for Text Classification 14 Nov 2021 · 0 repositories · arXiv:2111.07367
-
SocialBERT -- Transformers for Online SocialNetwork Language Modelling 13 Nov 2021 · 0 repositories · arXiv:2111.07148
-
MS-LaTTE: A Dataset of Where and When To-do Tasks are Completed 12 Nov 2021 · 1 repository · arXiv:2111.06902
-
Character-level HyperNetworks for Hate Speech Detection 11 Nov 2021 · 1 repository · arXiv:2111.06336
-
Improving Large-scale Language Models and Resources for Filipino 11 Nov 2021 · 0 repositories · arXiv:2111.06053
-
Amazon SageMaker Model Parallelism: A General and Flexible Framework for Large Model Training 10 Nov 2021 · 0 repositories · arXiv:2111.05972
-
BagBERT: BERT-based bagging-stacking for multi-topic classification 10 Nov 2021 · 1 repository · arXiv:2111.05808
-
CEHR-BERT: Incorporating temporal information from structured EHR data to improve prediction tasks 10 Nov 2021 · 0 repositories · arXiv:2111.08585
-
Prune Once for All: Sparse Pre-Trained Language Models 10 Nov 2021 · 2 repositories · arXiv:2111.05754
-
DSBERT:Unsupervised Dialogue Structure learning with BERT 9 Nov 2021 · 0 repositories · arXiv:2111.04933
-
FPM: A Collection of Large-scale Foundation Pre-trained Language Models 9 Nov 2021 · 0 repositories · arXiv:2111.04909
-
Human-in-the-Loop Disinformation Detection: Stance, Sentiment, or Something Else? 9 Nov 2021 · 0 repositories · arXiv:2111.05139
-
AI-UPV at IberLEF-2021 DETOXIS task: Toxicity Detection in Immigration-Related Web News Comments Using Transformers and Statistical Models 8 Nov 2021 · 1 repository · arXiv:2111.04530
-
Chemical detection and indexing in PubMed full text articles using deep learning and rule-based methods 8 Nov 2021 · 0 repositories
-
Detecting Depression in Thai Blog Posts: a Dataset and a Baseline 8 Nov 2021 · 0 repositories · arXiv:2111.04574
-
Guiding Multi-Step Rearrangement Tasks with Natural Language Instructions 8 Nov 2021 · 2 repositories
-
Sexism Prediction in Spanish and English Tweets Using Monolingual and Multilingual BERT and Ensemble Models 8 Nov 2021 · 1 repository · arXiv:2111.04551
-
TACCL: Guiding Collective Algorithm Synthesis using Communication Sketches 8 Nov 2021 · 2 repositories · arXiv:2111.04867
-
TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning 7 Nov 2021 · 2 repositories · arXiv:2111.04198
-
Profitable Trade-Off Between Memory and Performance In Multi-Domain Chatbot Architectures 6 Nov 2021 · 0 repositories · arXiv:2111.03963
-
Context-Aware Transformer Transducer for Speech Recognition 5 Nov 2021 · 0 repositories · arXiv:2111.03250
-
Effective Cross-Utterance Language Modeling for Conversational Speech Recognition 5 Nov 2021 · 0 repositories · arXiv:2111.03333
-
IBERT: Idiom Cloze-style reading comprehension with Attention 5 Nov 2021 · 0 repositories · arXiv:2112.02994
-
Sexism Identification in Tweets and Gabs using Deep Neural Networks 5 Nov 2021 · 0 repositories · arXiv:2111.03612
-
A text autoencoder from transformer for fast encoding language representation 4 Nov 2021 · 0 repositories · arXiv:2111.02844
-
An Empirical Study of the Effectiveness of an Ensemble of Stand-alone Sentiment Detection Tools for Software Engineering Datasets 4 Nov 2021 · 1 repository · arXiv:2111.03196
-
Conformal prediction for text infilling and part-of-speech prediction 4 Nov 2021 · 1 repository · arXiv:2111.02592
-
An Empirical Study of Training End-to-End Vision-and-Language Transformers 3 Nov 2021 · 3 repositories · arXiv:2111.02387Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
BERT-DRE: BERT with Deep Recursive Encoder for Natural Language Sentence Matching 3 Nov 2021 · 0 repositories · arXiv:2111.02188
-
Detection of Hate Speech using BERT and Hate Speech Word Embedding with Deep Model 2 Nov 2021 · 0 repositories · arXiv:2111.01515
-
Sentence encoding for Dialogue Act classification 2 Nov 2021 · 1 repository
-
UQuAD1.0: Development of an Urdu Question Answering Training Data for Machine Reading Comprehension 2 Nov 2021 · 0 repositories · arXiv:2111.01543
-
Comparative Study of Long Document Classification 1 Nov 2021 · 0 repositories · arXiv:2111.00702
-
Identifying causal relations in tweets using deep learning: Use case on diabetes-related tweets from 2017-2021 1 Nov 2021 · 1 repository · arXiv:2111.01225
-
MAPLE – MAsking words to generate blackout Poetry using sequence-to-sequence LEarning 1 Nov 2021 · 1 repository
-
Recent Advances in Natural Language Processing via Large Pre-Trained Language Models: A Survey 1 Nov 2021 · 0 repositories · arXiv:2111.01243
-
FinEAS: Financial Embedding Analysis of Sentiment 31 Oct 2021 · 1 repository · arXiv:2111.00526
-
Backdoor Pre-trained Models Can Transfer to All 30 Oct 2021 · 1 repository · arXiv:2111.00197
-
DSEE: Dually Sparsity-embedded Efficient Tuning of Pre-trained Language Models 30 Oct 2021 · 1 repository · arXiv:2111.00160Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 10 pointer-only (licence)
-
Magic Pyramid: Accelerating Inference with Early Exiting and Token Pruning 30 Oct 2021 · 0 repositories · arXiv:2111.00230
-
ICDM 2020 Knowledge Graph Contest: Consumer Event-Cause Extraction 28 Oct 2021 · 0 repositories · arXiv:2110.15722
-
A Sequence to Sequence Model for Extracting Multiple Product Name Entities from Dialog 28 Oct 2021 · 0 repositories · arXiv:2110.14843
-
Bridge the Gap Between CV and NLP! A Gradient-based Textual Adversarial Attack Framework 28 Oct 2021 · 1 repository · arXiv:2110.15317
-
Colossal-AI: A Unified Deep Learning System For Large-Scale Parallel Training 28 Oct 2021 · 1 repository · arXiv:2110.14883
-
Pruning Attention Heads of Transformer Models Using A* Search: A Novel Approach to Compress Big NLP Architectures 28 Oct 2021 · 0 repositories · arXiv:2110.15225
-
Semi-Siamese Bi-encoder Neural Ranking Model Using Lightweight Fine-Tuning 28 Oct 2021 · 1 repository · arXiv:2110.14943
-
Anomaly-Injected Deep Support Vector Data Description for Text Outlier Detection 27 Oct 2021 · 0 repositories · arXiv:2110.14729
-
CLAUSEREC: A Clause Recommendation Framework for AI-aided Contract Authoring 26 Oct 2021 · 0 repositories · arXiv:2110.15794
-
Post-processing for Individual Fairness 26 Oct 2021 · 1 repository · arXiv:2110.13796Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
s2s-ft: Fine-Tuning Pretrained Transformer Encoders for Sequence-to-Sequence Learning 26 Oct 2021 · 1 repository · arXiv:2110.13640
-
TriBERT: Full-body Human-centric Audio-visual Representation Learning for Visual Sound Separation 26 Oct 2021 · 1 repository · arXiv:2110.13412Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Fine-tuning of Pre-trained Transformers for Hate, Offensive, and Profane Content Detection in English and Marathi 25 Oct 2021 · 1 repository · arXiv:2110.12687
-
Paradigm Shift in Language Modeling: Revisiting CNN for Modeling Sanskrit Originated Bengali and Hindi Language 25 Oct 2021 · 0 repositories · arXiv:2110.13032
-
Hate and Offensive Speech Detection in Hindi and Marathi 23 Oct 2021 · 0 repositories · arXiv:2110.12200
-
Double Trouble: How to not explain a text classifier's decisions using counterfactuals synthesized by masked language models? 22 Oct 2021 · 1 repository · arXiv:2110.11929
-
Learning Text-Image Joint Embedding for Efficient Cross-Modal Retrieval with Deep Feature Engineering 22 Oct 2021 · 1 repository · arXiv:2110.11592
-
CLOOB: Modern Hopfield Networks with InfoLOOB Outperform CLIP 21 Oct 2021 · 1 repository · arXiv:2110.11316Syntology official (archive's flag): 4 ran · 7 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Fast Model Editing at Scale 21 Oct 2021 · 3 repositories · arXiv:2110.11309Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Modeling Performance in Open-Domain Dialogue with PARADISE 21 Oct 2021 · 0 repositories · arXiv:2110.11164
-
Distributionally Robust Classifiers in Sentiment Analysis 20 Oct 2021 · 1 repository · arXiv:2110.10372
-
SLAM: A Unified Encoder for Speech and Language Modeling via Speech-Text Joint Pre-Training 20 Oct 2021 · 0 repositories · arXiv:2110.10329
-
Ensemble ALBERT on SQuAD 2.0 19 Oct 2021 · 1 repository · arXiv:2110.09665
-
Risks of AI Foundation Models in Education 19 Oct 2021 · 0 repositories · arXiv:2110.10024
-
A Data Bootstrapping Recipe for Low Resource Multilingual Relation Classification 18 Oct 2021 · 0 repositories · arXiv:2110.09570
-
BERMo: What can BERT learn from ELMo? 18 Oct 2021 · 0 repositories · arXiv:2110.15802
-
Ceasing hate withMoH: Hate Speech Detection in Hindi-English Code-Switched Language 18 Oct 2021 · 0 repositories · arXiv:2110.09393
-
Contextual Hate Speech Detection in Code Mixed Text using Transformer Based Approaches 18 Oct 2021 · 0 repositories · arXiv:2110.09338
-
ViraPart: A Text Refinement Framework for Automatic Speech Recognition and Natural Language Processing Tasks in Persian 18 Oct 2021 · 0 repositories · arXiv:2110.09086
-
BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models 16 Oct 2021 · 0 repositories
-
EncT5: A Framework for Fine-tuning T5 as Non-autoregressive Models 16 Oct 2021 · 1 repository · arXiv:2110.08426
-
Hierarchical Transformer Networks for Long-sequence and Multiple Clinical Documents Classification 16 Oct 2021 · 0 repositories
-
IMPLI: Investigating NLI Models' Performance on Figurative Language 16 Oct 2021 · 0 repositories
-
Models In a Spelling Bee: Language Models Implicitly Learn the Character Composition of Tokens 16 Oct 2021 · 0 repositories
-
Old BERT, New Tricks: Artificial Language Learning for Pre-Trained Language Models 16 Oct 2021 · 1 repository