Methods › Natural Language Processing › Autoencoding Transformers › BERT › Papers, page 32
BERT
Papers archive 2025-07-28
archive papers tagged: 6,938 · with a code link: 2,862 · where Syntology ran a sample: 640 (520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (640 of 6,938 tagged: 520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument)
Page 32 of 70: papers 3,101 to 3,200 of 6,938, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
L3Cube-HindBERT and DevBERT: Pre-Trained BERT Transformer models for Devanagari based Hindi and Marathi Languages 21 Nov 2022 · 0 repositories · arXiv:2211.11418
-
L3Cube-MahaSBERT and HindSBERT: Sentence BERT Models and Benchmarking BERT Sentence Representations for Hindi and Marathi 21 Nov 2022 · 1 repository · arXiv:2211.11187
-
TCBERT: A Technical Report for Chinese Topic Classification BERT 21 Nov 2022 · 0 repositories · arXiv:2211.11304
-
Conceptor-Aided Debiasing of Large Language Models 20 Nov 2022 · 0 repositories · arXiv:2211.11087
-
Detecting Conspiracy Theory Against COVID-19 Vaccines 20 Nov 2022 · 0 repositories · arXiv:2211.13003
-
Feature Weaken: Vicinal Data Augmentation for Classification 20 Nov 2022 · 0 repositories · arXiv:2211.10944
-
Understanding and Improving Knowledge Distillation for Quantization-Aware Training of Large Transformer Encoders 20 Nov 2022 · 1 repository · arXiv:2211.11014
-
A survey on knowledge-enhanced multimodal learning 19 Nov 2022 · 0 repositories · arXiv:2211.12328
-
Entity-Assisted Language Models for Identifying Check-worthy Sentences 19 Nov 2022 · 0 repositories · arXiv:2211.10678
-
Leveraging Users' Social Network Embeddings for Fake News Detection on Twitter 19 Nov 2022 · 0 repositories · arXiv:2211.10672
-
Metadata Might Make Language Models Better 18 Nov 2022 · 0 repositories · arXiv:2211.10086
-
Where did you tweet from? Inferring the origin locations of tweets based on contextual information 18 Nov 2022 · 0 repositories · arXiv:2211.16506
-
LongFNT: Long-form Speech Recognition with Factorized Neural Transducer 17 Nov 2022 · 0 repositories · arXiv:2211.09412
-
ProtSi: Prototypical Siamese Network with Data Augmentation for Few-Shot Subjective Answer Evaluation 17 Nov 2022 · 1 repository · arXiv:2211.09855
-
Random-LTD: Random and Layerwise Token Dropping Brings Efficient Training for Large-scale Transformers 17 Nov 2022 · 1 repository · arXiv:2211.11586
-
Fast and Accurate FSA System Using ELBERT: An Efficient and Lightweight BERT 16 Nov 2022 · 0 repositories · arXiv:2211.08842
-
An FNet based Auto Encoder for Long Sequence News Story Generation 15 Nov 2022 · 1 repository · arXiv:2211.08295
-
Empowering Language Models with Knowledge Graph Reasoning for Question Answering 15 Nov 2022 · 0 repositories · arXiv:2211.08380
-
RobBERT-2022: Updating a Dutch Language Model to Account for Evolving Language Use 15 Nov 2022 · 0 repositories · arXiv:2211.08192
-
GreenPLM: Cross-Lingual Transfer of Monolingual Pre-Trained Language Models at Almost No Cost 13 Nov 2022 · 1 repository · arXiv:2211.06993
-
Xu at SemEval-2022 Task 4: Pre-BERT Neural Network Methods vs Post-BERT RoBERTa Approach for Patronizing and Condescending Language Detection 13 Nov 2022 · 1 repository · arXiv:2211.06874
-
Dark patterns in e-commerce: a dataset and its baseline evaluations 12 Nov 2022 · 1 repository · arXiv:2211.06543
-
Using Persuasive Writing Strategies to Explain and Detect Health Misinformation 11 Nov 2022 · 1 repository · arXiv:2211.05985
-
BERT-Based Combination of Convolutional and Recurrent Neural Network for Indonesian Sentiment Analysis 10 Nov 2022 · 0 repositories · arXiv:2211.05273
-
BERT in Plutarch's Shadows 10 Nov 2022 · 0 repositories · arXiv:2211.05673
-
Biomedical Multi-hop Question Answering Using Knowledge Graph Embeddings and Language Models 10 Nov 2022 · 0 repositories · arXiv:2211.05351
-
PAD-Net: An Efficient Framework for Dynamic Networks 10 Nov 2022 · 1 repository · arXiv:2211.05528
-
Syntax-Guided Domain Adaptation for Aspect-based Sentiment Analysis 10 Nov 2022 · 0 repositories · arXiv:2211.05457
-
Collateral facilitation in humans and language models 9 Nov 2022 · 1 repository · arXiv:2211.05198Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Cross-lingual Transfer Learning for Check-worthy Claim Identification over Twitter 9 Nov 2022 · 0 repositories · arXiv:2211.05087
-
Mask More and Mask Later: Efficient Pre-training of Masked Language Models by Disentangling the [MASK] Token 9 Nov 2022 · 1 repository · arXiv:2211.04898
-
Sentiment Analysis of Persian Language: Review of Algorithms, Approaches and Datasets 9 Nov 2022 · 0 repositories · arXiv:2212.06041
-
A Multimodal Approach for Dementia Detection from Spontaneous Speech with Tensor Fusion Layer 8 Nov 2022 · 0 repositories · arXiv:2211.04368
-
Discover, Explanation, Improvement: An Automatic Slice Detection Framework for Natural Language Processing 8 Nov 2022 · 0 repositories · arXiv:2211.04476
-
AD-BERT: Using Pre-trained contextualized embeddings to Predict the Progression from Mild Cognitive Impairment to Alzheimer's Disease 7 Nov 2022 · 0 repositories · arXiv:2212.06042
-
Suffix Retrieval-Augmented Language Modeling 6 Nov 2022 · 1 repository · arXiv:2211.03053
-
BERT-Deep CNN: State-of-the-Art for Sentiment Analysis of COVID-19 Tweets 4 Nov 2022 · 0 repositories · arXiv:2211.09733
-
BERT for Long Documents: A Case Study of Automated ICD Coding 4 Nov 2022 · 0 repositories · arXiv:2211.02519
-
Continuous Prompt Tuning Based Textual Entailment Model for E-commerce Entity Typing 4 Nov 2022 · 1 repository · arXiv:2211.02483
-
Fine-Tuning Language Models via Epistemic Neural Networks 3 Nov 2022 · 1 repository · arXiv:2211.01568Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder 2 Nov 2022 · 0 repositories · arXiv:2211.00792
-
Multi-level Distillation of Semantic Knowledge for Pre-training Multilingual Language Model 2 Nov 2022 · 0 repositories · arXiv:2211.01200
-
Investigating Content-Aware Neural Text-To-Speech MOS Prediction Using Prosodic and Linguistic Features 1 Nov 2022 · 0 repositories · arXiv:2211.00342
-
Reduce, Reuse, Recycle: Improving Training Efficiency with Distillation 1 Nov 2022 · 0 repositories · arXiv:2211.00683
-
Efficient Document Retrieval by End-to-End Refining and Quantizing BERT Embedding with Contrastive Product Quantization 31 Oct 2022 · 1 repository · arXiv:2210.17170
-
Leveraging Pre-trained Models for Failure Analysis Triplets Generation 31 Oct 2022 · 0 repositories · arXiv:2210.17497
-
QuaLA-MiniLM: a Quantized Length Adaptive MiniLM 31 Oct 2022 · 2 repositories · arXiv:2210.17114
-
SDCL: Self-Distillation Contrastive Learning for Chinese Spell Checking 31 Oct 2022 · 0 repositories · arXiv:2210.17168
-
Parameter-Efficient Tuning Makes a Good Classification Head 30 Oct 2022 · 1 repository · arXiv:2210.16771
-
BERT Meets CTC: New Formulation of End-to-End Speech Recognition with Pre-trained Masked Language Model 29 Oct 2022 · 0 repositories · arXiv:2210.16663
-
Empirical Evaluation of Post-Training Quantization Methods for Language Tasks 29 Oct 2022 · 0 repositories · arXiv:2210.16621
-
Exploiting prompt learning with pre-trained language models for Alzheimer's Disease detection 29 Oct 2022 · 1 repository · arXiv:2210.16539
-
BEBERT: Efficient and Robust Binary Ensemble BERT 28 Oct 2022 · 1 repository · arXiv:2210.15976
-
Feature Engineering vs BERT on Twitter Data 28 Oct 2022 · 0 repositories · arXiv:2210.16168
-
On the Use of Modality-Specific Large-Scale Pre-Trained Encoders for Multimodal Sentiment Analysis 28 Oct 2022 · 0 repositories · arXiv:2210.15937
-
Probing for targeted syntactic knowledge through grammatical error detection 28 Oct 2022 · 1 repository · arXiv:2210.16228
-
BERT-Flow-VAE: A Weakly-supervised Model for Multi-Label Text Classification 27 Oct 2022 · 0 repositories · arXiv:2210.15225
-
COCO-DR: Combating Distribution Shifts in Zero-Shot Dense Retrieval with Contrastive and Distributionally Robust Learning 27 Oct 2022 · 1 repository · arXiv:2210.15212Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
COST-EFF: Collaborative Optimization of Spatial and Temporal Efficiency with Slenderized Multi-exit Language Models 27 Oct 2022 · 1 repository · arXiv:2210.15523
-
Fast DistilBERT on CPUs 27 Oct 2022 · 1 repository · arXiv:2211.07715
-
FCTalker: Fine and Coarse Grained Context Modeling for Expressive Conversational Speech Synthesis 27 Oct 2022 · 1 repository · arXiv:2210.15360
-
Masked Vision-Language Transformer in Fashion 27 Oct 2022 · 1 repository · arXiv:2210.15110
-
Unsupervised Boundary-Aware Language Model Pretraining for Chinese Sequence Labeling 27 Oct 2022 · 2 repositories · arXiv:2210.15231
-
Automatic extraction of materials and properties from superconductors scientific literature 26 Oct 2022 · 2 repositories · arXiv:2210.15600
-
Bi-Link: Bridging Inductive Link Predictions from Text via Contrastive Learning of Transformers and Prompts 26 Oct 2022 · 0 repositories · arXiv:2210.14463
-
Don't Prompt, Search! Mining-based Zero-Shot Learning with Language Models 26 Oct 2022 · 0 repositories · arXiv:2210.14803
-
Exploring Robustness of Prefix Tuning in Noisy Data: A Case Study in Financial Sentiment Analysis 26 Oct 2022 · 0 repositories · arXiv:2211.05584
-
IELM: An Open Information Extraction Benchmark for Pre-Trained Language Models 25 Oct 2022 · 0 repositories · arXiv:2210.14128
-
Entity-level Sentiment Analysis in Contact Center Telephone Conversations 24 Oct 2022 · 0 repositories · arXiv:2210.13401
-
Explaining Translationese: why are Neural Classifiers Better and what do they Learn? 24 Oct 2022 · 0 repositories · arXiv:2210.13391
-
Exploring Euphemism Detection in Few-Shot and Zero-Shot Settings 24 Oct 2022 · 1 repository · arXiv:2210.12926
-
The Better Your Syntax, the Better Your Semantics? Probing Pretrained Language Models for the English Comparative Correlative 24 Oct 2022 · 0 repositories · arXiv:2210.13181
-
A BERT-based Deep Learning Approach for Reputation Analysis in Social Media 23 Oct 2022 · 0 repositories · arXiv:2211.01954
-
Data Augmentation for Automated Essay Scoring using Transformer Models 23 Oct 2022 · 0 repositories · arXiv:2210.12809
-
Meta-learning Pathologies from Radiology Reports using Variance Aware Prototypical Networks 22 Oct 2022 · 0 repositories · arXiv:2210.13979
-
Amos: An Adam-style Optimizer with Adaptive Weight Decay towards Model-Oriented Scale 21 Oct 2022 · 1 repository · arXiv:2210.11693Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples)
-
Discovering Differences in the Representation of People using Contextualized Semantic Axes 21 Oct 2022 · 1 repository · arXiv:2210.12170
-
Probing with Noise: Unpicking the Warp and Weft of Embeddings 21 Oct 2022 · 1 repository · arXiv:2210.12206
-
SpaBERT: A Pretrained Language Model from Geographic Data for Geo-Entity Representation 21 Oct 2022 · 0 repositories · arXiv:2210.12213
-
A Unified Neural Network Model for Readability Assessment with Feature Projection and Length-Balanced Loss 19 Oct 2022 · 1 repository · arXiv:2210.10305
-
BioGPT: Generative Pre-trained Transformer for Biomedical Text Generation and Mining 19 Oct 2022 · 4 repositories · arXiv:2210.10341
-
Language Model Decomposition: Quantifying the Dependency and Correlation of Language Models 19 Oct 2022 · 1 repository · arXiv:2210.10289
-
Tempo: Accelerating Transformer-Based Model Training through Memory Footprint Reduction 19 Oct 2022 · 1 repository · arXiv:2210.10246Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ELASTIC: Numerical Reasoning with Adaptive Symbolic Compiler 18 Oct 2022 · 1 repository · arXiv:2210.10105Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
CAN-BERT do it? Controller Area Network Intrusion Detection System based on BERT Language Model 17 Oct 2022 · 0 repositories · arXiv:2210.09439
-
iDNA-ABF: multi-scale deep biological language learning model for the interpretable prediction of DNA methylations 17 Oct 2022 · 2 repositories
-
Using Bottleneck Adapters to Identify Cancer in Clinical Notes under Low-Resource Constraints 17 Oct 2022 · 1 repository · arXiv:2210.09440
-
Zero-Shot Ranking Socio-Political Texts with Transformer Language Models to Reduce Close Reading Time 17 Oct 2022 · 0 repositories · arXiv:2210.09179
-
Acoustic-aware Non-autoregressive Spell Correction with Mask Sample Decoding 16 Oct 2022 · 0 repositories · arXiv:2210.08665
-
CTCBERT: Advancing Hidden-unit BERT with CTC Objectives 16 Oct 2022 · 0 repositories · arXiv:2210.08603
-
Improving Semantic Matching through Dependency-Enhanced Pre-trained Model with Adaptive Fusion 16 Oct 2022 · 0 repositories · arXiv:2210.08471
-
AraLegal-BERT: A pretrained language model for Arabic Legal text 15 Oct 2022 · 0 repositories · arXiv:2210.08284
-
DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation 14 Oct 2022 · 2 repositories · arXiv:2210.07558Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Kernel-Whitening: Overcome Dataset Bias with Isotropic Sentence Embedding 14 Oct 2022 · 2 repositories · arXiv:2210.07547
-
Overlooked Video Classification in Weakly Supervised Video Anomaly Detection 13 Oct 2022 · 1 repository · arXiv:2210.06688
-
SQuAT: Sharpness- and Quantization-Aware Training for BERT 13 Oct 2022 · 0 repositories · arXiv:2210.07171
-
Tone prediction and orthographic conversion for Basaa 13 Oct 2022 · 0 repositories · arXiv:2210.06986
-
Foundation Transformers 12 Oct 2022 · 4 repositories · arXiv:2210.06423Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
GMP*: Well-Tuned Gradual Magnitude Pruning Can Outperform Most BERT-Pruning Methods 12 Oct 2022 · 0 repositories · arXiv:2210.06384
-
On Text Style Transfer via Style Masked Language Models 12 Oct 2022 · 0 repositories · arXiv:2210.06394