Methods › Natural Language Processing › Autoencoding Transformers › BERT › Papers, page 27
BERT
Papers archive 2025-07-28
archive papers tagged: 6,938 · with a code link: 2,862 · where Syntology ran a sample: 640 (520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (640 of 6,938 tagged: 520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument)
Page 27 of 70: papers 2,601 to 2,700 of 6,938, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
UIT-Saviors at MEDVQA-GI 2023: Improving Multimodal Learning with Image Enhancement for Gastrointestinal Visual Question Answering 6 Jul 2023 · 0 repositories · arXiv:2307.02783
-
CAME: Confidence-guided Adaptive Memory Efficient Optimization 5 Jul 2023 · 2 repositories · arXiv:2307.02047Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Emoji Prediction in Tweets using BERT 5 Jul 2023 · 1 repository · arXiv:2307.02054
-
Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations 5 Jul 2023 · 0 repositories · arXiv:2307.03678
-
Named Entity Inclusion in Abstractive Text Summarization 5 Jul 2023 · 0 repositories · arXiv:2307.02570
-
KDSTM: Neural Semi-supervised Topic Modeling with Knowledge Distillation 4 Jul 2023 · 0 repositories · arXiv:2307.01878Syntology 6 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
ALBERTI, a Multilingual Domain Specific Language Model for Poetry Analysis 3 Jul 2023 · 0 repositories · arXiv:2307.01387
-
Improving Language Plasticity via Pretraining with Active Forgetting 3 Jul 2023 · 1 repository · arXiv:2307.01163
-
Interpretability and Transparency-Driven Detection and Transformation of Textual Adversarial Examples (IT-DT) 3 Jul 2023 · 0 repositories · arXiv:2307.01225
-
How far is Language Model from 100% Few-shot Named Entity Recognition in Medical Domain 1 Jul 2023 · 1 repository · arXiv:2307.00186
-
Ticket-BERT: Labeling Incident Management Tickets with Language Models 30 Jun 2023 · 0 repositories · arXiv:2307.00108
-
Classifying Crime Types using Judgment Documents from Social Media 29 Jun 2023 · 0 repositories · arXiv:2306.17020
-
Harnessing the Power of Hugging Face Transformers for Predicting Mental Health Disorders in Social Networks 29 Jun 2023 · 0 repositories · arXiv:2306.16891
-
An Efficient Sparse Inference Software Accelerator for Transformer-based Language Models on CPUs 28 Jun 2023 · 1 repository · arXiv:2306.16601
-
Beyond the Hype: Assessing the Performance, Trustworthiness, and Clinical Suitability of GPT3.5 28 Jun 2023 · 0 repositories · arXiv:2306.15887
-
Multi-Site Clinical Federated Learning using Recursive and Attentive Models and NVFlare 28 Jun 2023 · 0 repositories · arXiv:2306.16367
-
Gender Bias in BERT -- Measuring and Analysing Biases through Sentiment Rating in a Realistic Downstream Classification Task 27 Jun 2023 · 0 repositories · arXiv:2306.15298
-
Investigating Cross-Domain Behaviors of BERT in Review Understanding 27 Jun 2023 · 0 repositories · arXiv:2306.15123
-
MAT: Mixed-Strategy Game of Adversarial Training in Fine-tuning 27 Jun 2023 · 0 repositories · arXiv:2306.15826
-
SparseOptimizer: Sparsify Language Models through Moreau-Yosida Regularization and Accelerate via Compiler Co-design 27 Jun 2023 · 0 repositories · arXiv:2306.15656
-
Unleashing the Power of User Reviews: Exploring Airline Choices at Catania Airport, Italy 27 Jun 2023 · 0 repositories · arXiv:2306.15541
-
Constraint-aware and Ranking-distilled Token Pruning for Efficient Transformer Inference 26 Jun 2023 · 1 repository · arXiv:2306.14393
-
Addressing Cold Start Problem for End-to-end Automatic Speech Scoring 25 Jun 2023 · 0 repositories · arXiv:2306.14310
-
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices 25 Jun 2023 · 0 repositories · arXiv:2306.14263
-
Switch-BERT: Learning to Model Multimodal Interactions by Switching Attention and Input 25 Jun 2023 · 0 repositories · arXiv:2306.14182
-
Comparison of Pre-trained Language Models for Turkish Address Parsing 24 Jun 2023 · 0 repositories · arXiv:2306.13947
-
IERL: Interpretable Ensemble Representation Learning -- Combining CrowdSourced Knowledge and Distributed Semantic Representations 24 Jun 2023 · 0 repositories · arXiv:2306.13865
-
L3Cube-MahaSent-MD: A Multi-domain Marathi Sentiment Analysis Dataset and Transformer Models 24 Jun 2023 · 1 repository · arXiv:2306.13888
-
Math Word Problem Solving by Generating Linguistic Variants of Problem Statements 24 Jun 2023 · 1 repository · arXiv:2306.13899
-
My Boli: Code-mixed Marathi-English Corpora, Pretrained Language Models and Evaluation Benchmarks 24 Jun 2023 · 1 repository · arXiv:2306.14030
-
Partitioning-Guided K-Means: Extreme Empty Cluster Resolution for Extreme Model Compression 24 Jun 2023 · 0 repositories · arXiv:2306.14031
-
Resume Information Extraction via Post-OCR Text Processing 23 Jun 2023 · 0 repositories · arXiv:2306.13775
-
Named entity recognition in resumes 22 Jun 2023 · 0 repositories · arXiv:2306.13062
-
Investigating Pre-trained Language Models on Cross-Domain Datasets, a Step Closer to General AI 21 Jun 2023 · 0 repositories · arXiv:2306.12205
-
Fine-Tuning Language Models for Scientific Writing Support 19 Jun 2023 · 1 repository · arXiv:2306.10974
-
Instant Soup: Cheap Pruning Ensembles in A Single Pass Can Draw Lottery Tickets from Large Models 18 Jun 2023 · 1 repository · arXiv:2306.10460Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Investigating Masking-based Data Generation in Language Models 16 Jun 2023 · 0 repositories · arXiv:2307.00008
-
Revealing the impact of social circumstances on the selection of cancer therapy through natural language processing of social work notes 16 Jun 2023 · 0 repositories · arXiv:2306.09877
-
BED: Bi-Encoder-Based Detectors for Out-of-Distribution Detection 15 Jun 2023 · 1 repository · arXiv:2306.08852
-
Distillation Strategies for Discriminative Speech Recognition Rescoring 15 Jun 2023 · 0 repositories · arXiv:2306.09452
-
Mapping Researcher Activity based on Publication Data by means of Transformers 15 Jun 2023 · 0 repositories · arXiv:2306.09049
-
SLAMB: Accelerated Large Batch Training with Sparse Communication 15 Jun 2023 · 1 repository
-
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization 15 Jun 2023 · 0 repositories · arXiv:2306.09222
-
A semantically enhanced dual encoder for aspect sentiment triplet extraction 14 Jun 2023 · 1 repository · arXiv:2306.08373
-
Building a Corpus for Biomedical Relation Extraction of Species Mentions 14 Jun 2023 · 0 repositories · arXiv:2306.08403
-
Language models are not naysayers: An analysis of language models on negation benchmarks 14 Jun 2023 · 1 repository · arXiv:2306.08189
-
World-to-Words: Grounded Open Vocabulary Acquisition through Fast Mapping in Vision-Language Models 14 Jun 2023 · 1 repository · arXiv:2306.08685Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition 13 Jun 2023 · 0 repositories · arXiv:2306.07848
-
Improving Zero-Shot Detection of Low Prevalence Chest Pathologies using Domain Pre-trained Language Models 13 Jun 2023 · 1 repository · arXiv:2306.08000
-
Monolingual and Cross-Lingual Knowledge Transfer for Topic Classification 13 Jun 2023 · 0 repositories · arXiv:2306.07797
-
A Survey of Vision-Language Pre-training from the Lens of Multimodal Machine Translation 12 Jun 2023 · 0 repositories · arXiv:2306.07198
-
Imbalanced Multi-label Classification for Business-related Text with Moderately Large Label Spaces 12 Jun 2023 · 0 repositories · arXiv:2306.07046
-
Linear Classifier: An Often-Forgotten Baseline for Text Classification 12 Jun 2023 · 1 repository · arXiv:2306.07111Syntology official: no sample here; runs from other or unrecorded repositories · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Multimodal Audio-textual Architecture for Robust Spoken Language Understanding 12 Jun 2023 · 0 repositories · arXiv:2306.06819
-
EaSyGuide : ESG Issue Identification Framework leveraging Abilities of Generative Large Language Models 11 Jun 2023 · 1 repository · arXiv:2306.06662
-
RoBERTweet: A BERT Language Model for Romanian Tweets 11 Jun 2023 · 0 repositories · arXiv:2306.06598
-
Enhancing Low Resource NER Using Assisting Language And Transfer Learning 10 Jun 2023 · 0 repositories · arXiv:2306.06477
-
Medical Data Augmentation via ChatGPT: A Case Study on Medication Identification and Medication Event Classification 10 Jun 2023 · 0 repositories · arXiv:2306.07297
-
COVER: A Heuristic Greedy Adversarial Attack on Prompt-based Learning in Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.05659
-
End-to-End Neural Network Compression via ℓ₁/ℓ₂ Regularized Latency Surrogates 9 Jun 2023 · 0 repositories · arXiv:2306.05785
-
Implementing BERT and fine-tuned RobertA to detect AI generated news by ChatGPT 9 Jun 2023 · 0 repositories · arXiv:2306.07401
-
Prodigy: An Expeditiously Adaptive Parameter-Free Learner 9 Jun 2023 · 1 repository · arXiv:2306.06101
-
Understanding Telecom Language Through Large Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.07933
-
Augmenting Hessians with Inter-Layer Dependencies for Mixed-Precision Post-Training Quantization 8 Jun 2023 · 0 repositories · arXiv:2306.04879
-
Bias Against 93 Stigmatized Groups in Masked Language Models and Downstream Sentiment Classification Tasks 8 Jun 2023 · 1 repository · arXiv:2306.05550Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Extensive Evaluation of Transformer-based Architectures for Adverse Drug Events Extraction 8 Jun 2023 · 1 repository · arXiv:2306.05276
-
Leveraging Language Identification to Enhance Code-Mixed Text Classification 8 Jun 2023 · 0 repositories · arXiv:2306.04964
-
Mixture-of-Supernets: Improving Weight-Sharing Supernet Training with Architecture-Routed Mixture-of-Experts 8 Jun 2023 · 1 repository · arXiv:2306.04845
-
NOWJ at COLIEE 2023 -- Multi-Task and Ensemble Approaches in Legal Information Processing 8 Jun 2023 · 0 repositories · arXiv:2306.04903
-
An Empirical Analysis of Parameter-Efficient Methods for Debiasing Pre-Trained Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04067
-
Detecting Human Rights Violations on Social Media during Russia-Ukraine War 6 Jun 2023 · 0 repositories · arXiv:2306.05370
-
LEACE: Perfect linear concept erasure in closed form 6 Jun 2023 · 2 repositories · arXiv:2306.03819Syntology official (archive's flag): 3 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
On the Difference of BERT-style and CLIP-style Text Encoders 6 Jun 2023 · 1 repository · arXiv:2306.03678
-
COMET: Learning Cardinality Constrained Mixture of Experts with Trees and Local Search 5 Jun 2023 · 2 repositories · arXiv:2306.02824
-
On "Scientific Debt" in NLP: A Case for More Rigour in Language Model Pre-Training Research 5 Jun 2023 · 0 repositories · arXiv:2306.02870
-
Skill over Scale: The Case for Medium, Domain-Specific Models for SE 5 Jun 2023 · 0 repositories · arXiv:2306.03268
-
Using Sequences of Life-events to Predict Human Lives 5 Jun 2023 · 2 repositories · arXiv:2306.03009
-
SpellMapper: A non-autoregressive neural spellchecker for ASR customization with candidate retrieval based on n-gram mappings 4 Jun 2023 · 1 repository · arXiv:2306.02317
-
Financial sentiment analysis using FinBERT with application in predicting stock movement 3 Jun 2023 · 0 repositories · arXiv:2306.02136
-
MultiLegalPile: A 689GB Multilingual Legal Corpus 3 Jun 2023 · 0 repositories · arXiv:2306.02069
-
Concurrent Classifier Error Detection (CCED) in Large Scale Machine Learning Systems 2 Jun 2023 · 0 repositories · arXiv:2306.01820
-
Establishment of NLP-Based Greenwashing Pattern Detection Service 2 Jun 2023 · 0 repositories
-
Context selectivity with dynamic availability enables lifelong continual learning 2 Jun 2023 · 1 repository · arXiv:2306.01690
-
Word Embeddings for Banking Industry 2 Jun 2023 · 0 repositories · arXiv:2306.01807
-
Adapting Pre-trained Language Models to Vision-Language Tasks via Dynamic Visual Prompting 1 Jun 2023 · 1 repository · arXiv:2306.00409
-
Boosting the Performance of Transformer Architectures for Semantic Textual Similarity 1 Jun 2023 · 0 repositories · arXiv:2306.00708
-
Column Type Annotation using ChatGPT 1 Jun 2023 · 1 repository · arXiv:2306.00745
-
Feature Engineering-Based Detection of Buffer Overflow Vulnerability in Source Code Using Neural Networks 1 Jun 2023 · 0 repositories · arXiv:2306.07981
-
Make Pre-trained Model Reversible: From Parameter to Memory Efficient Fine-Tuning 1 Jun 2023 · 1 repository · arXiv:2306.00477Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Training-free Neural Architecture Search for RNNs and Transformers 1 Jun 2023 · 1 repository · arXiv:2306.00288Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
UCAS-IIE-NLP at SemEval-2023 Task 12: Enhancing Generalization of Multilingual BERT for Low-resource Sentiment Analysis 1 Jun 2023 · 1 repository · arXiv:2306.01093
-
Supplementary Features of BiLSTM for Enhanced Sequence Labeling 31 May 2023 · 1 repository · arXiv:2305.19928
-
Building Extractive Question Answering System to Support Human-AI Health Coaching Model for Sleep Domain 31 May 2023 · 0 repositories · arXiv:2305.19707
-
Catalysis distillation neural network for the few shot open catalyst challenge 31 May 2023 · 0 repositories · arXiv:2305.19545
-
DeepMerge: Deep-Learning-Based Region-Merging for Image Segmentation 31 May 2023 · 1 repository · arXiv:2305.19787
-
XPhoneBERT: A Pre-trained Multilingual Model for Phoneme Representations for Text-to-Speech 31 May 2023 · 2 repositories · arXiv:2305.19709Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples) · 10 pointer-only (licence)
-
Explaining Hate Speech Classification with Model Agnostic Methods 30 May 2023 · 0 repositories · arXiv:2306.00021
-
GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation 30 May 2023 · 0 repositories · arXiv:2305.18997
-
Multitask learning for recognizing stress and depression in social media 30 May 2023 · 0 repositories · arXiv:2305.18907
-
PreQuant: A Task-agnostic Quantization Approach for Pre-trained Language Models 30 May 2023 · 0 repositories · arXiv:2306.00014