Methods › Natural Language Processing › Autoencoding Transformers › BERT › Papers, page 25
BERT
Papers archive 2025-07-28
archive papers tagged: 6,938 · with a code link: 2,862 · where Syntology ran a sample: 640 (520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (640 of 6,938 tagged: 520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument)
Page 25 of 70: papers 2,401 to 2,500 of 6,938, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Auditing Gender Analyzers on Text Data 9 Oct 2023 · 0 repositories · arXiv:2310.06061
-
Cabbage Sweeter than Cake? Analysing the Potential of Large Language Models for Learning Conceptual Spaces 9 Oct 2023 · 0 repositories · arXiv:2310.05481
-
Foundation Models Meet Visualizations: Challenges and Opportunities 9 Oct 2023 · 0 repositories · arXiv:2310.05771
-
Transformer Fusion with Optimal Transport 9 Oct 2023 · 1 repository · arXiv:2310.05719Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection 8 Oct 2023 · 0 repositories · arXiv:2310.05115
-
Enhancing Pre-Trained Language Models with Sentence Position Embeddings for Rhetorical Roles Recognition in Legal Opinions 8 Oct 2023 · 0 repositories · arXiv:2310.05276
-
LLM4VV: Developing LLM-Driven Testsuite for Compiler Validation 8 Oct 2023 · 1 repository · arXiv:2310.04963
-
RAC-BERT: Character Radical Enhanced BERT for Ancient Chinese 8 Oct 2023 · 0 repositories
-
A Process for Topic Modelling Via Word Embeddings 6 Oct 2023 · 0 repositories · arXiv:2312.03705
-
Automatic Aspect Extraction from Scientific Texts 6 Oct 2023 · 1 repository · arXiv:2310.04074
-
Quantized Transformer Language Model Implementations on Edge Devices 6 Oct 2023 · 0 repositories · arXiv:2310.03971
-
Segmented Harmonic Loss: Handling Class-Imbalanced Multi-Label Clinical Data for Medical Coding with Large Language Models 6 Oct 2023 · 0 repositories · arXiv:2310.04595
-
COVID-19 South African Vaccine Hesitancy Models Show Boost in Performance Upon Fine-Tuning on M-pox Tweets 4 Oct 2023 · 0 repositories · arXiv:2310.04453
-
Memoria: Resolving Fateful Forgetting Problem through Human-Inspired Memory Architecture 4 Oct 2023 · 1 repository · arXiv:2310.03052Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 10 harvested samples)
-
Retrieval-augmented Generation to Improve Math Question-Answering: Trade-offs Between Groundedness and Human Preference 4 Oct 2023 · 2 repositories · arXiv:2310.03184Syntology official (archive's flag): 7 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 20 harvested samples)
-
Harnessing Pre-Trained Sentence Transformers for Offensive Language Detection in Indian Languages 3 Oct 2023 · 0 repositories · arXiv:2310.02249
-
Label Supervised LLaMA Finetuning 2 Oct 2023 · 2 repositories · arXiv:2310.01208
-
Natural Language Models for Data Visualization Utilizing nvBench Dataset 2 Oct 2023 · 0 repositories · arXiv:2310.00832
-
Target-Aware Contextual Political Bias Detection in News 2 Oct 2023 · 0 repositories · arXiv:2310.01138
-
Question-Answering Model for Schizophrenia Symptoms and Their Impact on Daily Life using Mental Health Forums Data 30 Sep 2023 · 0 repositories · arXiv:2310.00448
-
RelBERT: Embedding Relations with Language Models 30 Sep 2023 · 1 repository · arXiv:2310.00299
-
Intuitive or Dependent? Investigating LLMs' Behavior Style to Conflicting Prompts 29 Sep 2023 · 0 repositories · arXiv:2309.17415
-
MKRAG: Medical Knowledge Retrieval Augmented Generation for Medical Question Answering 27 Sep 2023 · 0 repositories · arXiv:2309.16035
-
CAPP-130: A Corpus of Chinese Application Privacy Policy Summarization and Interpretation 26 Sep 2023 · 1 repository
-
Low-rank Adaptation of Large Language Model Rescoring for Parameter-Efficient Speech Recognition 26 Sep 2023 · 0 repositories · arXiv:2309.15223
-
RAGAS: Automated Evaluation of Retrieval Augmented Generation 26 Sep 2023 · 3 repositories · arXiv:2309.15217Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Comprehensive Overview of Named Entity Recognition: Models, Domain-Specific Applications and Challenges 25 Sep 2023 · 0 repositories · arXiv:2309.14084
-
Accelerating Large Batch Training via Gradient Signal to Noise Ratio (GSNR) 24 Sep 2023 · 0 repositories · arXiv:2309.13681
-
Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models 24 Sep 2023 · 1 repository · arXiv:2309.13579
-
Lexical Squad@Multimodal Hate Speech Event Detection 2023: Multimodal Hate Speech Detection using Fused Ensemble Approach 23 Sep 2023 · 1 repository · arXiv:2309.13354
-
AMPLIFY:Attention-based Mixup for Performance Improvement and Label Smoothing in Transformer 22 Sep 2023 · 1 repository · arXiv:2309.12689
-
TOPFORMER: Topology-Aware Authorship Attribution of Deepfake Texts with Diverse Writing Styles 22 Sep 2023 · 1 repository · arXiv:2309.12934
-
Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection 21 Sep 2023 · 1 repository · arXiv:2309.12247
-
BayesTune: Bayesian Sparse Deep Model Fine-tuning 21 Sep 2023 · 1 repository
-
Implicit Differentiable Outlier Detection Enable Robust Deep Multimodal Analysis 21 Sep 2023 · 1 repository
-
Making Scalable Meta Learning Practical 21 Sep 2023 · 1 repository
-
Marich: A Query-efficient Distributionally Equivalent Model Extraction Attack 21 Sep 2023 · 1 repository
-
On the Relationship between Skill Neurons and Robustness in Prompt Tuning 21 Sep 2023 · 1 repository · arXiv:2309.12263
-
[Re] Exploring the Role of Grammar and Word Choice in Bias Toward African American English (AAE) in Hate Speech Classification 21 Sep 2023 · 0 repositories
-
SLHCat: Mapping Wikipedia Categories and Lists to DBpedia by Leveraging Semantic, Lexical, and Hierarchical Features 21 Sep 2023 · 0 repositories · arXiv:2309.11791
-
SPICED: News Similarity Detection Dataset with Multiple Topics and Complexity Levels 21 Sep 2023 · 0 repositories · arXiv:2309.13080
-
Stock Market Sentiment Classification and Backtesting via Fine-tuned BERT 21 Sep 2023 · 0 repositories · arXiv:2309.11979
-
The Cambridge Law Corpus: A Dataset for Legal AI Research 21 Sep 2023 · 0 repositories · arXiv:2309.12269
-
Towards Efficient Pre-Trained Language Model via Feature Correlation Distillation 21 Sep 2023 · 0 repositories
-
AttentionMix: Data augmentation method that relies on BERT attention mechanism 20 Sep 2023 · 0 repositories · arXiv:2309.11104
-
CoT-BERT: Enhancing Unsupervised Sentence Representation through Chain-of-Thought 20 Sep 2023 · 2 repositories · arXiv:2309.11143
-
CPLLM: Clinical Prediction with Large Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11295
-
GPT-MolBERTa: GPT Molecular Features Language Model for molecular property prediction 20 Sep 2023 · 0 repositories · arXiv:2310.03030
-
Sequence-to-Sequence Spanish Pre-trained Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11259
-
Mixed-Distil-BERT: Code-mixed Language Modeling for Bangla, English, and Hindi 19 Sep 2023 · 0 repositories · arXiv:2309.10272
-
Facilitating NSFW Text Detection in Open-Domain Dialogue Systems via Knowledge Distillation 18 Sep 2023 · 1 repository · arXiv:2309.09749
-
Proposition from the Perspective of Chinese Language: A Chinese Proposition Classification Evaluation Benchmark 18 Sep 2023 · 0 repositories · arXiv:2309.09602
-
Detecting covariate drift in text data using document embeddings and dimensionality reduction 17 Sep 2023 · 1 repository · arXiv:2309.10000
-
SplitEE: Early Exit in Deep Neural Networks with Split Computing 17 Sep 2023 · 1 repository · arXiv:2309.09195
-
Has Sentiment Returned to the Pre-pandemic Level? A Sentiment Analysis Using U.S. College Subreddit Data from 2019 to 2022 16 Sep 2023 · 1 repository · arXiv:2309.08845
-
AlbNER: A Corpus for Named Entity Recognition in Albanian 15 Sep 2023 · 0 repositories · arXiv:2309.08741
-
Detecting Relevant Information in High-Volume Chat Logs: Keyphrase Extraction for Grooming and Drug Dealing Forensic Analysis 15 Sep 2023 · 0 repositories · arXiv:2311.04905
-
Structural Self-Supervised Objectives for Transformers 15 Sep 2023 · 1 repository · arXiv:2309.08272
-
Transformer Based Punctuation Restoration for Turkish 15 Sep 2023 · 1 repository
-
VulnSense: Efficient Vulnerability Detection in Ethereum Smart Contracts by Multimodal Learning with Graph Neural Network and Language Model 15 Sep 2023 · 0 repositories · arXiv:2309.08474
-
Automatic Data Visualization Generation from Chinese Natural Language Questions 14 Sep 2023 · 0 repositories · arXiv:2309.07650
-
DebCSE: Rethinking Unsupervised Contrastive Sentence Embedding Learning in the Debiasing Perspective 14 Sep 2023 · 0 repositories · arXiv:2309.07396
-
EnCodecMAE: Leveraging neural codecs for universal audio representation learning 14 Sep 2023 · 2 repositories · arXiv:2309.07391Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Text Classification of Cancer Clinical Trial Eligibility Criteria 14 Sep 2023 · 0 repositories · arXiv:2309.07812
-
Balanced and Explainable Social Media Analysis for Public Health with Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05951
-
PRESTI: Predicting Repayment Effort of Self-Admitted Technical Debt Using Textual Information 12 Sep 2023 · 0 repositories · arXiv:2309.06020
-
Overview of Memotion 3: Sentiment and Emotion Analysis of Codemixed Hinglish Memes 12 Sep 2023 · 0 repositories · arXiv:2309.06517
-
Applying BioBERT to Extract Germline Gene-Disease Associations for Building a Knowledge Graph from the Biomedical Literature 11 Sep 2023 · 1 repository · arXiv:2309.13061
-
CrisisTransformers: Pre-trained language models and sentence encoders for crisis-related social media texts 11 Sep 2023 · 0 repositories · arXiv:2309.05494
-
Detecting Natural Language Biases with Prompt-based Learning 11 Sep 2023 · 0 repositories · arXiv:2309.05227
-
Learning Personalized User Preference from Cold Start in Multi-turn Conversations 10 Sep 2023 · 0 repositories · arXiv:2309.05127
-
Neural-Hidden-CRF: A Robust Weakly-Supervised Sequence Labeler 10 Sep 2023 · 1 repository · arXiv:2309.05086
-
RGAT: A Deeper Look into Syntactic Dependency Information for Coreference Resolution 10 Sep 2023 · 0 repositories · arXiv:2309.04977
-
Encoding Multi-Domain Scientific Papers by Ensembling Multiple CLS Tokens 8 Sep 2023 · 1 repository · arXiv:2309.04333Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Fuzzy Fingerprinting Transformer Language-Models for Emotion Recognition in Conversations 8 Sep 2023 · 0 repositories · arXiv:2309.04292
-
Leveraging Pretrained Image-text Models for Improving Audio-Visual Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04628
-
UQ at #SMM4H 2023: ALEX for Public Health Analysis with Social Media 8 Sep 2023 · 1 repository · arXiv:2309.04213
-
Certifying LLM Safety against Adversarial Prompting 6 Sep 2023 · 1 repository · arXiv:2309.02705
-
Leave no Place Behind: Improved Geolocation in Humanitarian Documents 6 Sep 2023 · 0 repositories · arXiv:2309.02914
-
Offensive Hebrew Corpus and Detection using BERT 6 Sep 2023 · 1 repository · arXiv:2309.02724
-
Self-Supervised Masked Digital Elevation Models Encoding for Low-Resource Downstream Tasks 6 Sep 2023 · 0 repositories · arXiv:2309.03367
-
Incorporating Dictionaries into a Neural Network Architecture to Extract COVID-19 Medical Concepts From Social Media 5 Sep 2023 · 0 repositories · arXiv:2309.02188
-
Leveraging BERT Language Models for Multi-Lingual ESG Issue Identification 5 Sep 2023 · 0 repositories · arXiv:2309.02189
-
Sample Size in Natural Language Processing within Healthcare Research 5 Sep 2023 · 0 repositories · arXiv:2309.02237
-
Benchmarking Large Language Models in Retrieval-Augmented Generation 4 Sep 2023 · 1 repository · arXiv:2309.01431Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Study on the Implementation of Generative AI Services Using an Enterprise Data-Based LLM Application Architecture 3 Sep 2023 · 0 repositories · arXiv:2309.01105
-
A Visual Interpretation-Based Self-Improved Classification System Using Virtual Adversarial Training 3 Sep 2023 · 0 repositories · arXiv:2309.01196
-
Knowledge Graph Embeddings for Multi-Lingual Structured Representations of Radiology Reports 2 Sep 2023 · 0 repositories · arXiv:2309.00917
-
Studying the impacts of pre-training using ChatGPT-generated text on downstream tasks 2 Sep 2023 · 0 repositories · arXiv:2309.05668
-
BatchPrompt: Accomplish more with less 1 Sep 2023 · 1 repository · arXiv:2309.00384
-
SortedNet: A Scalable and Generalized Framework for Training Modular Deep Neural Networks 1 Sep 2023 · 0 repositories · arXiv:2309.00255
-
Can humans help BERT gain "confidence"? 31 Aug 2023 · 0 repositories · arXiv:2309.06580
-
DictaBERT: A State-of-the-Art BERT Suite for Modern Hebrew 31 Aug 2023 · 0 repositories · arXiv:2308.16687
-
Linking microblogging sentiments to stock price movement: An application of GPT-4 31 Aug 2023 · 0 repositories · arXiv:2308.16771
-
Towards Improving the Expressiveness of Singing Voice Synthesis with BERT Derived Semantic Information 31 Aug 2023 · 0 repositories · arXiv:2308.16836
-
SP³: Enhancing Structured Pruning via PCA Projection 31 Aug 2023 · 1 repository · arXiv:2308.16475Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Analyzing Character and Consciousness in AI-Generated Social Content: A Case Study of Chirper, the AI Social Network 30 Aug 2023 · 0 repositories · arXiv:2309.08614
-
The DeepZen Speech Synthesis System for Blizzard Challenge 2023 30 Aug 2023 · 0 repositories · arXiv:2308.15945
-
SpikeBERT: A Language Spikformer Learned from BERT with Knowledge Distillation 29 Aug 2023 · 1 repository · arXiv:2308.15122Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
ANER: Arabic and Arabizi Named Entity Recognition using Transformer-Based Approach 28 Aug 2023 · 1 repository · arXiv:2308.14669