Methods › Natural Language Processing › Autoencoding Transformers › BERT › Papers, page 22
BERT
Papers archive 2025-07-28
archive papers tagged: 6,938 · with a code link: 2,862 · where Syntology ran a sample: 640 (520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (640 of 6,938 tagged: 520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument)
Page 22 of 70: papers 2,101 to 2,200 of 6,938, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
C-RAG: Certified Generation Risks for Retrieval-Augmented Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03181Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Enhancing textual textbook question answering with large language models and retrieval augmented generation 5 Feb 2024 · 1 repository · arXiv:2402.05128Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Financial Report Chunking for Effective Retrieval Augmented Generation 5 Feb 2024 · 1 repository · arXiv:2402.05131
-
LB-KBQA: Large-language-model and BERT based Knowledge-Based Question and Answering System 5 Feb 2024 · 0 repositories · arXiv:2402.05130
-
Multi-Lingual Malaysian Embedding: Leveraging Large Language Models for Semantic Representations 5 Feb 2024 · 0 repositories · arXiv:2402.03053
-
Breaking MLPerf Training: A Case Study on Optimizing BERT 4 Feb 2024 · 0 repositories · arXiv:2402.02447
-
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation 4 Feb 2024 · 0 repositories · arXiv:2402.14594
-
Data Quality Matters: Suicide Intention Detection on Social Media Posts Using RoBERTa-CNN 3 Feb 2024 · 0 repositories · arXiv:2402.02262
-
DE³-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks 3 Feb 2024 · 0 repositories · arXiv:2402.05948
-
Clarifying the Path to User Satisfaction: An Investigation into Clarification Usefulness 2 Feb 2024 · 1 repository · arXiv:2402.01934
-
LLM-Detector: Improving AI-Generated Chinese Text Detection with Open-Source LLM Instruction Tuning 2 Feb 2024 · 1 repository · arXiv:2402.01158
-
Predicting ATP binding sites in protein sequences using Deep Learning and Natural Language Processing 2 Feb 2024 · 0 repositories · arXiv:2402.01829
-
Retrieval Augmented End-to-End Spoken Dialog Models 2 Feb 2024 · 0 repositories · arXiv:2402.01828
-
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks 2 Feb 2024 · 0 repositories · arXiv:2402.01176
-
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA 1 Feb 2024 · 0 repositories · arXiv:2402.01767
-
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models 1 Feb 2024 · 1 repository · arXiv:2402.00794
-
Self-Supervised Contrastive Pre-Training for Multivariate Point Processes 1 Feb 2024 · 0 repositories · arXiv:2402.00987
-
RAG-Fusion: a New Take on Retrieval-Augmented Generation 31 Jan 2024 · 0 repositories · arXiv:2402.03367
-
Arabic Tweet Act: A Weighted Ensemble Pre-Trained Transformer Model for Classifying Arabic Speech Acts on Twitter 30 Jan 2024 · 0 repositories · arXiv:2401.17373
-
Breaking Free Transformer Models: Task-specific Context Attribution Promises Improved Generalizability Without Fine-tuning Pre-trained LLMs 30 Jan 2024 · 1 repository · arXiv:2401.16638
-
CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models 30 Jan 2024 · 1 repository · arXiv:2401.17043Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Detecting mental disorder on social media: a ChatGPT-augmented explainable approach 30 Jan 2024 · 1 repository · arXiv:2401.17477
-
Detecting Racist Text in Bengali: An Ensemble Deep Learning Framework 30 Jan 2024 · 0 repositories · arXiv:2401.16748
-
Fine-tuning Transformer-based Encoder for Turkish Language Understanding Tasks 30 Jan 2024 · 0 repositories · arXiv:2401.17396
-
Large Multi-Modal Models (LMMs) as Universal Foundation Models for AI-Native Wireless Systems 30 Jan 2024 · 0 repositories · arXiv:2402.01748
-
Single Word Change is All You Need: Designing Attacks and Defenses for Text Classifiers 30 Jan 2024 · 0 repositories · arXiv:2401.17196
-
Towards Generating Informative Textual Description for Neurons in Language Models 30 Jan 2024 · 0 repositories · arXiv:2401.16731
-
Credit Risk Meets Large Language Models: Building a Risk Indicator from Loan Descriptions in P2P Lending 29 Jan 2024 · 0 repositories · arXiv:2401.16458
-
Development and Testing of a Novel Large Language Model-Based Clinical Decision Support Systems for Medication Safety in 12 Clinical Specialties 29 Jan 2024 · 0 repositories · arXiv:2402.01741
-
Development and Testing of Retrieval Augmented Generation in Large Language Models -- A Case Study Report 29 Jan 2024 · 0 repositories · arXiv:2402.01733
-
BPDec: Unveiling the Potential of Masked Language Modeling Decoder in BERT pretraining 29 Jan 2024 · 0 repositories · arXiv:2401.15861
-
Multi-class Regret Detection in Hindi Devanagari Script 29 Jan 2024 · 0 repositories · arXiv:2401.16561
-
Contrastive Learning and Mixture of Experts Enables Precise Vector Embeddings 28 Jan 2024 · 1 repository · arXiv:2401.15713
-
UnMASKed: Quantifying Gender Biases in Masked Language Models through Linguistically Informed Job Market Prompts 28 Jan 2024 · 0 repositories · arXiv:2401.15798
-
Enhancing Large Language Model Performance To Answer Questions and Extract Information More Accurately 27 Jan 2024 · 0 repositories · arXiv:2402.01722
-
MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries 27 Jan 2024 · 2 repositories · arXiv:2401.15391Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
From RAG to QA-RAG: Integrating Generative AI for Pharmaceutical Regulatory Compliance Process 26 Jan 2024 · 1 repository · arXiv:2402.01717
-
The Power of Noise: Redefining Retrieval for RAG Systems 26 Jan 2024 · 3 repositories · arXiv:2401.14887Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
(Chat)GPT v BERT: Dawn of Justice for Semantic Change Detection 25 Jan 2024 · 1 repository · arXiv:2401.14040
-
Socially Aware Synthetic Data Generation for Suicidal Ideation Detection Using Large Language Models 25 Jan 2024 · 0 repositories · arXiv:2402.01712
-
Proactive Emotion Tracker: AI-Driven Continuous Mood and Emotion Monitoring 24 Jan 2024 · 0 repositories · arXiv:2401.13722
-
Segment Any Cell: A SAM-based Auto-prompting Fine-tuning Framework for Nuclei Segmentation 24 Jan 2024 · 0 repositories · arXiv:2401.13220
-
Contrastive Learning in Distilled Models 23 Jan 2024 · 1 repository · arXiv:2401.12472
-
Fast Adversarial Training against Textual Adversarial Attacks 23 Jan 2024 · 0 repositories · arXiv:2401.12461
-
Revolutionizing Retrieval-Augmented Generation with Enhanced PDF Structure Recognition 23 Jan 2024 · 0 repositories · arXiv:2401.12599
-
APT: Adaptive Pruning and Tuning Pretrained Language Models for Efficient Training and Inference 22 Jan 2024 · 1 repository · arXiv:2401.12200Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Keep Decoding Parallel with Effective Knowledge Distillation from Language Models to End-to-end Speech Recognisers 22 Jan 2024 · 0 repositories · arXiv:2401.11700
-
Zero-Space Cost Fault Tolerance for Transformer-based Language Models on ReRAM 22 Jan 2024 · 0 repositories · arXiv:2401.11664
-
Confidence Preservation Property in Knowledge Distillation Abstractions 21 Jan 2024 · 0 repositories · arXiv:2401.11365
-
SEBERTNets: Sequence Enhanced BERT Networks for Event Entity Extraction Tasks Oriented to the Finance Field 21 Jan 2024 · 1 repository · arXiv:2401.11408
-
Drop your Decoder: Pre-training with Bag-of-Word Prediction for Dense Passage Retrieval 20 Jan 2024 · 3 repositories · arXiv:2401.11248Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Prompt-RAG: Pioneering Vector Embedding-Free Retrieval-Augmented Generation in Niche Domains, Exemplified by Korean Medicine 20 Jan 2024 · 0 repositories · arXiv:2401.11246
-
Unfair TOS: An Automated Approach using Customized BERT 20 Jan 2024 · 0 repositories · arXiv:2401.11207
-
Mining experimental data from Materials Science literature with Large Language Models: an evaluation study 19 Jan 2024 · 1 repository · arXiv:2401.11052
-
ChatQA: Surpassing GPT-4 on Conversational QA and RAG 18 Jan 2024 · 0 repositories · arXiv:2401.10225
-
BERTologyNavigator: Advanced Question Answering with BERT-based Semantics 17 Jan 2024 · 0 repositories · arXiv:2401.09553
-
Efficient slot labelling 17 Jan 2024 · 0 repositories · arXiv:2401.09343
-
Improving Classification Performance With Human Feedback: Label a few, we label the rest 17 Jan 2024 · 0 repositories · arXiv:2401.09555
-
A Reproducibility Study of Goldilocks: Just-Right Tuning of BERT for TAR 16 Jan 2024 · 1 repository · arXiv:2401.08104
-
RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture 16 Jan 2024 · 0 repositories · arXiv:2401.08406
-
A character-based steganography using masked language modeling 15 Jan 2024 · 1 repository
-
Graph database while computationally efficient filters out quickly the ESG integrated equities in investment management 15 Jan 2024 · 0 repositories · arXiv:2401.07483
-
Towards Efficient Methods in Medical Question Answering using Knowledge Graph Embeddings 15 Jan 2024 · 1 repository · arXiv:2401.07977Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Leveraging the power of transformers for guilt detection in text 15 Jan 2024 · 0 repositories · arXiv:2401.07414
-
SemEval-2017 Task 4: Sentiment Analysis in Twitter using BERT 15 Jan 2024 · 1 repository · arXiv:2401.07944
-
The Chronicles of RAG: The Retriever, the Chunk and the Generator 15 Jan 2024 · 0 repositories · arXiv:2401.07883
-
Understanding YTHDF2-mediated mRNA Degradation By m6A-BERT-Deg 15 Jan 2024 · 1 repository · arXiv:2401.08004
-
Promptformer: Prompted Conformer Transducer for ASR 14 Jan 2024 · 0 repositories · arXiv:2401.07360
-
Bridging the Preference Gap between Retrievers and LLMs 13 Jan 2024 · 0 repositories · arXiv:2401.06954
-
An investigation of structures responsible for gender bias in BERT and DistilBERT 12 Jan 2024 · 0 repositories · arXiv:2401.06495
-
Improved Learned Sparse Retrieval with Corpus-Specific Vocabularies 12 Jan 2024 · 1 repository · arXiv:2401.06703
-
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation 12 Jan 2024 · 1 repository · arXiv:2401.06583
-
Analyzing Regional Impacts of Climate Change using Natural Language Processing Techniques 11 Jan 2024 · 0 repositories · arXiv:2401.06817
-
Prompt-based mental health screening from social media text 11 Jan 2024 · 0 repositories · arXiv:2401.05912
-
Reinforcement Learning for Optimizing RAG for Domain Chatbots 10 Jan 2024 · 0 repositories · arXiv:2401.06800
-
An Assessment on Comprehending Mental Health through Large Language Models 9 Jan 2024 · 0 repositories · arXiv:2401.04592
-
DepressionEmo: A novel dataset for multilabel classification of depression emotions 9 Jan 2024 · 1 repository · arXiv:2401.04655
-
Language Detection for Transliterated Content 9 Jan 2024 · 0 repositories · arXiv:2401.04619
-
Phishing Website Detection through Multi-Model Analysis of HTML Content 9 Jan 2024 · 0 repositories · arXiv:2401.04820
-
Anatomy of Neural Language Models 8 Jan 2024 · 1 repository · arXiv:2401.03797
-
Advancing bioinformatics with large language models: components, applications and perspectives 8 Jan 2024 · 0 repositories · arXiv:2401.04155
-
RoBERTurk: Adjusting RoBERTa for Turkish 7 Jan 2024 · 0 repositories · arXiv:2401.03515
-
PIXAR: Auto-Regressive Language Modeling in Pixel Space 6 Jan 2024 · 0 repositories · arXiv:2401.03321
-
Natural Language Programming in Medicine: Administering Evidence Based Clinical Workflows with Autonomous Agents Powered by Generative Large Language Models 5 Jan 2024 · 0 repositories · arXiv:2401.02851
-
German Text Embedding Clustering Benchmark 5 Jan 2024 · 1 repository · arXiv:2401.02709
-
Beyond Extraction: Contextualising Tabular Data for Efficient Summarisation by Language Models 4 Jan 2024 · 0 repositories · arXiv:2401.02333
-
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages 4 Jan 2024 · 1 repository · arXiv:2401.02254
-
Studying and Recommending Information Highlighting in Stack Overflow Answers 3 Jan 2024 · 1 repository · arXiv:2401.01472
-
Enhancing Multilingual Information Retrieval in Mixed Human Resources Environments: A RAG Model Implementation for Multicultural Enterprise 3 Jan 2024 · 0 repositories · arXiv:2401.01511
-
Iterative Mask Filling: An Effective Text Augmentation Method Using Masked Language Modeling 3 Jan 2024 · 0 repositories · arXiv:2401.01830
-
MLPs Compass: What is learned when MLPs are combined with PLMs? 3 Jan 2024 · 0 repositories · arXiv:2401.01667
-
Natural Language Processing and Multimodal Stock Price Prediction 3 Jan 2024 · 0 repositories · arXiv:2401.01487
-
Revisiting Counterfactual Problems in Referring Expression Comprehension 1 Jan 2024 · 1 repository
-
An Analysis of Embedding Layers and Similarity Scores using Siamese Neural Networks 31 Dec 2023 · 0 repositories · arXiv:2401.00582
-
RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models 31 Dec 2023 · 3 repositories · arXiv:2401.00396Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation 30 Dec 2023 · 1 repository · arXiv:2401.00280
-
Why is the User Interface a Dark Pattern? : Explainable Auto-Detection and its Analysis 30 Dec 2023 · 1 repository · arXiv:2401.04119
-
MosaicBERT: A Bidirectional Encoder Optimized for Fast Pretraining 29 Dec 2023 · 1 repository · arXiv:2312.17482
-
TuPy-E: detecting hate speech in Brazilian Portuguese social media with a novel dataset and comprehensive analysis of models 29 Dec 2023 · 1 repository · arXiv:2312.17704
-
Language Model as an Annotator: Unsupervised Context-aware Quality Phrase Generation 28 Dec 2023 · 0 repositories · arXiv:2312.17349