Methods › Natural Language Processing › Autoencoding Transformers › BERT › Papers, page 17
BERT
Papers archive 2025-07-28
archive papers tagged: 6,938 · with a code link: 2,862 · where Syntology ran a sample: 640 (520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (640 of 6,938 tagged: 520 with a run with no instrument failure, 120 where every run was a failure of Syntology's instrument)
Page 17 of 70: papers 1,601 to 1,700 of 6,938, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MeMemo: On-device Retrieval Augmentation for Private and Personalized Text Generation 2 Jul 2024 · 1 repository · arXiv:2407.01972
-
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs 2 Jul 2024 · 0 repositories · arXiv:2407.02485
-
The Solution for The PST-KDD-2024 OAG-Challenge 2 Jul 2024 · 0 repositories · arXiv:2407.12827
-
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation 1 Jul 2024 · 1 repository · arXiv:2407.01102Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Face4RAG: Factual Consistency Evaluation for Retrieval Augmented Generation in Chinese 1 Jul 2024 · 0 repositories · arXiv:2407.01080
-
Ground Every Sentence: Improving Retrieval-Augmented LLMs with Interleaved Reference-Claim Generation 1 Jul 2024 · 0 repositories · arXiv:2407.01796
-
Hybrid RAG-empowered Multi-modal LLM for Secure Data Management in Internet of Medical Things: A Diffusion-based Contract Approach 1 Jul 2024 · 0 repositories · arXiv:2407.00978
-
Multi-Modal Fusion-Based Multi-Task Semantic Communication System 1 Jul 2024 · 0 repositories · arXiv:2407.00964
-
Retrieval-augmented generation in multilingual settings 1 Jul 2024 · 1 repository · arXiv:2407.01463Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Searching for Best Practices in Retrieval-Augmented Generation 1 Jul 2024 · 1 repository · arXiv:2407.01219Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Summary of a Haystack: A Challenge to Long-Context LLMs and RAG Systems 1 Jul 2024 · 1 repository · arXiv:2407.01370Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Memory³: Language Modeling with Explicit Memory 1 Jul 2024 · 0 repositories · arXiv:2407.01178
-
Characterizing Stereotypical Bias from Privacy-preserving Pre-Training 30 Jun 2024 · 0 repositories · arXiv:2407.00764
-
LegalTurk Optimized BERT for Multi-Label Text Classification and NER 30 Jun 2024 · 0 repositories · arXiv:2407.00648
-
Parm: Efficient Training of Large Sparsely-Activated Models with Dedicated Schedules 30 Jun 2024 · 1 repository · arXiv:2407.00599
-
Answering real-world clinical questions using large language model based systems 29 Jun 2024 · 0 repositories · arXiv:2407.00541
-
From RAG to RICHES: Retrieval Interlaced with Sequence Generation 29 Jun 2024 · 0 repositories · arXiv:2407.00361
-
LLM-Generated Natural Language Meets Scaling Laws: New Explorations and Data Augmentation Methods 29 Jun 2024 · 0 repositories · arXiv:2407.00322
-
Uncertainty Quantification in Large Language Models Through Convex Hull Analysis 28 Jun 2024 · 0 repositories · arXiv:2406.19712
-
AutoPureData: Automated Filtering of Undesirable Web Data to Update LLM Knowledge 27 Jun 2024 · 1 repository · arXiv:2406.19271
-
AutoRAG-HP: Automatic Online Hyper-Parameter Tuning for Retrieval-Augmented Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19251
-
Historia Magistra Vitae: Dynamic Topic Modeling of Roman Literature using Neural Embeddings 27 Jun 2024 · 0 repositories · arXiv:2406.18907
-
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language 27 Jun 2024 · 0 repositories · arXiv:2406.19349
-
RAVEN: Multitask Retrieval Augmented Vision-Language Learning 27 Jun 2024 · 0 repositories · arXiv:2406.19150
-
SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation 27 Jun 2024 · 1 repository · arXiv:2406.19215Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Generating Is Believing: Membership Inference Attacks against Retrieval-Augmented Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19234
-
Evaluating Quality of Answers for Retrieval-Augmented Generation: A Strong LLM Is All You Need 26 Jun 2024 · 0 repositories · arXiv:2406.18064
-
"Glue pizza and eat rocks" -- Exploiting Vulnerabilities in Retrieval-Augmented Generative Models 26 Jun 2024 · 0 repositories · arXiv:2406.19417
-
Knowledge graph enhanced retrieval-augmented generation for failure mode and effects analysis 26 Jun 2024 · 1 repository · arXiv:2406.18114
-
Multi-step Inference over Unstructured Data 26 Jun 2024 · 0 repositories · arXiv:2406.17987
-
Poisoned LangChain: Jailbreak LLMs by LangChain 26 Jun 2024 · 0 repositories · arXiv:2406.18122
-
ResumeAtlas: Revisiting Resume Classification with Large-Scale Datasets and Large Language Models 26 Jun 2024 · 1 repository · arXiv:2406.18125
-
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation 26 Jun 2024 · 1 repository · arXiv:2406.18676Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets 26 Jun 2024 · 0 repositories · arXiv:2406.18239
-
SetBERT: Enhancing Retrieval Performance for Boolean Logic and Set Operation Queries 25 Jun 2024 · 0 repositories · arXiv:2406.17282
-
CTBench: A Comprehensive Benchmark for Evaluating Language Model Capabilities in Clinical Trial Design 25 Jun 2024 · 1 repository · arXiv:2406.17888
-
LumberChunker: Long-Form Narrative Document Segmentation 25 Jun 2024 · 1 repository · arXiv:2406.17526Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
RAGBench: Explainable Benchmark for Retrieval-Augmented Generation Systems 25 Jun 2024 · 0 repositories · arXiv:2407.11005
-
This Paper Had the Smartest Reviewers -- Flattery Detection Utilising an Audio-Textual Transformer-Based Approach 25 Jun 2024 · 1 repository · arXiv:2406.17667
-
Unlocking Continual Learning Abilities in Language Models 25 Jun 2024 · 1 repository · arXiv:2406.17245Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Attention Instruction: Amplifying Attention in the Middle via Prompting 24 Jun 2024 · 1 repository · arXiv:2406.17095
-
On the Role of Long-tail Knowledge in Retrieval Augmented Large Language Models 24 Jun 2024 · 0 repositories · arXiv:2406.16367
-
Panza: Design and Analysis of a Fully-Local Personalized Text Writing Assistant 24 Jun 2024 · 1 repository · arXiv:2407.10994
-
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track 24 Jun 2024 · 2 repositories · arXiv:2406.16828Syntology official (archive's flag): 9 ran · 19 ran (of which 0 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 23 harvested samples)
-
MixTex: Unambiguous Recognition Should Not Rely Solely on Real Data 24 Jun 2024 · 1 repository · arXiv:2406.17148
-
Evaluating Ensemble Methods for News Recommender Systems 23 Jun 2024 · 0 repositories · arXiv:2406.16106
-
A multi-speaker multi-lingual voice cloning system based on vits2 for limmits 2024 challenge 22 Jun 2024 · 0 repositories · arXiv:2406.17801
-
A Tale of Trust and Accuracy: Base vs. Instruct LLMs in RAG Systems 21 Jun 2024 · 1 repository · arXiv:2406.14972
-
GiusBERTo: A Legal Language Model for Personal Data De-identification in Italian Court of Auditors Decisions 21 Jun 2024 · 0 repositories · arXiv:2406.15032
-
LongRAG: Enhancing Retrieval-Augmented Generation with Long-context LLMs 21 Jun 2024 · 0 repositories · arXiv:2406.15319
-
Pistis-RAG: Enhancing Retrieval-Augmented Generation with Human Feedback 21 Jun 2024 · 0 repositories · arXiv:2407.00072
-
QPaug: Question and Passage Augmentation for Open-Domain Question Answering of LLMs 20 Jun 2024 · 1 repository · arXiv:2406.14277Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
CodeRAG-Bench: Can Retrieval Augment Code Generation? 20 Jun 2024 · 1 repository · arXiv:2406.14497Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
DIRAS: Efficient LLM Annotation of Document Relevance in Retrieval Augmented Generation 20 Jun 2024 · 1 repository · arXiv:2406.14162
-
Evaluating RAG-Fusion with RAGElo: an Automated Elo-based Framework 20 Jun 2024 · 1 repository · arXiv:2406.14783Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Healing Powers of BERT: How Task-Specific Fine-Tuning Recovers Corrupted Language Models 20 Jun 2024 · 0 repositories · arXiv:2406.14459
-
Relation Extraction with Fine-Tuned Large Language Models in Retrieval Augmented Generation Frameworks 20 Jun 2024 · 0 repositories · arXiv:2406.14745
-
Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More? 19 Jun 2024 · 1 repository · arXiv:2406.13121Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Fine-Tuning BERTs for Definition Extraction from Mathematical Text 19 Jun 2024 · 0 repositories · arXiv:2406.13827
-
FoRAG: Factuality-optimized Retrieval Augmented Generation for Web-enhanced Long-form Question Answering 19 Jun 2024 · 0 repositories · arXiv:2406.13779
-
InstructRAG: Instructing Retrieval-Augmented Generation via Self-Synthesized Rationales 19 Jun 2024 · 1 repository · arXiv:2406.13629Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Model Internals-based Answer Attribution for Trustworthy Retrieval-Augmented Generation 19 Jun 2024 · 1 repository · arXiv:2406.13663Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multi-Meta-RAG: Improving RAG for Multi-Hop Queries using Database Filtering with LLM-Extracted Metadata 19 Jun 2024 · 1 repository · arXiv:2406.13213
-
R^2AG: Incorporating Retrieval Information into Retrieval Augmented Generation 19 Jun 2024 · 1 repository · arXiv:2406.13249Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
WikiContradict: A Benchmark for Evaluating LLMs on Real-World Knowledge Conflicts from Wikipedia 19 Jun 2024 · 0 repositories · arXiv:2406.13805
-
From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries 18 Jun 2024 · 0 repositories · arXiv:2406.12824
-
Intermediate Distillation: Data-Efficient Distillation from Black-Box LLMs for Information Retrieval 18 Jun 2024 · 0 repositories · arXiv:2406.12169
-
PlanRAG: A Plan-then-Retrieval Augmented Generation for Generative Large Language Models as Decision Makers 18 Jun 2024 · 1 repository · arXiv:2406.12430Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Retrieval-Augmented Generation for Generative Artificial Intelligence in Medicine 18 Jun 2024 · 0 repositories · arXiv:2406.12449
-
RichRAG: Crafting Rich Responses for Multi-faceted Queries in Retrieval-Augmented Generation 18 Jun 2024 · 0 repositories · arXiv:2406.12566
-
Unified Active Retrieval for Retrieval Augmented Generation 18 Jun 2024 · 1 repository · arXiv:2406.12534
-
What Makes Two Language Models Think Alike? 18 Jun 2024 · 0 repositories · arXiv:2406.12620
-
Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance 17 Jun 2024 · 0 repositories · arXiv:2406.11139
-
CrAM: Credibility-Aware Attention Modification in LLMs for Combating Misinformation in RAG 17 Jun 2024 · 1 repository · arXiv:2406.11497
-
SeRTS: Self-Rewarding Tree Search for Biomedical Retrieval-Augmented Generation 17 Jun 2024 · 0 repositories · arXiv:2406.11258
-
Evaluating the Efficacy of Open-Source LLMs in Enterprise-Specific RAG Systems: A Comparative Study of Performance and Scalability 17 Jun 2024 · 1 repository · arXiv:2406.11424
-
Fine-Tuning or Fine-Failing? Debunking Performance Myths in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11201
-
Iterative Utility Judgment Framework via LLMs Inspired by Relevance in Philosophy 17 Jun 2024 · 0 repositories · arXiv:2406.11290
-
R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11681
-
Satyrn: A Platform for Analytics Augmented Generation 17 Jun 2024 · 1 repository · arXiv:2406.12069
-
Refiner: Restructure Retrieval Content Efficiently to Advance Question-Answering Capabilities 17 Jun 2024 · 1 repository · arXiv:2406.11357Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TRACE the Evidence: Constructing Knowledge-Grounded Reasoning Chains for Retrieval-Augmented Generation 17 Jun 2024 · 2 repositories · arXiv:2406.11460Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions 17 Jun 2024 · 1 repository · arXiv:2406.12058Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Predicting the Understandability of Computational Notebooks through Code Metrics Analysis 16 Jun 2024 · 1 repository · arXiv:2406.10989
-
ShareLoRA: Parameter Efficient and Robust Large Language Model Fine-tuning via Shared Low-Rank Adaptation 16 Jun 2024 · 1 repository · arXiv:2406.10785Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
A Comprehensive Survey of Foundation Models in Medicine 15 Jun 2024 · 0 repositories · arXiv:2406.10729
-
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation 15 Jun 2024 · 0 repositories · arXiv:2406.10561
-
Bag of Lies: Robustness in Continuous Pre-training BERT 14 Jun 2024 · 0 repositories · arXiv:2406.09967
-
HIRO: Hierarchical Information Retrieval Optimization 14 Jun 2024 · 1 repository · arXiv:2406.09979
-
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models 14 Jun 2024 · 1 repository · arXiv:2406.10130
-
Analyzing Gender Polarity in Short Social Media Texts with BERT: The Role of Emojis and Emoticons 13 Jun 2024 · 0 repositories · arXiv:2406.09573
-
BPE-knockout: Pruning Pre-existing BPE Tokenisers with Backwards-compatible Morphological Semi-supervision 13 Jun 2024 · 1 repository
-
PC-LoRA: Low-Rank Adaptation for Progressive Model Compression with Knowledge Distillation 13 Jun 2024 · 0 repositories · arXiv:2406.09117
-
Ad Auctions for LLMs via Retrieval Augmented Generation 12 Jun 2024 · 0 repositories · arXiv:2406.09459
-
Exploring Fact Memorization and Style Imitation in LLMs Using QLoRA: An Experimental Study and Quality Assessment Methods 12 Jun 2024 · 0 repositories · arXiv:2406.08582
-
Label-aware Hard Negative Sampling Strategies with Momentum Contrastive Learning for Implicit Hate Speech Detection 12 Jun 2024 · 1 repository · arXiv:2406.07886
-
Leveraging Large Language Models for Web Scraping 12 Jun 2024 · 0 repositories · arXiv:2406.08246
-
Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation 12 Jun 2024 · 0 repositories · arXiv:2406.08328
-
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis 11 Jun 2024 · 0 repositories · arXiv:2406.10273
-
COVID-19 Twitter Sentiment Classification Using Hybrid Deep Learning Model Based on Grid Search Methodology 11 Jun 2024 · 0 repositories · arXiv:2406.10266