Methods › General › Regularization › Attention Dropout › Papers, page 13
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 13 of 109: papers 1,201 to 1,300 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RAG with Differential Privacy 26 Dec 2024 · 1 repository · arXiv:2412.19291
-
Sentiment trading with large language models 26 Dec 2024 · 0 repositories · arXiv:2412.19245
-
HAND: Hierarchical Attention Network for Multi-Scale Handwritten Document Recognition and Layout Analysis 25 Dec 2024 · 0 repositories · arXiv:2412.18981
-
Injecting Bias into Text Classification Models using Backdoor Attacks 25 Dec 2024 · 0 repositories · arXiv:2412.18975
-
Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks 25 Dec 2024 · 0 repositories · arXiv:2412.18729
-
Resource-Efficient Transformer Architecture: Optimizing Memory and Execution Time for Real-Time Applications 25 Dec 2024 · 0 repositories · arXiv:2501.00042
-
SAFLITE: Fuzzing Autonomous Systems via Large Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18727
-
Whose Morality Do They Speak? Unraveling Cultural Bias in Multilingual Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18863
-
Advancing Explainability in Neural Machine Translation: Analytical Metrics for Attention and Alignment Consistency 24 Dec 2024 · 0 repositories · arXiv:2412.18669
-
Comprehensive Assessment of BERT-Based Methods for Predicting Antimicrobial Peptides 24 Dec 2024 · 1 repository
-
Do Language Models Understand the Cognitive Tasks Given to Them? Investigations with the N-Back Paradigm 24 Dec 2024 · 0 repositories · arXiv:2412.18120
-
GeAR: Graph-enhanced Agent for Retrieval-augmented Generation 24 Dec 2024 · 0 repositories · arXiv:2412.18431
-
HTR-JAND: Handwritten Text Recognition with Joint Attention Network and Knowledge Distillation 24 Dec 2024 · 0 repositories · arXiv:2412.18524
-
Improving Factuality with Explicit Working Memory 24 Dec 2024 · 0 repositories · arXiv:2412.18069
-
Molly: Making Large Language Model Agents Solve Python Problem More Logically 24 Dec 2024 · 0 repositories · arXiv:2412.18093
-
Pirates of the RAG: Adaptively Attacking LLMs to Leak Knowledge Bases 24 Dec 2024 · 0 repositories · arXiv:2412.18295
-
Research on the Proximity Relationships of Psychosomatic Disease Knowledge Graph Modules Extracted by Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18419
-
Unlocking the Potential of Multiple BERT Models for Bangla Question Answering in NCTB Textbooks 24 Dec 2024 · 0 repositories · arXiv:2412.18440
-
A Survey of Query Optimization in Large Language Models 23 Dec 2024 · 0 repositories · arXiv:2412.17558
-
Comparative Analysis of Document-Level Embedding Methods for Similarity Scoring on Shakespeare Sonnets and Taylor Swift Lyrics 23 Dec 2024 · 0 repositories · arXiv:2412.17552
-
Efficient fine-tuning methodology of text embedding models for information retrieval: contrastive learning penalty (clp) 23 Dec 2024 · 1 repository · arXiv:2412.17364
-
A Reality Check on Context Utilisation for Retrieval-Augmented Generation 22 Dec 2024 · 1 repository · arXiv:2412.17031
-
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models 22 Dec 2024 · 0 repositories · arXiv:2501.03246
-
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora 22 Dec 2024 · 0 repositories · arXiv:2412.16976
-
PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental Health 22 Dec 2024 · 1 repository · arXiv:2412.16882
-
Robustness of Large Language Models Against Adversarial Attacks 22 Dec 2024 · 0 repositories · arXiv:2412.17011
-
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction 22 Dec 2024 · 0 repositories · arXiv:2412.16919
-
AlzheimerRAG: Multimodal Retrieval Augmented Generation for PubMed articles 21 Dec 2024 · 0 repositories · arXiv:2412.16701
-
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans? 21 Dec 2024 · 0 repositories · arXiv:2412.16772
-
Distilling Large Language Models for Efficient Clinical Information Extraction 21 Dec 2024 · 0 repositories · arXiv:2501.00031
-
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16486
-
Formal Language Knowledge Corpus for Retrieval Augmented Generation 21 Dec 2024 · 0 repositories · arXiv:2412.16689
-
Identifying Cyberbullying Roles in Social Media 21 Dec 2024 · 0 repositories · arXiv:2412.16417
-
Improving FIM Code Completions via Context & Curriculum Based Learning 21 Dec 2024 · 0 repositories · arXiv:2412.16589
-
Quantum-Like Contextuality in Large Language Models 21 Dec 2024 · 1 repository · arXiv:2412.16806
-
Research on Violent Text Detection System Based on BERT-fasttext Model 21 Dec 2024 · 0 repositories · arXiv:2412.16455
-
TimeRAG: BOOSTING LLM Time Series Forecasting via Retrieval-Augmented Generation 21 Dec 2024 · 0 repositories · arXiv:2412.16643
-
Towards More Robust Retrieval-Augmented Generation: Evaluating RAG Under Adversarial Poisoning Attacks 21 Dec 2024 · 1 repository · arXiv:2412.16708
-
Adversarial Robustness through Dynamic Ensemble Learning 20 Dec 2024 · 0 repositories · arXiv:2412.16254
-
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation 20 Dec 2024 · 0 repositories · arXiv:2412.16135
-
Decoding Linguistic Nuances in Mental Health Text Classification Using Expressive Narrative Stories 20 Dec 2024 · 0 repositories · arXiv:2412.16302
-
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks 20 Dec 2024 · 1 repository · arXiv:2412.15605Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context 20 Dec 2024 · 0 repositories · arXiv:2412.16359
-
HybGRAG: Hybrid Retrieval-Augmented Generation on Textual and Relational Knowledge Bases 20 Dec 2024 · 0 repositories · arXiv:2412.16311
-
Linguistic Features Extracted by GPT-4 Improve Alzheimer's Disease Detection based on Spontaneous Speech 20 Dec 2024 · 1 repository · arXiv:2412.15772
-
The First Multilingual Model For The Detection of Suicide Texts 20 Dec 2024 · 0 repositories · arXiv:2412.15498
-
Towards Interpretable Radiology Report Generation via Concept Bottlenecks using a Multi-Agentic RAG 20 Dec 2024 · 1 repository · arXiv:2412.16086
-
XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation 20 Dec 2024 · 1 repository · arXiv:2412.15529Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Analysis and Visualization of Linguistic Structures in Large Language Models: Neural Representations of Verb-Particle Constructions in BERT 19 Dec 2024 · 0 repositories · arXiv:2412.14670
-
CORD: Balancing COnsistency and Rank Distillation for Robust Retrieval-Augmented Generation 19 Dec 2024 · 0 repositories · arXiv:2412.14581
-
Decade of Natural Language Processing in Chronic Pain: A Systematic Review 19 Dec 2024 · 0 repositories · arXiv:2412.15360
-
Dehallucinating Parallel Context Extension for Retrieval-Augmented Generation 19 Dec 2024 · 0 repositories · arXiv:2412.14905
-
DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs 19 Dec 2024 · 0 repositories · arXiv:2412.14838
-
Graph-Convolutional Networks: Named Entity Recognition and Large Language Model Embedding in Document Clustering 19 Dec 2024 · 0 repositories · arXiv:2412.14867
-
How good is GPT at writing political speeches for the White House? 19 Dec 2024 · 0 repositories · arXiv:2412.14617
-
Knowledge Injection via Prompt Distillation 19 Dec 2024 · 0 repositories · arXiv:2412.14964
-
LLMs as mediators: Can they diagnose conflicts accurately? 19 Dec 2024 · 0 repositories · arXiv:2412.14675
-
PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization 19 Dec 2024 · 1 repository · arXiv:2412.14510
-
Query pipeline optimization for cancer patient question answering systems 19 Dec 2024 · 0 repositories · arXiv:2412.14751
-
Relational Programming with Foundation Models 19 Dec 2024 · 0 repositories · arXiv:2412.14515
-
ResoFilter: Fine-grained Synthetic Data Filtering for Large Language Models through Data-Parameter Resonance Analysis 19 Dec 2024 · 1 repository · arXiv:2412.14809
-
Review-Then-Refine: A Dynamic Framework for Multi-Hop Question Answering with Temporal Adaptability 19 Dec 2024 · 0 repositories · arXiv:2412.15101
-
SKETCH: Structured Knowledge Enhanced Text Comprehension for Holistic Retrieval 19 Dec 2024 · 0 repositories · arXiv:2412.15443
-
Till the Layers Collapse: Compressing a Deep Neural Network through the Lenses of Batch Normalization Layers 19 Dec 2024 · 1 repository · arXiv:2412.15077
-
TOMG-Bench: Evaluating LLMs on Text-based Open Molecule Generation 19 Dec 2024 · 1 repository · arXiv:2412.14642
-
VISA: Retrieval Augmented Generation with Visual Source Attribution 19 Dec 2024 · 0 repositories · arXiv:2412.14457
-
Autonomous Microscopy Experiments through Large Language Model Agents 18 Dec 2024 · 1 repository · arXiv:2501.10385
-
Enhancing Rhetorical Figure Annotation: An Ontology-Based Web Application with RAG Integration 18 Dec 2024 · 1 repository · arXiv:2412.13799
-
EvoWiki: Evaluating LLMs on Evolving Knowledge 18 Dec 2024 · 0 repositories · arXiv:2412.13582
-
FarExStance: Explainable Stance Detection for Farsi 18 Dec 2024 · 2 repositories · arXiv:2412.14008
-
Federated Learning and RAG Integration: A Scalable Approach for Medical Large Language Models 18 Dec 2024 · 0 repositories · arXiv:2412.13720
-
Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN 18 Dec 2024 · 1 repository · arXiv:2412.13795
-
RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment 18 Dec 2024 · 1 repository · arXiv:2412.13746
-
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference 18 Dec 2024 · 2 repositories · arXiv:2412.13663Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
A MapReduce Approach to Effectively Utilize Long Context Information in Retrieval Augmented Language Models 17 Dec 2024 · 0 repositories · arXiv:2412.15271
-
Adaptations of AI models for querying the LandMatrix database in natural language 17 Dec 2024 · 1 repository · arXiv:2412.12961
-
C-FedRAG: A Confidential Federated Retrieval-Augmented Generation System 17 Dec 2024 · 0 repositories · arXiv:2412.13163
-
Chinese SafetyQA: A Safety Short-form Factuality Benchmark for Large Language Models 17 Dec 2024 · 0 repositories · arXiv:2412.15265
-
Detecting Document-level Paraphrased Machine Generated Content: Mimicking Human Writing Style and Involving Discourse Features 17 Dec 2024 · 0 repositories · arXiv:2412.12679
-
EXIT: Context-Aware Extractive Compression for Enhancing Retrieval-Augmented Generation 17 Dec 2024 · 1 repository · arXiv:2412.12559
-
LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning 17 Dec 2024 · 0 repositories · arXiv:2412.12541
-
LLMs are Also Effective Embedding Models: An In-depth Overview 17 Dec 2024 · 0 repositories · arXiv:2412.12591
-
OmniEval: An Omnidirectional and Automatic RAG Evaluation Benchmark in Financial Domain 17 Dec 2024 · 1 repository · arXiv:2412.13018
-
PERC: Plan-As-Query Example Retrieval for Underrepresented Code Generation 17 Dec 2024 · 0 repositories · arXiv:2412.12447
-
RAG-Star: Enhancing Deliberative Reasoning with Retrieval Augmented Verification and Refinement 17 Dec 2024 · 0 repositories · arXiv:2412.12881
-
RemoteRAG: A Privacy-Preserving LLM Cloud RAG Service 17 Dec 2024 · 0 repositories · arXiv:2412.12775
-
SimGRAG: Leveraging Similar Subgraphs for Knowledge Graphs Driven Retrieval-Augmented Generation 17 Dec 2024 · 1 repository · arXiv:2412.15272
-
What External Knowledge is Preferred by LLMs? Characterizing and Exploring Chain of Evidence in Imperfect Context 17 Dec 2024 · 0 repositories · arXiv:2412.12632
-
A Benchmark and Robustness Study of In-Context-Learning with Large Language Models in Music Entity Detection 16 Dec 2024 · 1 repository · arXiv:2412.11851
-
BioBridge: Unified Bio-Embedding with Bridging Modality in Code-Switched EMR 16 Dec 2024 · 1 repository · arXiv:2412.11671
-
Causal Diffusion Transformers for Generative Modeling 16 Dec 2024 · 1 repository · arXiv:2412.12095Syntology official (archive's flag): 7 ran · 8 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Glimpse: Enabling White-Box Methods to Use Proprietary Models for Zero-Shot LLM-Generated Text Detection 16 Dec 2024 · 1 repository · arXiv:2412.11506Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Graph-Guided Textual Explanation Generation Framework 16 Dec 2024 · 0 repositories · arXiv:2412.12318
-
Investigating Mixture of Experts in Dense Retrieval 16 Dec 2024 · 0 repositories · arXiv:2412.11864
-
Look Ahead Text Understanding and LLM Stitching 16 Dec 2024 · 1 repository · arXiv:2412.17836
-
No More Adam: Learning Rate Scaling at Initialization is All You Need 16 Dec 2024 · 1 repository · arXiv:2412.11768
-
Optimized Quran Passage Retrieval Using an Expanded QA Dataset and Fine-Tuned Language Models 16 Dec 2024 · 0 repositories · arXiv:2412.11431
-
Priority-Aware Model-Distributed Inference at Edge Networks 16 Dec 2024 · 0 repositories · arXiv:2412.12371
-
RAG Playground: A Framework for Systematic Evaluation of Retrieval Strategies and Prompt Engineering in RAG Systems 16 Dec 2024 · 1 repository · arXiv:2412.12322
-
Unanswerability Evaluation for Retrieval Augmented Generation 16 Dec 2024 · 0 repositories · arXiv:2412.12300