Methods › General › Regularization › Attention Dropout › Papers, page 2
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 2 of 109: papers 101 to 200 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning 5 Jun 2025 · 0 repositories · arXiv:2506.04527
-
Knowledgeable-r1: Policy Optimization for Knowledge Exploration in Retrieval-Augmented Generation 5 Jun 2025 · 1 repository · arXiv:2506.05154Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
LotusFilter: Fast Diverse Nearest Neighbor Search via a Learned Cutoff Table 5 Jun 2025 · 1 repository · arXiv:2506.04790Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Mathematical Reasoning for Unmanned Aerial Vehicles: A RAG-Based Approach for Complex Arithmetic Reasoning 5 Jun 2025 · 1 repository · arXiv:2506.04998
-
Micro-Act: Mitigate Knowledge Conflict in Question Answering via Actionable Self-Reasoning 5 Jun 2025 · 1 repository · arXiv:2506.05278
-
Multiple-Choice Question Generation Using Large Language Models: Methodology and Educator Insights 5 Jun 2025 · 0 repositories · arXiv:2506.04851
-
On Automating Security Policies with Contemporary LLMs 5 Jun 2025 · 0 repositories · arXiv:2506.04838
-
Predicting ICU In-Hospital Mortality Using Adaptive Transformer Layer Fusion 5 Jun 2025 · 1 repository · arXiv:2506.04924
-
The NTNU System at the S&I Challenge 2025 SLA Open Track 5 Jun 2025 · 0 repositories · arXiv:2506.05121
-
Facts are Harder Than Opinions -- A Multilingual, Comparative Analysis of LLM-Based Fact-Checking Reliability 4 Jun 2025 · 0 repositories · arXiv:2506.03655
-
Magic Mushroom: A Customizable Benchmark for Fine-grained Analysis of Retrieval Noise Erosion in RAG Systems 4 Jun 2025 · 0 repositories · arXiv:2506.03901
-
Privacy and Security Threat for OpenAI GPTs 4 Jun 2025 · 0 repositories · arXiv:2506.04036
-
R-Search: Empowering LLM Reasoning with Search via Multi-Reward Reinforcement Learning 4 Jun 2025 · 1 repository · arXiv:2506.04185
-
TracLLM: A Generic Framework for Attributing Long Context LLMs 4 Jun 2025 · 1 repository · arXiv:2506.04202
-
Retrieval-Augmented Generation as Noisy In-Context Learning: A Unified Theory and Risk Bounds 3 Jun 2025 · 0 repositories · arXiv:2506.03100
-
An Exploratory Framework for Future SETI Applications: Detecting Generative Reactivity via Language Models 3 Jun 2025 · 0 repositories · arXiv:2506.02730
-
CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG 3 Jun 2025 · 0 repositories · arXiv:2506.02544
-
Enhancing Automatic PT Tagging for MEDLINE Citations Using Transformer-Based Models 3 Jun 2025 · 0 repositories · arXiv:2506.03321
-
Rethinking the effects of data contamination in Code Intelligence 3 Jun 2025 · 0 repositories · arXiv:2506.02791
-
Hybrid AI for Responsive Multi-Turn Online Conversations with Novel Dynamic Routing and Feedback Adaptation 2 Jun 2025 · 0 repositories · arXiv:2506.02097
-
LLMs as World Models: Data-Driven and Human-Centered Pre-Event Simulation for Disaster Impact Assessment 2 Jun 2025 · 0 repositories · arXiv:2506.06355
-
Retrieval-Augmented Generation of Ontologies from Relational Databases 2 Jun 2025 · 0 repositories · arXiv:2506.01232
-
A Graph-Retrieval-Augmented Generation Framework Enhances Decision-Making in the Circular Economy 1 Jun 2025 · 0 repositories · arXiv:2506.04252
-
How Neural Networks Organize Concepts: Introducing Concept Trajectory Analysis for Deep Learning Interpretability 1 Jun 2025 · 1 repository
-
RARE: Retrieval-Aware Robustness Evaluation for Retrieval-Augmented Generation Systems 1 Jun 2025 · 1 repository · arXiv:2506.00789
-
FinBERT2: A Specialized Bidirectional Encoder for Bridging the Gap in Finance-Specific Deployment of Large Language Models 31 May 2025 · 0 repositories · arXiv:2506.06335
-
Power-of-Two (PoT) Weights in Large Language Models (LLMs) 31 May 2025 · 0 repositories · arXiv:2506.00315
-
Adversarial Threat Vectors and Risk Mitigation for Retrieval-Augmented Generation Systems 30 May 2025 · 0 repositories · arXiv:2506.00281
-
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks 30 May 2025 · 1 repository · arXiv:2505.24876
-
ClueAnchor: Clue-Anchored Knowledge Reasoning Exploration and Optimization for Retrieval-Augmented Generation 30 May 2025 · 1 repository · arXiv:2505.24388
-
E^2GraphRAG: Streamlining Graph-based RAG for High Efficiency and Effectiveness 30 May 2025 · 0 repositories · arXiv:2505.24226
-
Interpretable phenotyping of Heart Failure patients with Dutch discharge letters 30 May 2025 · 0 repositories · arXiv:2505.24619
-
LPASS: Linear Probes as Stepping Stones for vulnerability detection using compressed LLMs 30 May 2025 · 0 repositories · arXiv:2505.24451
-
MOFGPT: Generative Design of Metal-Organic Frameworks using Language Models 30 May 2025 · 1 repository · arXiv:2506.00198
-
RealDrive: Retrieval-Augmented Driving with Diffusion Models 30 May 2025 · 0 repositories · arXiv:2505.24808
-
When GPT Spills the Tea: Comprehensive Assessment of Knowledge File Leakage in GPTs 30 May 2025 · 0 repositories · arXiv:2506.00197
-
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning 29 May 2025 · 0 repositories · arXiv:2505.23298
-
Critical Batch Size Revisited: A Simple Empirical Approach to Large-Batch Language Model Training 29 May 2025 · 0 repositories · arXiv:2505.23971
-
Data-efficient Meta-models for Evaluation of Context-based Questions and Answers in LLMs 29 May 2025 · 0 repositories · arXiv:2505.23299
-
Daunce: Data Attribution through Uncertainty Estimation 29 May 2025 · 0 repositories · arXiv:2505.23223
-
Decom-Renorm-Merge: Model Merging on the Right Space Improves Multitasking 29 May 2025 · 0 repositories · arXiv:2505.23117
-
Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach 29 May 2025 · 0 repositories · arXiv:2505.23953
-
Evaluating AI capabilities in detecting conspiracy theories on YouTube 29 May 2025 · 1 repository · arXiv:2505.23570
-
Matryoshka Model Learning for Improved Elastic Student Models 29 May 2025 · 0 repositories · arXiv:2505.23337
-
MCP Safety Training: Learning to Refuse Falsely Benign MCP Exploits using Improved Preference Alignment 29 May 2025 · 0 repositories · arXiv:2505.23634
-
mRAG: Elucidating the Design Space of Multi-modal Retrieval-Augmented Generation 29 May 2025 · 0 repositories · arXiv:2505.24073
-
Query Routing for Retrieval-Augmented Language Models 29 May 2025 · 0 repositories · arXiv:2505.23052
-
Reducing Latency in LLM-Based Natural Language Commands Processing for Robot Navigation 29 May 2025 · 0 repositories · arXiv:2506.00075
-
Agent-UniRAG: A Trainable Open-Source LLM Agent Framework for Unified Retrieval-Augmented Generation Systems 28 May 2025 · 0 repositories · arXiv:2505.22571
-
Breaking the Cloak! Unveiling Chinese Cloaked Toxicity with Homophone Graph and Toxic Lexicon 28 May 2025 · 0 repositories · arXiv:2505.22184
-
Climate Finance Bench 28 May 2025 · 1 repository · arXiv:2505.22752
-
Contextual Memory Intelligence -- A Foundational Paradigm for Human-AI Collaboration and Reflective Generative AI Systems 28 May 2025 · 0 repositories · arXiv:2506.05370
-
Cross-modal RAG: Sub-dimensional Retrieval-Augmented Text-to-Image Generation 28 May 2025 · 1 repository · arXiv:2505.21956
-
Improving QA Efficiency with DistilBERT: Fine-Tuning and Inference on mobile Intel CPUs 28 May 2025 · 0 repositories · arXiv:2505.22937
-
Multi-MLLM Knowledge Distillation for Out-of-Context News Detection 28 May 2025 · 0 repositories · arXiv:2505.22517
-
RAGPPI: RAG Benchmark for Protein-Protein Interactions in Drug Discovery 28 May 2025 · 1 repository · arXiv:2505.23823
-
Say What You Mean: Natural Language Access Control with Large Language Models for Internet of Things 28 May 2025 · 0 repositories · arXiv:2505.23835
-
SkewRoute: Training-Free LLM Routing for Knowledge Graph Retrieval-Augmented Generation via Score Skewness of Retrieved Context 28 May 2025 · 0 repositories · arXiv:2505.23841
-
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning 28 May 2025 · 1 repository · arXiv:2505.22019Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Diagnosing and Resolving Cloud Platform Instability with Multi-modal RAG LLMs 27 May 2025 · 0 repositories · arXiv:2505.21419
-
Explainability of Large Language Models using SMILE: Statistical Model-agnostic Interpretability with Local Explanations 27 May 2025 · 1 repository · arXiv:2505.21657
-
From prosthetic memory to prosthetic denial: Auditing whether large language models are prone to mass atrocity denialism 27 May 2025 · 0 repositories · arXiv:2505.21753
-
Long Context Scaling: Divide and Conquer via Multi-Agent Question-driven Collaboration 27 May 2025 · 0 repositories · arXiv:2505.20625
-
Privacy-Preserving Chest X-ray Report Generation via Multimodal Federated Learning with ViT and GPT-2 27 May 2025 · 0 repositories · arXiv:2505.21715
-
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents 26 May 2025 · 0 repositories · arXiv:2505.19494
-
Automated evaluation of children's speech fluency for low-resource languages 26 May 2025 · 0 repositories · arXiv:2505.19671
-
Benchmarking Multimodal Knowledge Conflict for Large Multimodal Models 26 May 2025 · 1 repository · arXiv:2505.19509
-
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages 26 May 2025 · 0 repositories · arXiv:2505.19851
-
Calibrating Pre-trained Language Classifiers on LLM-generated Noisy Labels via Iterative Refinement 26 May 2025 · 1 repository · arXiv:2505.19675
-
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language 26 May 2025 · 0 repositories · arXiv:2505.19971
-
Detection of Suicidal Risk on Social Media: A Hybrid Model 26 May 2025 · 0 repositories · arXiv:2505.23797
-
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems 26 May 2025 · 0 repositories · arXiv:2505.19847
-
DoctorRAG: Medical RAG Fusing Knowledge with Patient Analogy through Textual Gradients 26 May 2025 · 0 repositories · arXiv:2505.19538
-
Emotion Classification In-Context in Spanish 26 May 2025 · 0 repositories · arXiv:2505.20571
-
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining 26 May 2025 · 0 repositories · arXiv:2505.19893
-
KnowTrace: Bootstrapping Iterative Retrieval-Augmented Generation with Structured Knowledge Tracing 26 May 2025 · 1 repository · arXiv:2505.20245
-
MA-RAG: Multi-Agent Retrieval-Augmented Generation via Collaborative Chain-of-Thought Reasoning 26 May 2025 · 0 repositories · arXiv:2505.20096
-
NeuSym-RAG: Hybrid Neural Symbolic Retrieval with Multiview Structuring for PDF Question Answering 26 May 2025 · 1 repository · arXiv:2505.19754
-
R3-RAG: Learning Step-by-Step Reasoning and Retrieval for LLMs via Reinforcement Learning 26 May 2025 · 1 repository · arXiv:2505.23794
-
syftr: Pareto-Optimal Generative AI 26 May 2025 · 1 repository · arXiv:2505.20266
-
AI4Math: A Native Spanish Benchmark for University-Level Mathematical Reasoning in Large Language Models 25 May 2025 · 0 repositories · arXiv:2505.18978
-
Conventional Contrastive Learning Often Falls Short: Improving Dense Retrieval with Cross-Encoder Listwise Distillation and Synthetic Data 25 May 2025 · 1 repository · arXiv:2505.19274
-
Hermes@DravidianLangTech 2025: Sentiment Analysis of Dravidian Languages using XLM-RoBERTa 25 May 2025 · 1 repository
-
Hypercube-RAG: Hypercube-Based Retrieval-Augmented Generation for In-domain Scientific Question-Answering 25 May 2025 · 1 repository · arXiv:2505.19288
-
Investigating Pedagogical Teacher and Student LLM Agents: Genetic Adaptation Meets Retrieval Augmented Generation Across Learning Style 25 May 2025 · 0 repositories · arXiv:2505.19173
-
Optimized Text Embedding Models and Benchmarks for Amharic Passage Retrieval 25 May 2025 · 1 repository · arXiv:2505.19356
-
POQD: Performance-Oriented Query Decomposer for Multi-vector retrieval 25 May 2025 · 1 repository · arXiv:2505.19189Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Retrieval-Augmented Generation for Service Discovery: Chunking Strategies and Benchmarking 25 May 2025 · 0 repositories · arXiv:2505.19310
-
A Survey of LLM × DATA 24 May 2025 · 2 repositories · arXiv:2505.18458
-
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation 24 May 2025 · 0 repositories · arXiv:2505.18543
-
BRIT: Bidirectional Retrieval over Unified Image-Text Graph 24 May 2025 · 0 repositories · arXiv:2505.18450
-
Federated Retrieval-Augmented Generation: A Systematic Mapping Study 24 May 2025 · 0 repositories · arXiv:2505.18906
-
From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data 24 May 2025 · 0 repositories · arXiv:2505.18464
-
GainRAG: Preference Alignment in Retrieval-Augmented Generation through Gain Signal Synthesis 24 May 2025 · 1 repository · arXiv:2505.18710
-
Is Attention Required for Transformer Inference? Explore Function-preserving Attention Replacement 24 May 2025 · 0 repositories · arXiv:2505.21535
-
LLMs for Supply Chain Management 24 May 2025 · 0 repositories · arXiv:2505.18597
-
Pruning for Performance: Efficient Idiom and Metaphor Classification in Low-Resource Konkani Using mBERT 24 May 2025 · 0 repositories · arXiv:2506.02005
-
Removal of Hallucination on Hallucination: Debate-Augmented RAG 24 May 2025 · 1 repository · arXiv:2505.18581Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
Strong Membership Inference Attacks on Massive Datasets and (Moderately) Large Language Models 24 May 2025 · 0 repositories · arXiv:2505.18773
-
The Silent Saboteur: Imperceptible Adversarial Attacks against Black-Box Retrieval-Augmented Generation Systems 24 May 2025 · 0 repositories · arXiv:2505.18583