Methods › General › Regularization › Attention Dropout › Papers, page 10
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 10 of 109: papers 901 to 1,000 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Dataset Protection via Watermarked Canaries in Retrieval-Augmented LLMs 15 Feb 2025 · 0 repositories · arXiv:2502.10673
-
Evolving Hate Speech Online: An Adaptive Framework for Detection and Mitigation 15 Feb 2025 · 0 repositories · arXiv:2502.10921
-
CAE-Net: Generalized Deepfake Image Detection using Convolution and Attention Mechanisms with Spatial and Frequency Domain Features 15 Feb 2025 · 0 repositories · arXiv:2502.10682
-
NitiBench: A Comprehensive Studies of LLM Frameworks Capabilities for Thai Legal Question Answering 15 Feb 2025 · 1 repository · arXiv:2502.10868
-
Order-agnostic Identifier for Large Language Model-based Generative Recommendation 15 Feb 2025 · 0 repositories · arXiv:2502.10833
-
The underlying structures of self-attention: symmetry, directionality, and emergent dynamics in Transformer training 15 Feb 2025 · 1 repository · arXiv:2502.10927
-
ArchRAG: Attributed Community-based Hierarchical Retrieval-Augmented Generation 14 Feb 2025 · 0 repositories · arXiv:2502.09891
-
Do Large Language Models Reason Causally Like Us? Even Better? 14 Feb 2025 · 0 repositories · arXiv:2502.10215
-
EmbBERT-Q: Breaking Memory Barriers in Embedded NLP 14 Feb 2025 · 0 repositories · arXiv:2502.10001
-
Hallucinations and Truth: A Comprehensive Accuracy Evaluation of RAG, LoRA and DoRA 14 Feb 2025 · 0 repositories · arXiv:2502.10497
-
LaRA: Benchmarking Retrieval-Augmented Generation and Long-Context LLMs - No Silver Bullet for LC or RAG Routing 14 Feb 2025 · 0 repositories · arXiv:2502.09977
-
Post-training an LLM for RAG? Train on Self-Generated Demonstrations 14 Feb 2025 · 0 repositories · arXiv:2502.10596
-
Enhancing RAG with Active Learning on Conversation Records: Reject Incapables and Answer Capables 13 Feb 2025 · 0 repositories · arXiv:2502.09073
-
KIMAs: A Configurable Knowledge Integrated Multi-Agent System 13 Feb 2025 · 0 repositories · arXiv:2502.09596
-
Mechanistic Unveiling of Transformer Circuits: Self-Influence as a Key to Model Reasoning 13 Feb 2025 · 0 repositories · arXiv:2502.09022
-
Predicting Cognitive Decline: A Multimodal AI Approach to Dementia Screening from Speech 13 Feb 2025 · 0 repositories · arXiv:2502.08862
-
Can Vision-Language Models Infer Speaker's Ignorance? The Role of Visual and Linguistic Cues 13 Feb 2025 · 0 repositories · arXiv:2502.09120
-
Utilizing Pre-trained and Large Language Models for 10-K Items Segmentation 13 Feb 2025 · 0 repositories · arXiv:2502.08875
-
Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation 12 Feb 2025 · 1 repository · arXiv:2502.08826
-
InTAR: Inter-Task Auto-Reconfigurable Accelerator Design for High Data Volume Variation in DNNs 12 Feb 2025 · 1 repository · arXiv:2502.08807
-
ParetoRAG: Leveraging Sentence-Context Attention for Robust and Efficient Retrieval-Augmented Generation 12 Feb 2025 · 0 repositories · arXiv:2502.08178
-
Systematic Knowledge Injection into Large Language Models via Diverse Augmentation for Domain-Specific RAG 12 Feb 2025 · 1 repository · arXiv:2502.08356Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation 12 Feb 2025 · 0 repositories · arXiv:2502.10467
-
A Large-Scale Benchmark for Vietnamese Sentence Paraphrases 11 Feb 2025 · 1 repository · arXiv:2502.07188
-
An Advanced NLP Framework for Automated Medical Diagnosis with DeBERTa and Dynamic Contextual Positional Gating 11 Feb 2025 · 0 repositories · arXiv:2502.07755
-
Automated Capability Discovery via Model Self-Exploration 11 Feb 2025 · 2 repositories · arXiv:2502.07577
-
FoQA: A Faroese Question-Answering Dataset 11 Feb 2025 · 0 repositories · arXiv:2502.07642
-
Grammar Control in Dialogue Response Generation for Language Learning Chatbots 11 Feb 2025 · 1 repository · arXiv:2502.07544
-
Graph RAG-Tool Fusion 11 Feb 2025 · 1 repository · arXiv:2502.07223
-
Making Language Models Robust Against Negation 11 Feb 2025 · 0 repositories · arXiv:2502.07717
-
Tractable Transformers for Flexible Conditional Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07616
-
C-3PO: Compact Plug-and-Play Proxy Optimization to Achieve Human-like Retrieval-Augmented Generation 10 Feb 2025 · 0 repositories · arXiv:2502.06205Syntology 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
ConMeC: A Dataset for Metonymy Resolution with Common Nouns 10 Feb 2025 · 1 repository · arXiv:2502.06087
-
DebateBench: A Challenging Long Context Reasoning Benchmark For Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06279
-
Find Central Dogma Again: Leveraging Multilingual Transfer in Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06253
-
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights 10 Feb 2025 · 0 repositories · arXiv:2502.07049
-
Leveraging GPT-4o Efficiency for Detecting Rework Anomaly in Business Processes 10 Feb 2025 · 0 repositories · arXiv:2502.06918
-
Multimodal Task Representation Memory Bank vs. Catastrophic Forgetting in Anomaly Detection 10 Feb 2025 · 0 repositories · arXiv:2502.06194
-
Optimizing Knowledge Integration in Retrieval-Augmented Generation with Self-Selection 10 Feb 2025 · 0 repositories · arXiv:2502.06148
-
RALLRec: Improving Retrieval Augmented Large Language Model Recommendation with Representation Learning 10 Feb 2025 · 1 repository · arXiv:2502.06101
-
Towards Copyright Protection for Knowledge Bases of Retrieval-augmented Language Models via Reasoning 10 Feb 2025 · 0 repositories · arXiv:2502.10440
-
Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models 9 Feb 2025 · 0 repositories · arXiv:2502.06039
-
Emergence of Episodic Memory in Transformers: Characterizing Changes in Temporal Structure of Attention Scores During Training 9 Feb 2025 · 0 repositories · arXiv:2502.06902
-
Enhancing Financial Time-Series Forecasting with Retrieval-Augmented Large Language Models 9 Feb 2025 · 0 repositories · arXiv:2502.05878
-
Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End" 9 Feb 2025 · 0 repositories · arXiv:2502.06898
-
ScaffoldGPT: A Scaffold-based GPT Model for Drug Optimization 9 Feb 2025 · 0 repositories · arXiv:2502.06891
-
APE: Faster and Longer Context-Augmented Generation via Adaptive Parallel Encoding 8 Feb 2025 · 1 repository · arXiv:2502.05431Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Forbidden Science: Dual-Use AI Challenge Benchmark and Scientific Refusal Tests 8 Feb 2025 · 0 repositories · arXiv:2502.06867
-
Knowledge Graph-Guided Retrieval Augmented Generation 8 Feb 2025 · 1 repository · arXiv:2502.06864
-
On the Effectiveness of Large Language Models in Automating Categorization of Scientific Texts 8 Feb 2025 · 0 repositories · arXiv:2502.15745
-
The Odyssey of the Fittest: Can Agents Survive and Still Be Good? 8 Feb 2025 · 1 repository · arXiv:2502.05442
-
Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey 8 Feb 2025 · 1 repository · arXiv:2502.06872
-
Can Large Language Models Understand Intermediate Representations? 7 Feb 2025 · 0 repositories · arXiv:2502.06854
-
Detection of LLM-Generated Java Code Using Discretized Nested Bigrams 7 Feb 2025 · 0 repositories · arXiv:2502.15740
-
EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification 7 Feb 2025 · 0 repositories · arXiv:2502.06852
-
Efficient Knowledge Feeding to Language Models: A Novel Integrated Encoder-Decoder Architecture 7 Feb 2025 · 0 repositories · arXiv:2502.05233
-
Probing Internal Representations of Multi-Word Verbs in Large Language Models 7 Feb 2025 · 0 repositories · arXiv:2502.04789
-
A Classification System Approach in Predicting Chinese Censorship 6 Feb 2025 · 0 repositories · arXiv:2502.04234
-
Experiments with Large Language Models on Retrieval-Augmented Generation for Closed-Source Simulation Software 6 Feb 2025 · 0 repositories · arXiv:2502.03916
-
Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis 6 Feb 2025 · 1 repository · arXiv:2502.04128
-
LLMs to Support a Domain Specific Knowledge Assistant 6 Feb 2025 · 0 repositories · arXiv:2502.04095
-
MD-BERT: Action Recognition in Dark Videos via Dynamic Multi-Stream Fusion and Temporal Modeling 6 Feb 2025 · 1 repository · arXiv:2502.03724
-
MedRAG: Enhancing Retrieval-augmented Generation with Knowledge Graph-Elicited Reasoning for Healthcare Copilot 6 Feb 2025 · 1 repository · arXiv:2502.04413
-
MRAMG-Bench: A Comprehensive Benchmark for Advancing Multimodal Retrieval-Augmented Multimodal Generation 6 Feb 2025 · 1 repository · arXiv:2502.04176
-
PsyPlay: Personality-Infused Role-Playing Conversational Agents 6 Feb 2025 · 0 repositories · arXiv:2502.03821
-
Cache-Craft: Managing Chunk-Caches for Efficient Retrieval-Augmented Generation 5 Feb 2025 · 0 repositories · arXiv:2502.15734
-
MARAGE: Transferable Multi-Model Adversarial Attack for Retrieval-Augmented Generation Data Extraction 5 Feb 2025 · 0 repositories · arXiv:2502.04360
-
OPTIC: Optimizing Patient-Provider Triaging & Improving Communications in Clinical Operations using GPT-4 Data Labeling and Model Distillation 5 Feb 2025 · 0 repositories · arXiv:2503.05701
-
Path Planning for Masked Diffusion Model Sampling 5 Feb 2025 · 0 repositories · arXiv:2502.03540
-
Aligning Human and Machine Attention for Enhanced Supervised Learning 4 Feb 2025 · 0 repositories · arXiv:2502.06811
-
CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance 4 Feb 2025 · 1 repository · arXiv:2502.04350
-
Conversation AI Dialog for Medicare powered by Finetuning and Retrieval Augmented Generation 4 Feb 2025 · 0 repositories · arXiv:2502.02249
-
Open Foundation Models in Healthcare: Challenges, Paradoxes, and Opportunities with GenAI Driven Personalized Prescription 4 Feb 2025 · 0 repositories · arXiv:2502.04356
-
OverThink: Slowdown Attacks on Reasoning LLMs 4 Feb 2025 · 1 repository · arXiv:2502.02542
-
Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation 4 Feb 2025 · 1 repository · arXiv:2502.02464
-
Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign 4 Feb 2025 · 0 repositories · arXiv:2502.02068
-
Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions 4 Feb 2025 · 0 repositories · arXiv:2502.18470
-
Topic Modeling in Marathi 4 Feb 2025 · 0 repositories · arXiv:2502.02100
-
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation 3 Feb 2025 · 0 repositories · arXiv:2502.01697
-
GFM-RAG: Graph Foundation Model for Retrieval Augmented Generation 3 Feb 2025 · 1 repository · arXiv:2502.01113Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Harmonic Loss Trains Interpretable AI Models 3 Feb 2025 · 1 repository · arXiv:2502.01628
-
Polynomial, trigonometric, and tropical activations 3 Feb 2025 · 1 repository · arXiv:2502.01247
-
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models 3 Feb 2025 · 0 repositories · arXiv:2502.01386
-
VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos 3 Feb 2025 · 1 repository · arXiv:2502.01549
-
DeepGate4: Efficient and Effective Representation Learning for Circuit Design at Scale 2 Feb 2025 · 1 repository · arXiv:2502.01681Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Explainability in Practice: A Survey of Explainable NLP Across Various Domains 2 Feb 2025 · 0 repositories · arXiv:2502.00837
-
LIBRA: Measuring Bias of Large Language Model from a Local Context 2 Feb 2025 · 0 repositories · arXiv:2502.01679
-
CoddLLM: Empowering Large Language Models for Data Analytics 1 Feb 2025 · 0 repositories · arXiv:2502.00329
-
Fast Solvers for Discrete Diffusion Models: Theory and Applications of High-Order Algorithms 1 Feb 2025 · 0 repositories · arXiv:2502.00234
-
Riddle Me This! Stealthy Membership Inference for Retrieval-Augmented Generation 1 Feb 2025 · 1 repository · arXiv:2502.00306
-
Can AI Solve the Peer Review Crisis? A Large Scale Cross Model Experiment of LLMs' Performance and Biases in Evaluating over 1000 Economics Papers 31 Jan 2025 · 0 repositories · arXiv:2502.00070
-
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search 31 Jan 2025 · 1 repository · arXiv:2501.18922Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
Large Language Models' Accuracy in Emulating Human Experts' Evaluation of Public Sentiments about Heated Tobacco Products on Social Media 31 Jan 2025 · 0 repositories · arXiv:2502.01658
-
Through the Looking Glass: LLM-Based Analysis of AR/VR Android Applications Privacy Policies 31 Jan 2025 · 0 repositories · arXiv:2501.19223
-
AlphaAdam:Asynchronous Masked Optimization with Dynamic Alpha for Selective Updates 30 Jan 2025 · 0 repositories · arXiv:2501.18094
-
Can we Retrieve Everything All at Once? ARM: An Alignment-Oriented LLM-based Retrieval Method 30 Jan 2025 · 0 repositories · arXiv:2501.18539
-
Economic Rationality under Specialization: Evidence of Decision Bias in AI Agents 30 Jan 2025 · 0 repositories · arXiv:2501.18190
-
General Embedding vs. Task-Specific Embedding: A Comparative Approach to Enhancing NLP Performance 30 Jan 2025 · 0 repositories
-
Israel-Hamas war through Telegram, Reddit and Twitter 30 Jan 2025 · 0 repositories · arXiv:2502.00060
-
Leveraging LLM Agents for Automated Optimization Modeling for SASP Problems: A Graph-RAG based Approach 30 Jan 2025 · 0 repositories · arXiv:2501.18320