Methods › General › Regularization › Attention Dropout › Papers, page 8
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 8 of 109: papers 701 to 800 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A LongFormer-Based Framework for Accurate and Efficient Medical Text Summarization 10 Mar 2025 · 0 repositories · arXiv:2503.06888
-
CtrlRAG: Black-box Adversarial Attacks Based on Masked Language Models in Retrieval-Augmented Language Generation 10 Mar 2025 · 0 repositories · arXiv:2503.06950
-
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings 10 Mar 2025 · 0 repositories · arXiv:2503.06980
-
Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning 10 Mar 2025 · 0 repositories · arXiv:2503.07591
-
Fully Autonomous Programming using Iterative Multi-Agent Debugging with Large Language Models 10 Mar 2025 · 0 repositories · arXiv:2503.07693
-
Implicit Reasoning in Transformers is Reasoning through Shortcuts 10 Mar 2025 · 1 repository · arXiv:2503.07604Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Improving cognitive diagnostics in pathology: a deep learning approach for augmenting perceptional understanding of histopathology images 10 Mar 2025 · 0 repositories · arXiv:2503.06894
-
MapQA: Open-domain Geospatial Question Answering on Map Data 10 Mar 2025 · 0 repositories · arXiv:2503.07871
-
Roamify: Designing and Evaluating an LLM Based Google Chrome Extension for Personalised Itinerary Planning 10 Mar 2025 · 1 repository · arXiv:2504.10489
-
Taking Notes Brings Focus? Towards Multi-Turn Multimodal Dialogue Learning 10 Mar 2025 · 0 repositories · arXiv:2503.07002
-
Talking to GDELT Through Knowledge Graphs 10 Mar 2025 · 0 repositories · arXiv:2503.07584
-
Effectiveness of Zero-shot-CoT in Japanese Prompts 9 Mar 2025 · 0 repositories · arXiv:2503.06765
-
Human Cognition Inspired RAG with Knowledge Graph for Complex Problem Solving 9 Mar 2025 · 0 repositories · arXiv:2503.06567
-
Multimodal Emotion Recognition and Sentiment Analysis in Multi-Party Conversation Contexts 9 Mar 2025 · 0 repositories · arXiv:2503.06805
-
Seeing Delta Parameters as JPEG Images: Data-Free Delta Compression with Discrete Cosine Transform 9 Mar 2025 · 0 repositories · arXiv:2503.06676
-
UniGenX: Unified Generation of Sequence and Structure with Autoregressive Diffusion 9 Mar 2025 · 0 repositories · arXiv:2503.06687
-
Breaking Free from MMI: A New Frontier in Rationalization by Probing Input Utilization 8 Mar 2025 · 1 repository · arXiv:2503.06202Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Constructions are Revealed in Word Distributions 8 Mar 2025 · 1 repository · arXiv:2503.06048
-
LimTopic: LLM-based Topic Modeling and Text Summarization for Analyzing Scientific Articles limitations 8 Mar 2025 · 1 repository · arXiv:2503.10658
-
MoEMoE: Question Guided Dense and Scalable Sparse Mixture-of-Expert for Multi-source Multi-modal Answering 8 Mar 2025 · 0 repositories · arXiv:2503.06296
-
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation 8 Mar 2025 · 0 repositories · arXiv:2503.06254
-
Evaluating Large Language Models in Code Generation: INFINITE Methodology for Defining the Inference Index 7 Mar 2025 · 0 repositories · arXiv:2503.05852
-
Simulating and Analysing Human Survey Responses with Large Language Models: A Case Study in Energy Stated Preference 7 Mar 2025 · 0 repositories · arXiv:2503.10652
-
Evaluating open-source Large Language Models for automated fact-checking 7 Mar 2025 · 0 repositories · arXiv:2503.05565
-
FMT:A Multimodal Pneumonia Detection Model Based on Stacking MOE Framework 7 Mar 2025 · 0 repositories · arXiv:2503.05626
-
Leveraging Approximate Caching for Faster Retrieval-Augmented Generation 7 Mar 2025 · 0 repositories · arXiv:2503.05530
-
Leveraging Semantic Type Dependencies for Clinical Named Entity Recognition 7 Mar 2025 · 0 repositories · arXiv:2503.05373
-
Quantifying the Robustness of Retrieval-Augmented Language Models Against Spurious Features in Grounding Data 7 Mar 2025 · 0 repositories · arXiv:2503.05587
-
R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning 7 Mar 2025 · 5 repositories · arXiv:2503.05592Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Zero-shot Medical Event Prediction Using a Generative Pre-trained Transformer on Electronic Health Records 7 Mar 2025 · 0 repositories · arXiv:2503.05893
-
Beyond RAG: Task-Aware KV Cache Compression for Comprehensive Knowledge Reasoning 6 Mar 2025 · 0 repositories · arXiv:2503.04973
-
Collapse of Dense Retrievers: Short, Early, and Literal Biases Outranking Factual Evidence 6 Mar 2025 · 0 repositories · arXiv:2503.05037
-
Compositional Causal Reasoning Evaluation in Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04556
-
HILGEN: Hierarchically-Informed Data Generation for Biomedical NER Using Knowledgebases and Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04930
-
In-depth Analysis of Graph-based RAG in a Unified Framework 6 Mar 2025 · 0 repositories · arXiv:2503.04338
-
Incentivizing Multi-Tenant Split Federated Learning for Foundation Models at the Network Edge 6 Mar 2025 · 0 repositories · arXiv:2503.04971
-
Can Frontier LLMs Replace Annotators in Biomedical Text Mining? Analyzing Challenges and Exploring Solutions 5 Mar 2025 · 1 repository · arXiv:2503.03261
-
Intermediate-Task Transfer Learning: Leveraging Sarcasm Detection for Stance Detection 5 Mar 2025 · 0 repositories · arXiv:2503.03172
-
Large language models in finance : what is financial sentiment? 5 Mar 2025 · 0 repositories · arXiv:2503.03612
-
Personalized Federated Fine-tuning for Heterogeneous Data: An Automatic Rank Learning Approach via Two-Level LoRA 5 Mar 2025 · 0 repositories · arXiv:2503.03920
-
Sarcasm Detection as a Catalyst: Improving Stance Detection with Cross-Target Capabilities 5 Mar 2025 · 0 repositories · arXiv:2503.03787
-
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation 5 Mar 2025 · 1 repository · arXiv:2503.03308
-
A Transformer Model for Predicting Chemical Reaction Products from Generic Templates 4 Mar 2025 · 0 repositories · arXiv:2503.05810
-
Effectively Steer LLM To Follow Preference via Building Confident Directions 4 Mar 2025 · 0 repositories · arXiv:2503.02989
-
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02917
-
LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning 4 Mar 2025 · 0 repositories · arXiv:2503.04812
-
Optimizing open-domain question answering with graph-based retrieval augmented generation 4 Mar 2025 · 0 repositories · arXiv:2503.02922
-
PennyLang: Pioneering LLM-Based Quantum Code Generation with a Novel PennyLane-Centric Dataset 4 Mar 2025 · 0 repositories · arXiv:2503.02497
-
Use Me Wisely: AI-Driven Assessment for LLM Prompting Skills Development 4 Mar 2025 · 0 repositories · arXiv:2503.02532
-
Weak-to-Strong Generalization Even in Random Feature Networks, Provably 4 Mar 2025 · 0 repositories · arXiv:2503.02877
-
Wikipedia in the Era of LLMs: Evolution and Risks 4 Mar 2025 · 1 repository · arXiv:2503.02879
-
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders 4 Mar 2025 · 0 repositories · arXiv:2503.02993
-
Enhancing Social Media Rumor Detection: A Semantic and Graph Neural Network Approach for the 2024 Global Election 3 Mar 2025 · 0 repositories · arXiv:2503.01394
-
Machine Learners Should Acknowledge the Legal Implications of Large Language Models as Personal Data 3 Mar 2025 · 0 repositories · arXiv:2503.01630
-
SAGE: A Framework of Precise Retrieval for RAG 3 Mar 2025 · 0 repositories · arXiv:2503.01713
-
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01814
-
Attention Condensation via Sparsity Induced Regularized Training 3 Mar 2025 · 0 repositories · arXiv:2503.01564
-
Boolean-aware Attention for Dense Retrieval 3 Mar 2025 · 0 repositories · arXiv:2503.01753
-
Cancer Type, Stage and Prognosis Assessment from Pathology Reports using LLMs 3 Mar 2025 · 1 repository · arXiv:2503.01194
-
HoH: A Dynamic Benchmark for Evaluating the Impact of Outdated Information on Retrieval-Augmented Generation 3 Mar 2025 · 0 repositories · arXiv:2503.04800Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG 3 Mar 2025 · 1 repository · arXiv:2503.01222
-
SePer: Measure Retrieval Utility Through The Lens Of Semantic Perplexity Reduction 3 Mar 2025 · 1 repository · arXiv:2503.01478Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
SRAG: Structured Retrieval-Augmented Generation for Multi-Entity Question Answering over Wikipedia Graph 3 Mar 2025 · 0 repositories · arXiv:2503.01346
-
Using (Not so) Large Language Models for Generating Simulation Models in a Formal DSL -- A Study on Reaction Networks 3 Mar 2025 · 0 repositories · arXiv:2503.01675
-
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking 2 Mar 2025 · 1 repository · arXiv:2503.00955
-
ER-RAG: Enhance RAG with ER-Based Unified Modeling of Heterogeneous Data Sources 2 Mar 2025 · 0 repositories · arXiv:2504.06271
-
Optimizing Multi-Hop Document Retrieval Through Intermediate Representations 2 Mar 2025 · 0 repositories · arXiv:2503.04796
-
Pseudo-Knowledge Graph: Meta-Path Guided Retrieval and In-Graph Text for RAG-Equipped LLM 1 Mar 2025 · 0 repositories · arXiv:2503.00309
-
BadJudge: Backdoor Vulnerabilities of LLM-as-a-Judge 1 Mar 2025 · 0 repositories · arXiv:2503.00596
-
Hierarchical Multi-Stage BERT Fusion Framework with Dual Attention for Enhanced Cyberbullying Detection in Social Media 1 Mar 2025 · 0 repositories · arXiv:2503.00342
-
Psychological Counseling Ability of Large Language Models 1 Mar 2025 · 0 repositories · arXiv:2503.07627
-
U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack 1 Mar 2025 · 1 repository · arXiv:2503.00353
-
SafeAuto: Knowledge-Enhanced Safe Autonomous Driving with Multimodal Foundation Models 28 Feb 2025 · 1 repository · arXiv:2503.00211
-
Beyond Words: A Latent Memory Approach to Internal Reasoning in LLMs 28 Feb 2025 · 0 repositories · arXiv:2502.21030
-
CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation 28 Feb 2025 · 1 repository · arXiv:2502.21074Syntology 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Ext2Gen: Alignment through Unified Extraction and Generation for Robust Retrieval-Augmented Generation 28 Feb 2025 · 0 repositories · arXiv:2503.04789
-
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation 28 Feb 2025 · 1 repository · arXiv:2502.20640
-
NutriGen: Personalized Meal Plan Generator Leveraging Large Language Models to Enhance Dietary and Nutritional Adherence 28 Feb 2025 · 1 repository · arXiv:2502.20601
-
Retrieval Augmented Generation for Topic Modeling in Organizational Research: An Introduction with Empirical Demonstration 28 Feb 2025 · 0 repositories · arXiv:2502.20963
-
RuCCoD: Towards Automated ICD Coding in Russian 28 Feb 2025 · 1 repository · arXiv:2502.21263
-
SuperRAG: Beyond RAG with Layout-Aware Graph Modeling 28 Feb 2025 · 0 repositories · arXiv:2503.04790
-
TeleRAG: Efficient Retrieval-Augmented Generation Inference with Lookahead Retrieval 28 Feb 2025 · 0 repositories · arXiv:2502.20969
-
The RAG Paradox: A Black-Box Attack Exploiting Unintentional Vulnerabilities in Retrieval-Augmented Generation Systems 28 Feb 2025 · 0 repositories · arXiv:2502.20995
-
Advanced Deep Learning Techniques for Analyzing Earnings Call Transcripts: Methodologies and Applications 27 Feb 2025 · 0 repositories · arXiv:2503.01886
-
An exploration of features to improve the generalisability of fake news detection models 27 Feb 2025 · 0 repositories · arXiv:2502.20299
-
Bridging Legal Knowledge and AI: Retrieval-Augmented Generation with Vector Stores, Knowledge Graphs, and Hierarchical Non-negative Matrix Factorization 27 Feb 2025 · 1 repository · arXiv:2502.20364
-
Long-Context Inference with Retrieval-Augmented Speculative Decoding 27 Feb 2025 · 1 repository · arXiv:2502.20330
-
Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think 27 Feb 2025 · 1 repository · arXiv:2502.20172
-
SeisMoLLM: Advancing Seismic Monitoring via Cross-modal Transfer with Pre-trained Large Language Model 27 Feb 2025 · 1 repository · arXiv:2502.19960Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Cognitive networks highlight differences and similarities in the STEM mindsets of human and LLM-simulated trainees, experts and academics 26 Feb 2025 · 0 repositories · arXiv:2502.19529
-
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation 26 Feb 2025 · 0 repositories · arXiv:2502.18726
-
Efficient Federated Search for Retrieval-Augmented Generation 26 Feb 2025 · 0 repositories · arXiv:2502.19280
-
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering 26 Feb 2025 · 0 repositories · arXiv:2502.18993
-
NeoBERT: A Next-Generation BERT 26 Feb 2025 · 1 repository · arXiv:2502.19587
-
The Sharpness Disparity Principle in Transformers for Accelerating Language Model Pre-Training 26 Feb 2025 · 0 repositories · arXiv:2502.19002
-
Assessing Large Language Models in Agentic Multilingual National Bias 25 Feb 2025 · 0 repositories · arXiv:2502.17945
-
Broadening Discovery through Structural Models: Multimodal Combination of Local and Structural Properties for Predicting Chemical Features 25 Feb 2025 · 0 repositories · arXiv:2502.17986
-
Detecting Knowledge Boundary of Vision Large Language Models by Sampling-Based Inference 25 Feb 2025 · 0 repositories · arXiv:2502.18023
-
Enhancing Text Classification with a Novel Multi-Agent Collaboration Framework Leveraging BERT 25 Feb 2025 · 0 repositories · arXiv:2502.18653
-
Faster, Cheaper, Better: Multi-Objective Hyperparameter Optimization for LLM and RAG Systems 25 Feb 2025 · 0 repositories · arXiv:2502.18635