Methods › General › Regularization › Attention Dropout › Papers, page 11
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 11 of 109: papers 1,001 to 1,100 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RbFT: Robust Fine-tuning for Retrieval-Augmented Generation against Retrieval Defects 30 Jan 2025 · 1 repository · arXiv:2501.18365
-
Retrieval Augmented Generation Based LLM Evaluation For Protocol State Machine Inference With Chain-of-Thought Reasoning 30 Jan 2025 · 0 repositories · arXiv:2502.15727
-
Scalable and Cost-Efficient ML Inference: Parallel Batch Processing with Serverless Functions 30 Jan 2025 · 0 repositories · arXiv:2502.12017
-
Structure Development in List-Sorting Transformers 30 Jan 2025 · 0 repositories · arXiv:2501.18666
-
Unraveling the Capabilities of Language Models in News Summarization 30 Jan 2025 · 1 repository · arXiv:2501.18128
-
WILDCHAT-50M: A Deep Dive Into the Role of Synthetic Data in Post-Training 30 Jan 2025 · 1 repository · arXiv:2501.18511Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
Hybrid Graphs for Table-and-Text based Question Answering using LLMs 29 Jan 2025 · 0 repositories · arXiv:2501.17767
-
Leveraging In-Context Learning and Retrieval-Augmented Generation for Automatic Question Generation in Educational Domains 29 Jan 2025 · 0 repositories · arXiv:2501.17397
-
Attribution analysis of legal language as used by LLM 28 Jan 2025 · 0 repositories · arXiv:2501.17330
-
Balancing Content Size in RAG-Text2SQL System 28 Jan 2025 · 0 repositories · arXiv:2502.15723
-
Detecting harassment and defamation in cyberbullying with emotion-adaptive training 28 Jan 2025 · 1 repository · arXiv:2501.16925
-
Graph of Attacks with Pruning: Optimizing Stealthy Jailbreak Prompt Generation for Enhanced LLM Content Moderation 28 Jan 2025 · 1 repository · arXiv:2501.18638
-
Multiple Abstraction Level Retrieve Augment Generation 28 Jan 2025 · 0 repositories · arXiv:2501.16952
-
Open-Source Retrieval Augmented Generation Framework for Retrieving Accurate Medication Insights from Formularies for African Healthcare Workers 28 Jan 2025 · 0 repositories · arXiv:2502.15722
-
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model 28 Jan 2025 · 1 repository · arXiv:2501.18636
-
A Comprehensive Study on Fine-Tuning Large Language Models for Medical Question Answering Using Classification Models and Comparative Analysis 27 Jan 2025 · 0 repositories · arXiv:2501.17190
-
Enhancing and Exploring Mild Cognitive Impairment Detection with W2V-BERT-2.0 27 Jan 2025 · 0 repositories · arXiv:2501.16201
-
Kernels of Selfhood: GPT-4o shows humanlike patterns of cognitive consistency moderated by free choice 27 Jan 2025 · 0 repositories · arXiv:2502.07088
-
LemmaHead: RAG Assisted Proof Generation Using Large Language Models 27 Jan 2025 · 0 repositories · arXiv:2501.15797
-
Parametric Retrieval Augmented Generation 27 Jan 2025 · 1 repository · arXiv:2501.15915
-
Provence: efficient and robust context pruning for retrieval-augmented generation 27 Jan 2025 · 0 repositories · arXiv:2501.16214
-
RelCAT: Advancing Extraction of Clinical Inter-Entity Relationships from Unstructured Electronic Health Records 27 Jan 2025 · 1 repository · arXiv:2501.16077
-
URAG: Implementing a Unified Hybrid RAG for Precise Answers in University Admission Chatbots -- A Case Study at HCMUT 27 Jan 2025 · 0 repositories · arXiv:2501.16276
-
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference 27 Jan 2025 · 0 repositories · arXiv:2501.15754
-
Decentralized Low-Rank Fine-Tuning of Large Language Models 26 Jan 2025 · 0 repositories · arXiv:2501.15361
-
Visualizing Uncertainty in Translation Tasks: An Evaluation of LLM Performance and Confidence Metrics 26 Jan 2025 · 1 repository · arXiv:2501.17187
-
Advanced Real-Time Fraud Detection Using RAG-Based LLMs 25 Jan 2025 · 0 repositories · arXiv:2501.15290
-
An AI-Driven Live Systematic Reviews in the Brain-Heart Interconnectome: Minimizing Research Waste and Advancing Evidence Synthesis 25 Jan 2025 · 1 repository · arXiv:2501.17181
-
An Attempt to Unraveling Token Prediction Refinement and Identifying Essential Layers of Large Language Models 25 Jan 2025 · 0 repositories · arXiv:2501.15054
-
ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval 25 Jan 2025 · 0 repositories · arXiv:2501.15245
-
CG-RAG: Research Question Answering by Citation Graph Retrieval-Augmented LLMs 25 Jan 2025 · 0 repositories · arXiv:2501.15067
-
Improving Retrieval-Augmented Generation through Multi-Agent Reinforcement Learning 25 Jan 2025 · 1 repository · arXiv:2501.15228Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Speech Translation Refinement using Large Language Models 25 Jan 2025 · 1 repository · arXiv:2501.15090
-
TrustDataFilter:Leveraging Trusted Knowledge Base Data for More Effective Filtering of Unknown Information 25 Jan 2025 · 0 repositories · arXiv:2502.15714
-
Causal Graphs Meet Thoughts: Enhancing Complex Reasoning in Graph-Augmented LLMs 24 Jan 2025 · 1 repository · arXiv:2501.14892
-
Chain-of-Retrieval Augmented Generation 24 Jan 2025 · 0 repositories · arXiv:2501.14342
-
DarkMind: Latent Chain-of-Thought Backdoor in Customized LLMs 24 Jan 2025 · 0 repositories · arXiv:2501.18617
-
Diffusion based Text-to-Music Generation with Global and Local Text based Conditioning 24 Jan 2025 · 0 repositories · arXiv:2501.14680
-
Fast Think-on-Graph: Wider, Deeper and Faster Reasoning of Large Language Model on Knowledge Graph 24 Jan 2025 · 1 repository · arXiv:2501.14300
-
GraPPI: A Retrieve-Divide-Solve GraphRAG Framework for Large-scale Protein-protein Interaction Exploration 24 Jan 2025 · 1 repository · arXiv:2501.16382
-
Idiom Detection in Sorani Kurdish Texts 24 Jan 2025 · 0 repositories · arXiv:2501.14528
-
Prompt-Based Cost-Effective Evaluation and Operation of ChatGPT as a Computer Programming Teaching Assistant 24 Jan 2025 · 0 repositories · arXiv:2501.17176
-
Rethinking Table Instruction Tuning 24 Jan 2025 · 1 repository · arXiv:2501.14693
-
A Transformer-based Autoregressive Decoder Architecture for Hierarchical Text Classification 23 Jan 2025 · 1 repository · arXiv:2501.13598
-
CAPRAG: A Large Language Model Solution for Customer Service and Automatic Reporting using Vector and Graph Retrieval-Augmented Generation 23 Jan 2025 · 0 repositories · arXiv:2501.13993
-
GraphRAG under Fire 23 Jan 2025 · 0 repositories · arXiv:2501.14050
-
LLMs are Vulnerable to Malicious Prompts Disguised as Scientific Language 23 Jan 2025 · 0 repositories · arXiv:2501.14073
-
Multi-Level Attention and Contrastive Learning for Enhanced Text Classification with an Optimized Transformer 23 Jan 2025 · 0 repositories · arXiv:2501.13467
-
Retrievals Can Be Detrimental: A Contrastive Backdoor Attack Paradigm on Retrieval-Augmented Diffusion Models 23 Jan 2025 · 0 repositories · arXiv:2501.13340
-
RPO: Retrieval Preference Optimization for Robust Retrieval-Augmented Generation 23 Jan 2025 · 0 repositories · arXiv:2501.13726
-
StreamingRAG: Real-time Contextual Retrieval and Generation Framework 23 Jan 2025 · 0 repositories · arXiv:2501.14101
-
Adaptive Retrieval Without Self-Knowledge? Bringing Uncertainty Back Home 22 Jan 2025 · 0 repositories · arXiv:2501.12835Syntology 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 5 pointer-only (licence)
-
EvidenceMap: Learning Evidence Analysis to Unleash the Power of Small Language Models for Biomedical Question Answering 22 Jan 2025 · 0 repositories · arXiv:2501.12746
-
Exploring GPT's Ability as a Judge in Music Understanding 22 Jan 2025 · 1 repository · arXiv:2501.13261
-
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana 22 Jan 2025 · 0 repositories · arXiv:2501.12789
-
RAG-Reward: Optimizing RAG with Reward Modeling and RLHF 22 Jan 2025 · 0 repositories · arXiv:2501.13264
-
A Survey of Graph Retrieval-Augmented Generation for Customized Large Language Models 21 Jan 2025 · 1 repository · arXiv:2501.13958
-
Academic Case Reports Lack Diversity: Assessing the Presence and Diversity of Sociodemographic and Behavioral Factors related to Post COVID-19 Condition 21 Jan 2025 · 0 repositories · arXiv:2501.12538
-
Advancing the Understanding and Evaluation of AR-Generated Scenes: When Vision-Language Models Shine and Stumble 21 Jan 2025 · 1 repository · arXiv:2501.13964
-
ALoFTRAG: Automatic Local Fine Tuning for Retrieval Augmented Generation 21 Jan 2025 · 1 repository · arXiv:2501.11929
-
Assisting Mathematical Formalization with A Learning-based Premise Retriever 21 Jan 2025 · 1 repository · arXiv:2501.13959
-
Comparative Approaches to Sentiment Analysis Using Datasets in Major European and Arabic Languages 21 Jan 2025 · 0 repositories · arXiv:2501.12540
-
Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation 21 Jan 2025 · 0 repositories · arXiv:2501.12432
-
FOCUS: First Order Concentrated Updating Scheme 21 Jan 2025 · 0 repositories · arXiv:2501.12243
-
Harnessing Generative Pre-Trained Transformer for Datacenter Packet Trace Generation 21 Jan 2025 · 0 repositories · arXiv:2501.12033
-
Med-R²: Crafting Trustworthy LLM Physicians via Retrieval and Reasoning of Evidence-Based Medicine 21 Jan 2025 · 1 repository · arXiv:2501.11885
-
Network-informed Prompt Engineering against Organized Astroturf Campaigns under Extreme Class Imbalance 21 Jan 2025 · 1 repository · arXiv:2501.11849
-
Vision-Language Models for Automated Chest X-ray Interpretation: Leveraging ViT and GPT-2 21 Jan 2025 · 0 repositories · arXiv:2501.12356
-
Explainable Lane Change Prediction for Near-Crash Scenarios Using Knowledge Graph Embeddings and Retrieval Augmented Generation 20 Jan 2025 · 0 repositories · arXiv:2501.11560
-
KEIR @ ECIR 2025: The Second Workshop on Knowledge-Enhanced Information Retrieval 20 Jan 2025 · 0 repositories · arXiv:2501.11499
-
PIKE-RAG: sPecIalized KnowledgE and Rationale Augmented Generation 20 Jan 2025 · 1 repository · arXiv:2501.11551Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Poison-RAG: Adversarial Data Poisoning Attacks on Retrieval-Augmented Generation in Recommender Systems 20 Jan 2025 · 1 repository · arXiv:2501.11759
-
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection 20 Jan 2025 · 0 repositories · arXiv:2501.11786
-
Trustformer: A Trusted Federated Transformer 20 Jan 2025 · 0 repositories · arXiv:2501.11706
-
TutorLLM: Customizing Learning Recommendations with Knowledge Tracing and Retrieval-Augmented Generation 20 Jan 2025 · 0 repositories · arXiv:2502.15709
-
From Arabic Text to Puzzles: LLM-Driven Development of Arabic Educational Crosswords 19 Jan 2025 · 0 repositories · arXiv:2501.11035
-
FSMoE: A Flexible and Scalable Training System for Sparse Mixture-of-Experts Models 18 Jan 2025 · 0 repositories · arXiv:2501.10714
-
GEC-RAG: Improving Generative Error Correction via Retrieval-Augmented Generation for Automatic Speech Recognition Systems 18 Jan 2025 · 0 repositories · arXiv:2501.10734
-
Visual RAG: Expanding MLLM visual knowledge without fine-tuning 18 Jan 2025 · 0 repositories · arXiv:2501.10834
-
4bit-Quantization in Vector-Embedding for RAG 17 Jan 2025 · 1 repository · arXiv:2501.10534
-
AirRAG: Activating Intrinsic Reasoning for Retrieval Augmented Generation via Tree-based Search 17 Jan 2025 · 0 repositories · arXiv:2501.10053
-
BBPOS: BERT-based Part-of-Speech Tagging for Uzbek 17 Jan 2025 · 0 repositories · arXiv:2501.10107
-
Bias in Decision-Making for AI's Ethical Dilemmas: A Comparative Study of ChatGPT and Claude 17 Jan 2025 · 1 repository · arXiv:2501.10484
-
Passage Segmentation of Documents for Extractive Question Answering 17 Jan 2025 · 0 repositories · arXiv:2501.09940
-
Confidence Estimation for Error Detection in Text-to-SQL Systems 16 Jan 2025 · 1 repository · arXiv:2501.09527
-
On Learning Informative Trajectory Embeddings for Imitation, Classification and Regression 16 Jan 2025 · 1 repository · arXiv:2501.09327
-
Perspective Transition of Large Language Models for Solving Subjective Tasks 16 Jan 2025 · 0 repositories · arXiv:2501.09265
-
Sentiment Analysis in Twitter Social Network Centered on Cryptocurrencies Using Machine Learning 16 Jan 2025 · 0 repositories · arXiv:2501.09777
-
Agentic Retrieval-Augmented Generation: A Survey on Agentic RAG 15 Jan 2025 · 1 repository · arXiv:2501.09136
-
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment 15 Jan 2025 · 0 repositories · arXiv:2501.09126
-
Expanding Vietnamese SentiWordNet to Improve Performance of Vietnamese Sentiment Analysis Models 15 Jan 2025 · 0 repositories · arXiv:2501.08758
-
Generative AI Takes a Statistics Exam: A Comparison of Performance between ChatGPT3.5, ChatGPT4, and ChatGPT4o-mini 15 Jan 2025 · 0 repositories · arXiv:2501.09171
-
The Impact of Big Five Personality Traits on AI Agent Decision-Making in Public Spaces: A Social Simulation Study 15 Jan 2025 · 0 repositories · arXiv:2503.15497
-
A Driver Advisory System Based on Large Language Model for High-speed Train 14 Jan 2025 · 0 repositories · arXiv:2501.07837
-
ASTRID -- An Automated and Scalable TRIaD for the Evaluation of RAG-based Clinical Question Answering Systems 14 Jan 2025 · 0 repositories · arXiv:2501.08208
-
Comparative Analysis of Efficient Adapter-Based Fine-Tuning of State-of-the-Art Transformer Models 14 Jan 2025 · 0 repositories · arXiv:2501.08271
-
Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models 14 Jan 2025 · 0 repositories · arXiv:2501.08248
-
Exploring Narrative Clustering in Large Language Models: A Layerwise Analysis of BERT 14 Jan 2025 · 0 repositories · arXiv:2501.08053
-
Exploring Robustness of Multilingual LLMs on Real-World Noisy Data 14 Jan 2025 · 1 repository · arXiv:2501.08322
-
Investigating Energy Efficiency and Performance Trade-offs in LLM Inference Across Tasks and DVFS Settings 14 Jan 2025 · 0 repositories · arXiv:2501.08219