Methods › General › Regularization › Attention Dropout › Papers, page 19
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 19 of 109: papers 1,801 to 1,900 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents 14 Oct 2024 · 1 repository · arXiv:2410.10594Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Will LLMs Replace the Encoder-Only Models in Temporal Relation Classification? 14 Oct 2024 · 1 repository · arXiv:2410.10476Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
A Comparative Study of PDF Parsing Tools Across Diverse Document Categories 13 Oct 2024 · 0 repositories · arXiv:2410.09871
-
Can In-context Learning Really Generalize to Out-of-distribution Tasks? 13 Oct 2024 · 0 repositories · arXiv:2410.09695
-
Evaluating Gender Bias of LLMs in Making Morality Judgements 13 Oct 2024 · 0 repositories · arXiv:2410.09992
-
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG 13 Oct 2024 · 0 repositories · arXiv:2410.09699
-
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs 13 Oct 2024 · 0 repositories · arXiv:2410.12864
-
Learning to Rank for Multiple Retrieval-Augmented Models through Iterative Utility Maximization 13 Oct 2024 · 0 repositories · arXiv:2410.09942
-
M2M-Gen: A Multimodal Framework for Automated Background Music Generation in Japanese Manga Using Large Language Models 13 Oct 2024 · 0 repositories · arXiv:2410.09928
-
Single Ground Truth Is Not Enough: Add Linguistic Variability to Aspect-based Sentiment Analysis Evaluation 13 Oct 2024 · 0 repositories · arXiv:2410.09807
-
Automatic Speech Recognition with BERT and CTC Transformers: A Review 12 Oct 2024 · 0 repositories · arXiv:2410.09456
-
\llinstruct: An Instruction-tuned model for English Language Proficiency Assessments 12 Oct 2024 · 0 repositories · arXiv:2410.09314
-
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation 12 Oct 2024 · 1 repository · arXiv:2410.09584Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
A Methodology for Evaluating RAG Systems: A Case Study On Configuration Dependency Validation 11 Oct 2024 · 1 repository · arXiv:2410.08801
-
AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation 11 Oct 2024 · 1 repository · arXiv:2410.09040
-
Enhancing Long Context Performance in LLMs Through Inner Loop Query Mechanism 11 Oct 2024 · 0 repositories · arXiv:2410.12859
-
Extra Global Attention Designation Using Keyword Detection in Sparse Transformer Architectures 11 Oct 2024 · 0 repositories · arXiv:2410.08971
-
Fine-Tuning In-House Large Language Models to Infer Differential Diagnosis from Radiology Reports 11 Oct 2024 · 0 repositories · arXiv:2410.09234
-
Humanity in AI: Detecting the Personality of Large Language Models 11 Oct 2024 · 0 repositories · arXiv:2410.08545
-
Long Range Named Entity Recognition for Marathi Documents 11 Oct 2024 · 0 repositories · arXiv:2410.09192
-
Observing the Southern US Culture of Honor Using Large-Scale Social Media Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.13887
-
Optimized Biomedical Question-Answering Services with LLM and Multi-BERT Integration 11 Oct 2024 · 0 repositories · arXiv:2410.12856
-
oRetrieval Augmented Generation for 10 Large Language Models and its Generalizability in Assessing Medical Fitness 11 Oct 2024 · 0 repositories · arXiv:2410.08431
-
Retriever-and-Memory: Towards Adaptive Note-Enhanced Retrieval-Augmented Generation 11 Oct 2024 · 1 repository · arXiv:2410.08821Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08698
-
StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization 11 Oct 2024 · 1 repository · arXiv:2410.08815
-
Synth-SONAR: Sonar Image Synthesis with Enhanced Diversity and Realism via Dual Diffusion Models and GPT Prompting 11 Oct 2024 · 1 repository · arXiv:2410.08612
-
Adam Exploits ℓ_∞-geometry of Loss Landscape via Coordinate-wise Adaptivity 10 Oct 2024 · 1 repository · arXiv:2410.08198Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models 10 Oct 2024 · 0 repositories · arXiv:2410.08207
-
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation 10 Oct 2024 · 0 repositories · arXiv:2410.08320
-
Fine-Tuning Language Models for Ethical Ambiguity: A Comparative Study of Alignment with Human Responses 10 Oct 2024 · 0 repositories · arXiv:2410.07826
-
FLIER: Few-shot Language Image Models Embedded with Latent Representations 10 Oct 2024 · 0 repositories · arXiv:2410.07648
-
News Reporter: A Multi-lingual LLM Framework for Broadcast T.V News 10 Oct 2024 · 0 repositories · arXiv:2410.07520
-
No Free Lunch: Retrieval-Augmented Generation Undermines Fairness in LLMs, Even for Vigilant Users 10 Oct 2024 · 0 repositories · arXiv:2410.07589
-
Privately Learning from Graphs with Applications in Fine-tuning Large Language Models 10 Oct 2024 · 1 repository · arXiv:2410.08299
-
Robust AI-Generated Text Detection by Restricted Embeddings 10 Oct 2024 · 1 repository · arXiv:2410.08113Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
The Rise of AI-Generated Content in Wikipedia 10 Oct 2024 · 1 repository · arXiv:2410.08044Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TurboRAG: Accelerating Retrieval-Augmented Generation with Precomputed KV Caches for Chunked Text 10 Oct 2024 · 1 repository · arXiv:2410.07590
-
A Two-Model Approach for Humour Style Recognition 9 Oct 2024 · 1 repository · arXiv:2410.12842
-
Astute RAG: Overcoming Imperfect Retrieval Augmentation and Knowledge Conflicts for Large Language Models 9 Oct 2024 · 0 repositories · arXiv:2410.07176
-
AutoFeedback: An LLM-based Framework for Efficient and Accurate API Request Generation 9 Oct 2024 · 0 repositories · arXiv:2410.06943
-
Capturing Bias Diversity in LLMs 9 Oct 2024 · 0 repositories · arXiv:2410.12839
-
Generative Model for Less-Resourced Language with 1 billion parameters 9 Oct 2024 · 0 repositories · arXiv:2410.06898
-
Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy 9 Oct 2024 · 0 repositories · arXiv:2410.09102
-
Large Language Models as Code Executors: An Exploratory Study 9 Oct 2024 · 0 repositories · arXiv:2410.06667
-
Mental Disorders Detection in the Era of Large Language Models 9 Oct 2024 · 0 repositories · arXiv:2410.07129
-
MentalArena: Self-play Training of Language Models for Diagnosis and Treatment of Mental Health Disorders 9 Oct 2024 · 1 repository · arXiv:2410.06845
-
SAGE: Scalable Ground Truth Evaluations for Large Sparse Autoencoders 9 Oct 2024 · 0 repositories · arXiv:2410.07456
-
SparseGrad: A Selective Method for Efficient Fine-tuning of MLP Layers 9 Oct 2024 · 1 repository · arXiv:2410.07383
-
The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models 9 Oct 2024 · 1 repository · arXiv:2410.06554Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
A Comparative Study of Hybrid Models in Health Misinformation Text Classification 8 Oct 2024 · 0 repositories · arXiv:2410.06311
-
A second-order-like optimizer with adaptive gradient scaling for deep learning 8 Oct 2024 · 1 repository · arXiv:2410.05871
-
Auto-Evolve: Enhancing Large Language Model's Performance via Self-Reasoning Framework 8 Oct 2024 · 0 repositories · arXiv:2410.06328
-
Coevolving with the Other You: Fine-Tuning LLM with Sequential Cooperative Multi-Agent Reinforcement Learning 8 Oct 2024 · 1 repository · arXiv:2410.06101Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Enhancing SPARQL Generation by Triplet-order-sensitive Pre-training 8 Oct 2024 · 1 repository · arXiv:2410.05731
-
Leveraging free energy in pretraining model selection for improved fine-tuning 8 Oct 2024 · 0 repositories · arXiv:2410.05612
-
LightRAG: Simple and Fast Retrieval-Augmented Generation 8 Oct 2024 · 1 repository · arXiv:2410.05779Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG 8 Oct 2024 · 0 repositories · arXiv:2410.05983
-
Retrieving, Rethinking and Revising: The Chain-of-Verification Can Improve Retrieval Augmented Generation 8 Oct 2024 · 0 repositories · arXiv:2410.05801
-
AnyAttack: Towards Large-scale Self-supervised Adversarial Attacks on Vision-language Models 7 Oct 2024 · 0 repositories · arXiv:2410.05346
-
Deciphering the Interplay of Parametric and Non-parametric Memory in Retrieval-augmented Language Models 7 Oct 2024 · 1 repository · arXiv:2410.05162Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
GARLIC: LLM-Guided Dynamic Progress Control with Hierarchical Weighted Graph for Long Document QA 7 Oct 2024 · 0 repositories · arXiv:2410.04790
-
LPZero: Language Model Zero-cost Proxy Search from Zero 7 Oct 2024 · 0 repositories · arXiv:2410.04808
-
Narrative-of-Thought: Improving Temporal Reasoning of Large Language Models via Recounted Narratives 7 Oct 2024 · 1 repository · arXiv:2410.05558
-
On Instruction-Finetuning Neural Machine Translation Models 7 Oct 2024 · 0 repositories · arXiv:2410.05553
-
FAMMA: A Benchmark for Financial Domain Multilingual Multimodal Question Answering 6 Oct 2024 · 1 repository · arXiv:2410.04526Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Inference Scaling for Long-Context Retrieval Augmented Generation 6 Oct 2024 · 0 repositories · arXiv:2410.04343
-
Large Language Model Inference Acceleration: A Comprehensive Hardware Perspective 6 Oct 2024 · 1 repository · arXiv:2410.04466
-
Large Language Models for Knowledge-Free Network Management: Feasibility Study and Opportunities 6 Oct 2024 · 0 repositories · arXiv:2410.17259
-
ProtocoLLM: Automatic Evaluation Framework of LLMs on Domain-Specific Scientific Protocol Formulation Tasks 6 Oct 2024 · 0 repositories · arXiv:2410.04601
-
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job? 5 Oct 2024 · 0 repositories · arXiv:2410.16285
-
Deep Transfer Learning Based Peer Review Aggregation and Meta-review Generation for Scientific Articles 5 Oct 2024 · 0 repositories · arXiv:2410.04202
-
Gamified crowd-sourcing of high-quality data for visual fine-tuning 5 Oct 2024 · 0 repositories · arXiv:2410.04038
-
Metadata-based Data Exploration with Retrieval-Augmented Generation for Large Language Models 5 Oct 2024 · 0 repositories · arXiv:2410.04231
-
Take It Easy: Label-Adaptive Self-Rationalization for Fact Verification and Explanation Generation 5 Oct 2024 · 1 repository · arXiv:2410.04002Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Auto-GDA: Automatic Domain Adaptation for Efficient Grounding Verification in Retrieval Augmented Generation 4 Oct 2024 · 0 repositories · arXiv:2410.03461
-
Crafting Narrative Closures: Zero-Shot Learning with SSM Mamba for Short Story Ending Generation 4 Oct 2024 · 0 repositories · arXiv:2410.10848
-
Cross-lingual Transfer for Automatic Question Generation by Learning Interrogative Structures in Target Languages 4 Oct 2024 · 0 repositories · arXiv:2410.03197
-
How Language Models Prioritize Contextual Grammatical Cues? 4 Oct 2024 · 1 repository · arXiv:2410.03447
-
Learning Semantic Structure through First-Order-Logic Translation 4 Oct 2024 · 0 repositories · arXiv:2410.03203
-
Steering Large Language Models between Code Execution and Textual Reasoning 4 Oct 2024 · 1 repository · arXiv:2410.03524
-
Structured List-Grounded Question Answering 4 Oct 2024 · 0 repositories · arXiv:2410.03950
-
SwiftKV: Fast Prefill-Optimized Inference with Knowledge-Preserving Model Transformation 4 Oct 2024 · 2 repositories · arXiv:2410.03960
-
Towards Linguistically-Aware and Language-Independent Tokenization for Large Language Models (LLMs) 4 Oct 2024 · 0 repositories · arXiv:2410.03568
-
Using Prompts to Guide Large Language Models in Imitating a Real Person's Language Style 4 Oct 2024 · 0 repositories · arXiv:2410.03848
-
Variational Language Concepts for Interpreting Foundation Language Models 4 Oct 2024 · 1 repository · arXiv:2410.03964
-
Vulnerability Detection via Topological Analysis of Attention Maps 4 Oct 2024 · 1 repository · arXiv:2410.03470
-
Ward: Provable RAG Dataset Inference via LLM Watermarks 4 Oct 2024 · 0 repositories · arXiv:2410.03537Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
A Comprehensive Survey of Retrieval-Augmented Generation (RAG): Evolution, Current Landscape and Future Directions 3 Oct 2024 · 0 repositories · arXiv:2410.12837
-
AlphaIntegrator: Transformer Action Search for Symbolic Integration Proofs 3 Oct 2024 · 0 repositories · arXiv:2410.02666
-
CodeJudge: Evaluating Code Generation with Large Language Models 3 Oct 2024 · 1 repository · arXiv:2410.02184Syntology official (archive's flag): 13 ran · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 23 harvested samples) · 2 pointer-only (licence)
-
Adversarial Decoding: Generating Readable Documents for Adversarial Objectives 3 Oct 2024 · 1 repository · arXiv:2410.02163
-
Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization 3 Oct 2024 · 0 repositories · arXiv:2410.02721
-
HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly 3 Oct 2024 · 1 repository · arXiv:2410.02694Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
How Much Can RAG Help the Reasoning of LLM? 3 Oct 2024 · 0 repositories · arXiv:2410.02338
-
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages? 3 Oct 2024 · 0 repositories · arXiv:2410.02611
-
Intrinsic Evaluation of RAG Systems for Deep-Logic Questions 3 Oct 2024 · 0 repositories · arXiv:2410.02932
-
L-CiteEval: Do Long-Context Models Truly Leverage Context for Responding? 3 Oct 2024 · 2 repositories · arXiv:2410.02115
-
LLaVA-Critic: Learning to Evaluate Multimodal Models 3 Oct 2024 · 0 repositories · arXiv:2410.02712
-
Morphological evaluation of subwords vocabulary used by BETO language model 3 Oct 2024 · 0 repositories · arXiv:2410.02283