Methods › General › Regularization › Attention Dropout › Papers, page 22
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 22 of 109: papers 2,101 to 2,200 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CA-BERT: Leveraging Context Awareness for Enhanced Multi-Turn Chat Interaction 5 Sep 2024 · 0 repositories · arXiv:2409.13701
-
CACER: Clinical Concept Annotations for Cancer Events and Relations 5 Sep 2024 · 1 repository · arXiv:2409.03905
-
Evaluating Open-Source Sparse Autoencoders on Disentangling Factual Knowledge in GPT-2 Small 5 Sep 2024 · 1 repository · arXiv:2409.04478Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MARAGS: A Multi-Adapter System for Multi-Task Retrieval Augmented Generation Question Answering 5 Sep 2024 · 0 repositories · arXiv:2409.03171
-
MaterialBENCH: Evaluating College-Level Materials Science Problem-Solving Abilities of Large Language Models 5 Sep 2024 · 0 repositories · arXiv:2409.03161
-
RAG based Question-Answering for Contextual Response Prediction System 5 Sep 2024 · 0 repositories · arXiv:2409.03708
-
Revolutionizing Database Q&A with Large Language Models: Comprehensive Benchmark and Evaluation 5 Sep 2024 · 1 repository · arXiv:2409.04475
-
Sketch: A Toolkit for Streamlining LLM Operations 5 Sep 2024 · 0 repositories · arXiv:2409.03346
-
A Comparative Study on Large Language Models for Log Parsing 4 Sep 2024 · 0 repositories · arXiv:2409.02474
-
A Data Selection Approach for Enhancing Low Resource Machine Translation Using Cross-Lingual Sentence Representations 4 Sep 2024 · 0 repositories · arXiv:2409.02712
-
GenDFIR: Advancing Cyber Incident Timeline Analysis Through Retrieval Augmented Generation and Large Language Models 4 Sep 2024 · 0 repositories · arXiv:2409.02572
-
Detecting Calls to Action in Multimodal Content: Analysis of the 2021 German Federal Election Campaign on Instagram 4 Sep 2024 · 0 repositories · arXiv:2409.02690
-
Diversify-verify-adapt: Efficient and Robust Retrieval-Augmented Ambiguous Question Answering 4 Sep 2024 · 0 repositories · arXiv:2409.02361
-
How Privacy-Savvy Are Large Language Models? A Case Study on Compliance and Privacy Technical Review 4 Sep 2024 · 0 repositories · arXiv:2409.02375
-
Irrelevant Alternatives Bias Large Language Model Hiring Decisions 4 Sep 2024 · 0 repositories · arXiv:2409.15299
-
Large Language Models as Efficient Reward Function Searchers for Custom-Environment Multi-Objective Reinforcement Learning 4 Sep 2024 · 0 repositories · arXiv:2409.02428
-
MoA is All You Need: Building LLM Research Team using Mixture of Agents 4 Sep 2024 · 0 repositories · arXiv:2409.07487
-
More is More: Addition Bias in Large Language Models 4 Sep 2024 · 1 repository · arXiv:2409.02569
-
OpenFact at CheckThat! 2024: Combining Multiple Attack Methods for Effective Adversarial Text Generation 4 Sep 2024 · 0 repositories · arXiv:2409.02649
-
Pre-training data selection for biomedical domain adaptation using journal impact metrics 4 Sep 2024 · 0 repositories · arXiv:2409.02725
-
Robust Text-to-Cypher Using Combination of BERT, GraphSAGE, and Transformer (CoBGT) Model 4 Sep 2024 · 0 repositories
-
Multi-Source Knowledge Pruning for Retrieval-Augmented Generation: A Benchmark and Empirical Study 3 Sep 2024 · 2 repositories · arXiv:2409.13694
-
You Only Use Reactive Attention Slice For Long Context Retrieval 3 Sep 2024 · 1 repository · arXiv:2409.13695
-
AdaComp: Extractive Context Compression with Adaptive Predictor for Retrieval-Augmented Large Language Models 3 Sep 2024 · 0 repositories · arXiv:2409.01579
-
BEAVER: An Enterprise Benchmark for Text-to-SQL 3 Sep 2024 · 0 repositories · arXiv:2409.02038
-
Benchmarking Cognitive Domains for LLMs: Insights from Taiwanese Hakka Culture 3 Sep 2024 · 0 repositories · arXiv:2409.01556
-
Dialogue You Can Trust: Human and AI Perspectives on Generated Conversations 3 Sep 2024 · 0 repositories · arXiv:2409.01808
-
In Defense of RAG in the Era of Long-Context Language Models 3 Sep 2024 · 0 repositories · arXiv:2409.01666
-
It is Time to Develop an Auditing Framework to Promote Value Aware Chatbots 3 Sep 2024 · 1 repository · arXiv:2409.01539
-
LifeGPT: Topology-Agnostic Generative Pretrained Transformer Model for Cellular Automata 3 Sep 2024 · 1 repository · arXiv:2409.12182
-
The Era of Foundation Models in Medical Imaging is Approaching : A Scoping Review of the Clinical Value of Large-Scale Generative AI Applications in Radiology 3 Sep 2024 · 0 repositories · arXiv:2409.12973
-
Masking The Bias : From Echo Chambers to Large Scale Aspect-Based Sentiment Analysis 2 Sep 2024 · 1 repository
-
Pairing Analogy-Augmented Generation with Procedural Memory for Procedural Q&A 2 Sep 2024 · 1 repository · arXiv:2409.01344
-
Self-Judge: Selective Instruction Following with Alignment Self-Evaluation 2 Sep 2024 · 1 repository · arXiv:2409.00935Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Deep Knowledge-Infusion For Explainable Depression Detection 1 Sep 2024 · 0 repositories · arXiv:2409.02122
-
Modeling Text-Label Alignment for Hierarchical Text Classification 1 Sep 2024 · 1 repository · arXiv:2409.00788
-
The Design of an LLM-powered Unstructured Analytics System 1 Sep 2024 · 0 repositories · arXiv:2409.00847
-
An Empirical Study on Information Extraction using Large Language Models 31 Aug 2024 · 0 repositories · arXiv:2409.00369
-
GenAI-powered Multi-Agent Paradigm for Smart Urban Mobility: Opportunities and Challenges for Integrating Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) with Intelligent Transportation Systems 31 Aug 2024 · 0 repositories · arXiv:2409.00494
-
Assessing Generative Language Models in Classification Tasks: Performance and Self-Evaluation Capabilities in the Environmental and Climate Change Domain 30 Aug 2024 · 1 repository · arXiv:2408.17362
-
Can Large Language Models Address Open-Target Stance Detection? 30 Aug 2024 · 0 repositories · arXiv:2409.00222
-
Improving Extraction of Clinical Event Contextual Properties from Electronic Health Records: A Comparative Study 30 Aug 2024 · 0 repositories · arXiv:2408.17181
-
MaFeRw: Query Rewriting with Multi-Aspect Feedbacks for Retrieval-Augmented Large Language Models 30 Aug 2024 · 1 repository · arXiv:2408.17072
-
ProGRes: Prompted Generative Rescoring on ASR n-Best 30 Aug 2024 · 1 repository · arXiv:2409.00217
-
Retrieval-Augmented Natural Language Reasoning for Explainable Visual Question Answering 30 Aug 2024 · 0 repositories · arXiv:2408.17006
-
RISSOLE: Parameter-efficient Diffusion Models via Block-wise Generation and Retrieval-Guidance 30 Aug 2024 · 0 repositories · arXiv:2408.17095
-
Training Ultra Long Context Language Model with Fully Pipelined Distributed Transformer 30 Aug 2024 · 1 repository · arXiv:2408.16978
-
Unintentional Security Flaws in Code: Automated Defense via Root Cause Analysis 30 Aug 2024 · 0 repositories · arXiv:2409.00199
-
Assessing Large Language Models for Online Extremism Research: Identification, Explanation, and New Knowledge 29 Aug 2024 · 0 repositories · arXiv:2408.16749
-
HyPA-RAG: A Hybrid Parameter Adaptive Retrieval-Augmented Generation System for AI Legal and Policy Applications 29 Aug 2024 · 0 repositories · arXiv:2409.09046
-
LLaVA-Chef: A Multi-modal Generative Model for Food Recipes 29 Aug 2024 · 1 repository · arXiv:2408.16889
-
A Simple Baseline with Single-encoder for Referring Image Segmentation 28 Aug 2024 · 0 repositories · arXiv:2408.15521
-
An Extremely Data-efficient and Generative LLM-based Reinforcement Learning Agent for Recommenders 28 Aug 2024 · 0 repositories · arXiv:2408.16032
-
CBF-LLM: Safe Control for LLM Alignment 28 Aug 2024 · 1 repository · arXiv:2408.15625
-
Conan-embedding: General Text Embedding with More and Better Negative Samples 28 Aug 2024 · 0 repositories · arXiv:2408.15710
-
FRACTURED-SORRY-Bench: Framework for Revealing Attacks in Conversational Turns Undermining Refusal Efficacy and Defenses over SORRY-Bench (Automated Multi-shot Jailbreaks) 28 Aug 2024 · 0 repositories · arXiv:2408.16163
-
Is Personality Prediction Possible Based on Reddit Comments? 28 Aug 2024 · 0 repositories · arXiv:2408.16089
-
Leveraging Large Language Models for Wireless Symbol Detection via In-Context Learning 28 Aug 2024 · 0 repositories · arXiv:2409.00124
-
LRP4RAG: Detecting Hallucinations in Retrieval-Augmented Generation via Layer-wise Relevance Propagation 28 Aug 2024 · 3 repositories · arXiv:2408.15533
-
MambaPlace:Text-to-Point-Cloud Cross-Modal Place Recognition with Attention Mamba Mechanisms 28 Aug 2024 · 1 repository · arXiv:2408.15740
-
Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation 28 Aug 2024 · 1 repository · arXiv:2408.15876
-
A Survey of Large Language Models for European Languages 27 Aug 2024 · 0 repositories · arXiv:2408.15040
-
Into the Unknown Unknowns: Engaged Human Learning through Participation in Language Model Agent Conversations 27 Aug 2024 · 1 repository · arXiv:2408.15232
-
Strategic Optimization and Challenges of Large Language Models in Object-Oriented Programming 27 Aug 2024 · 0 repositories · arXiv:2408.14834
-
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis 27 Aug 2024 · 0 repositories · arXiv:2409.00106
-
CHARTOM: A Visual Theory-of-Mind Benchmark for Multimodal Large Language Models 26 Aug 2024 · 1 repository · arXiv:2408.14419
-
Question answering system of bridge design specification based on large language model 26 Aug 2024 · 1 repository · arXiv:2408.13282
-
Bidirectional Awareness Induction in Autoregressive Seq2Seq Models 25 Aug 2024 · 0 repositories · arXiv:2408.13959
-
CodeGraph: Enhancing Graph Reasoning of LLMs with Code 25 Aug 2024 · 1 repository · arXiv:2408.13863
-
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task 25 Aug 2024 · 0 repositories · arXiv:2408.13909
-
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models 25 Aug 2024 · 1 repository · arXiv:2409.00084
-
Pandora's Box or Aladdin's Lamp: A Comprehensive Analysis Revealing the Role of RAG Noise in Large Language Models 24 Aug 2024 · 1 repository · arXiv:2408.13533Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
An In-Depth Investigation of Data Collection in LLM App Ecosystems 23 Aug 2024 · 0 repositories · arXiv:2408.13247
-
Enhancing Automated Program Repair with Solution Design 22 Aug 2024 · 0 repositories · arXiv:2408.12056
-
Enhancing Multi-hop Reasoning through Knowledge Erasure in Large Language Model Editing 22 Aug 2024 · 0 repositories · arXiv:2408.12456
-
GRATR: Zero-Shot Evidence Graph Retrieval-Augmented Trustworthiness Reasoning 22 Aug 2024 · 1 repository · arXiv:2408.12333
-
Large Language Models as Foundations for Next-Gen Dense Retrieval: A Comprehensive Empirical Assessment 22 Aug 2024 · 0 repositories · arXiv:2408.12194
-
LLMs are not Zero-Shot Reasoners for Biomedical Information Extraction 22 Aug 2024 · 0 repositories · arXiv:2408.12249
-
Optimizing Performance: How Compact Models Match or Exceed GPT's Classification Capabilities through Fine-Tuning 22 Aug 2024 · 0 repositories · arXiv:2409.11408
-
Unlearning Trojans in Large Language Models: A Comparison Between Natural Language and Source Code 22 Aug 2024 · 0 repositories · arXiv:2408.12416
-
A Quick, trustworthy spectral knowledge Q&A system leveraging retrieval-augmented generation on LLM 21 Aug 2024 · 1 repository · arXiv:2408.11557
-
Ancient Wisdom, Modern Tools: Exploring Retrieval-Augmented LLMs for Ancient Indian Philosophy 21 Aug 2024 · 1 repository · arXiv:2408.11903
-
Applying and Evaluating Large Language Models in Mental Health Care: A Scoping Review of Human-Assessed Generative Tasks 21 Aug 2024 · 0 repositories · arXiv:2408.11288
-
D-RMGPT: Robot-assisted collaborative tasks driven by large multimodal models 21 Aug 2024 · 0 repositories · arXiv:2408.11761
-
Mixed Sparsity Training: Achieving 4× FLOP Reduction for Transformer Pretraining 21 Aug 2024 · 0 repositories · arXiv:2408.11746
-
WeQA: A Benchmark for Retrieval Augmented Generation in Wind Energy Domain 21 Aug 2024 · 0 repositories · arXiv:2408.11800
-
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization 21 Aug 2024 · 0 repositories · arXiv:2408.12003
-
RAGLAB: A Modular and Research-Oriented Unified Framework for Retrieval-Augmented Generation 21 Aug 2024 · 1 repository · arXiv:2408.11381
-
The Self-Contained Negation Test Set 21 Aug 2024 · 0 repositories · arXiv:2408.11469
-
Unlocking Adversarial Suffix Optimization Without Affirmative Phrases: Efficient Black-box Jailbreaking via LLM as Optimizer 21 Aug 2024 · 1 repository · arXiv:2408.11313Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CTP-LLM: Clinical Trial Phase Transition Prediction Using Large Language Models 20 Aug 2024 · 0 repositories · arXiv:2408.10995
-
How Well Do Large Language Models Serve as End-to-End Secure Code Agents for Python? 20 Aug 2024 · 0 repositories · arXiv:2408.10495
-
Language Modeling on Tabular Data: A Survey of Foundations, Techniques and Evolution 20 Aug 2024 · 1 repository · arXiv:2408.10548
-
Reading with Intent 20 Aug 2024 · 0 repositories · arXiv:2408.11189
-
Reconciling Methodological Paradigms: Employing Large Language Models as Novice Qualitative Research Assistants in Talent Management Research 20 Aug 2024 · 0 repositories · arXiv:2408.11043
-
Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs 20 Aug 2024 · 1 repository · arXiv:2408.10902
-
Towardseffective teaching assistants: From intent-based chatbots to LLM-poweredteachingassistants 20 Aug 2024 · 0 repositories
-
Tracing Privacy Leakage of Language Models to Training Data via Adjusted Influence Functions 20 Aug 2024 · 0 repositories · arXiv:2408.10468
-
Security Attacks on LLM-based Code Completion Tools 20 Aug 2024 · 1 repository · arXiv:2408.11006
-
A Strategy to Combine 1stGen Transformers and Open LLMs for Automatic Text Classification 19 Aug 2024 · 0 repositories · arXiv:2408.09629