Methods › General › Regularization › Attention Dropout › Papers, page 26
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 26 of 109: papers 2,501 to 2,600 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos 8 Jul 2024 · 1 repository · arXiv:2407.06018
-
Personality Analysis for Social Media Users using Arabic language and its Effect on Sentiment Analysis 8 Jul 2024 · 0 repositories · arXiv:2407.06314
-
Potential of Multimodal Large Language Models for Data Mining of Medical Images and Free-text Reports 8 Jul 2024 · 0 repositories · arXiv:2407.05758
-
Surprising gender biases in GPT 8 Jul 2024 · 0 repositories · arXiv:2407.06003
-
Vision-Braille: An End-to-End Tool for Chinese Braille Image-to-Text Translation 8 Jul 2024 · 0 repositories · arXiv:2407.06048
-
How do you know that? Teaching Generative Language Models to Reference Answers to Biomedical Questions 6 Jul 2024 · 1 repository · arXiv:2407.05015
-
RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models 6 Jul 2024 · 1 repository · arXiv:2407.05131Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
SHINE: Saliency-aware HIerarchical NEgative Ranking for Compositional Temporal Grounding 6 Jul 2024 · 1 repository · arXiv:2407.05118Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Are LLMs Correctly Integrated into Software Systems? 6 Jul 2024 · 0 repositories · arXiv:2407.05138
-
Are Large Language Models Strategic Decision Makers? A Study of Performance and Bias in Two-Player Non-Zero-Sum Games 5 Jul 2024 · 0 repositories · arXiv:2407.04467
-
GPT vs RETRO: Exploring the Intersection of Retrieval and Parameter-Efficient Fine-Tuning 5 Jul 2024 · 0 repositories · arXiv:2407.04528
-
Using LLMs to label medical papers according to the CIViC evidence model 5 Jul 2024 · 1 repository · arXiv:2407.04466
-
Convolutional vs Large Language Models for Software Log Classification in Edge-Deployable Cellular Network Testing 4 Jul 2024 · 0 repositories · arXiv:2407.03759
-
Deep Content Understanding Toward Entity and Aspect Target Sentiment Analysis on Foundation Models 4 Jul 2024 · 1 repository · arXiv:2407.04050
-
DSLR: Document Refinement with Sentence-Level Re-ranking and Reconstruction to Enhance Retrieval-Augmented Generation 4 Jul 2024 · 0 repositories · arXiv:2407.03627
-
From Data to Commonsense Reasoning: The Use of Large Language Models for Explainable AI 4 Jul 2024 · 0 repositories · arXiv:2407.03778
-
QET: Enhancing Quantized LLM Parameters and KV cache Compression through Element Substitution and Residual Clustering 4 Jul 2024 · 0 repositories · arXiv:2407.03637
-
HYBRINFOX at CheckThat! 2024 -- Task 1: Enhancing Language Models with Structured Information for Check-Worthiness Estimation 4 Jul 2024 · 0 repositories · arXiv:2407.03850
-
HYBRINFOX at CheckThat! 2024 -- Task 2: Enriching BERT Models with the Expert System VAGO for Subjectivity Detection 4 Jul 2024 · 0 repositories · arXiv:2407.03770
-
NutriBench: A Dataset for Evaluating Large Language Models on Nutrition Estimation from Meal Descriptions 4 Jul 2024 · 0 repositories · arXiv:2407.12843
-
Controllable Conversations: Planning-Based Dialogue Agent with Large Language Models 4 Jul 2024 · 1 repository · arXiv:2407.03884
-
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks 4 Jul 2024 · 0 repositories · arXiv:2407.03624
-
Slice-100K: A Multimodal Dataset for Extrusion-based 3D Printing 4 Jul 2024 · 0 repositories · arXiv:2407.04180
-
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4 4 Jul 2024 · 0 repositories · arXiv:2407.04130
-
A Comparative Study of DSL Code Generation: Fine-Tuning vs. Optimized Retrieval Augmentation 3 Jul 2024 · 0 repositories · arXiv:2407.02742
-
AgentInstruct: Toward Generative Teaching with Agentic Flows 3 Jul 2024 · 0 repositories · arXiv:2407.03502
-
Gradient descent with generalized Newton's method 3 Jul 2024 · 1 repository · arXiv:2407.02772Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
CATT: Character-based Arabic Tashkeel Transformer 3 Jul 2024 · 1 repository · arXiv:2407.03236
-
Croppable Knowledge Graph Embedding 3 Jul 2024 · 0 repositories · arXiv:2407.02779
-
Improving LLM Abilities in Idiomatic Translation 3 Jul 2024 · 0 repositories · arXiv:2407.03518
-
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models 3 Jul 2024 · 0 repositories · arXiv:2407.02775
-
ObfuscaTune: Obfuscated Offsite Fine-tuning and Inference of Proprietary LLMs on Private Datasets 3 Jul 2024 · 0 repositories · arXiv:2407.02960
-
OSPC: Artificial VLM Features for Hateful Meme Detection 3 Jul 2024 · 0 repositories · arXiv:2407.12836
-
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring 3 Jul 2024 · 0 repositories · arXiv:2407.13781
-
Regurgitative Training: The Value of Real Data in Training Large Language Models 3 Jul 2024 · 0 repositories · arXiv:2407.12835
-
Assessing the Code Clone Detection Capability of Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02402
-
Beyond Numeric Awards: In-Context Dueling Bandits with LLM Agents 2 Jul 2024 · 0 repositories · arXiv:2407.01887
-
Extracting and Encoding: Leveraging Large Language Models and Medical Knowledge to Enhance Radiological Text Representation 2 Jul 2024 · 1 repository · arXiv:2407.01948
-
GPTCast: a weather language model for precipitation nowcasting 2 Jul 2024 · 1 repository · arXiv:2407.02089
-
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning 2 Jul 2024 · 0 repositories · arXiv:2407.01892
-
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation 2 Jul 2024 · 1 repository · arXiv:2407.02056Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
MeMemo: On-device Retrieval Augmentation for Private and Personalized Text Generation 2 Jul 2024 · 1 repository · arXiv:2407.01972
-
Language Model Alignment in Multilingual Trolley Problems 2 Jul 2024 · 2 repositories · arXiv:2407.02273Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs 2 Jul 2024 · 0 repositories · arXiv:2407.02485
-
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters 2 Jul 2024 · 1 repository · arXiv:2407.01902Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Solution for The PST-KDD-2024 OAG-Challenge 2 Jul 2024 · 0 repositories · arXiv:2407.12827
-
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation 1 Jul 2024 · 1 repository · arXiv:2407.01102Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Face4RAG: Factual Consistency Evaluation for Retrieval Augmented Generation in Chinese 1 Jul 2024 · 0 repositories · arXiv:2407.01080
-
Ground Every Sentence: Improving Retrieval-Augmented LLMs with Interleaved Reference-Claim Generation 1 Jul 2024 · 0 repositories · arXiv:2407.01796
-
Hybrid RAG-empowered Multi-modal LLM for Secure Data Management in Internet of Medical Things: A Diffusion-based Contract Approach 1 Jul 2024 · 0 repositories · arXiv:2407.00978
-
Increasing Model Capacity for Free: A Simple Strategy for Parameter Efficient Fine-tuning 1 Jul 2024 · 1 repository · arXiv:2407.01320Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 7 pointer-only (licence)
-
Multi-Modal Fusion-Based Multi-Task Semantic Communication System 1 Jul 2024 · 0 repositories · arXiv:2407.00964
-
Predicting DC-Link Capacitor Current Ripple in AC-DC Rectifier Circuits Using Fine-Tuned Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.01724
-
Retrieval-augmented generation in multilingual settings 1 Jul 2024 · 1 repository · arXiv:2407.01463Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Searching for Best Practices in Retrieval-Augmented Generation 1 Jul 2024 · 1 repository · arXiv:2407.01219Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Summary of a Haystack: A Challenge to Long-Context LLMs and RAG Systems 1 Jul 2024 · 1 repository · arXiv:2407.01370Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Memory³: Language Modeling with Explicit Memory 1 Jul 2024 · 0 repositories · arXiv:2407.01178
-
Characterizing Stereotypical Bias from Privacy-preserving Pre-Training 30 Jun 2024 · 0 repositories · arXiv:2407.00764
-
LegalTurk Optimized BERT for Multi-Label Text Classification and NER 30 Jun 2024 · 0 repositories · arXiv:2407.00648
-
Parm: Efficient Training of Large Sparsely-Activated Models with Dedicated Schedules 30 Jun 2024 · 1 repository · arXiv:2407.00599
-
Answering real-world clinical questions using large language model based systems 29 Jun 2024 · 0 repositories · arXiv:2407.00541
-
From RAG to RICHES: Retrieval Interlaced with Sequence Generation 29 Jun 2024 · 0 repositories · arXiv:2407.00361
-
LLM-Generated Natural Language Meets Scaling Laws: New Explorations and Data Augmentation Methods 29 Jun 2024 · 0 repositories · arXiv:2407.00322
-
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs 28 Jun 2024 · 0 repositories · arXiv:2406.20060
-
FRED: Flexible REduction-Distribution Interconnect and Communication Implementation for Wafer-Scale Distributed Training of DNN Models 28 Jun 2024 · 0 repositories · arXiv:2406.19580
-
Machine Learning Predictors for Min-Entropy Estimation 28 Jun 2024 · 1 repository · arXiv:2406.19983
-
ScaleBiO: Scalable Bilevel Optimization for LLM Data Reweighting 28 Jun 2024 · 0 repositories · arXiv:2406.19976
-
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents 28 Jun 2024 · 1 repository · arXiv:2407.00132Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Uncertainty Quantification in Large Language Models Through Convex Hull Analysis 28 Jun 2024 · 0 repositories · arXiv:2406.19712
-
AutoPureData: Automated Filtering of Undesirable Web Data to Update LLM Knowledge 27 Jun 2024 · 1 repository · arXiv:2406.19271
-
AutoRAG-HP: Automatic Online Hyper-Parameter Tuning for Retrieval-Augmented Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19251
-
Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads 27 Jun 2024 · 1 repository · arXiv:2406.19391
-
Fine-tuned network relies on generic representation to solve unseen cognitive task 27 Jun 2024 · 0 repositories · arXiv:2406.18926
-
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data 27 Jun 2024 · 1 repository · arXiv:2406.19292
-
Granite-Function Calling Model: Introducing Function Calling Abilities via Multi-task Learning of Granular Tasks 27 Jun 2024 · 0 repositories · arXiv:2407.00121
-
Historia Magistra Vitae: Dynamic Topic Modeling of Roman Literature using Neural Embeddings 27 Jun 2024 · 0 repositories · arXiv:2406.18907
-
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language 27 Jun 2024 · 0 repositories · arXiv:2406.19349
-
Predicting Depression and Anxiety Risk in Dutch Neighborhoods from Street-View Images 27 Jun 2024 · 0 repositories · arXiv:2407.09547
-
RAVEN: Multitask Retrieval Augmented Vision-Language Learning 27 Jun 2024 · 0 repositories · arXiv:2406.19150
-
SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation 27 Jun 2024 · 1 repository · arXiv:2406.19215Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Generating Is Believing: Membership Inference Attacks against Retrieval-Augmented Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19234
-
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models 27 Jun 2024 · 0 repositories · arXiv:2406.19358
-
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets 26 Jun 2024 · 0 repositories · arXiv:2406.18518
-
Evaluating Quality of Answers for Retrieval-Augmented Generation: A Strong LLM Is All You Need 26 Jun 2024 · 0 repositories · arXiv:2406.18064
-
FactFinders at CheckThat! 2024: Refining Check-worthy Statement Detection with LLMs through Data Pruning 26 Jun 2024 · 1 repository · arXiv:2406.18297
-
"Glue pizza and eat rocks" -- Exploiting Vulnerabilities in Retrieval-Augmented Generative Models 26 Jun 2024 · 0 repositories · arXiv:2406.19417
-
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources 26 Jun 2024 · 0 repositories · arXiv:2406.18049
-
Knowledge graph enhanced retrieval-augmented generation for failure mode and effects analysis 26 Jun 2024 · 1 repository · arXiv:2406.18114
-
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data 26 Jun 2024 · 3 repositories · arXiv:2406.18321Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Multi-step Inference over Unstructured Data 26 Jun 2024 · 0 repositories · arXiv:2406.17987
-
Poisoned LangChain: Jailbreak LLMs by LangChain 26 Jun 2024 · 0 repositories · arXiv:2406.18122
-
ResumeAtlas: Revisiting Resume Classification with Large-Scale Datasets and Large Language Models 26 Jun 2024 · 1 repository · arXiv:2406.18125
-
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation 26 Jun 2024 · 1 repository · arXiv:2406.18676Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets 26 Jun 2024 · 0 repositories · arXiv:2406.18239
-
SetBERT: Enhancing Retrieval Performance for Boolean Logic and Set Operation Queries 25 Jun 2024 · 0 repositories · arXiv:2406.17282
-
CTBench: A Comprehensive Benchmark for Evaluating Language Model Capabilities in Clinical Trial Design 25 Jun 2024 · 1 repository · arXiv:2406.17888
-
Interpreting Attention Layer Outputs with Sparse Autoencoders 25 Jun 2024 · 1 repository · arXiv:2406.17759Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
LumberChunker: Long-Form Narrative Document Segmentation 25 Jun 2024 · 1 repository · arXiv:2406.17526Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
RAGBench: Explainable Benchmark for Retrieval-Augmented Generation Systems 25 Jun 2024 · 0 repositories · arXiv:2407.11005
-
This Paper Had the Smartest Reviewers -- Flattery Detection Utilising an Audio-Textual Transformer-Based Approach 25 Jun 2024 · 1 repository · arXiv:2406.17667