Methods › General › Regularization › Attention Dropout › Papers, page 29
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 29 of 109: papers 2,801 to 2,900 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Focus on the Core: Efficient Attention via Pruned Token Compression for Document Classification 3 Jun 2024 · 0 repositories · arXiv:2406.01283
-
In-Context Learning of Physical Properties: Few-Shot Adaptation to Out-of-Distribution Molecular Graphs 3 Jun 2024 · 0 repositories · arXiv:2406.01808
-
Luna: An Evaluation Foundation Model to Catch Language Model Hallucinations with High Accuracy and Low Cost 3 Jun 2024 · 0 repositories · arXiv:2406.00975
-
Natural Language Interaction with a Household Electricity Knowledge-based Digital Twin 3 Jun 2024 · 0 repositories · arXiv:2406.06566
-
SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning 3 Jun 2024 · 1 repository · arXiv:2406.01006Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples)
-
SoccerRAG: Multimodal Soccer Information Retrieval via Natural Queries 3 Jun 2024 · 1 repository · arXiv:2406.01273
-
SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models 3 Jun 2024 · 1 repository · arXiv:2406.01584
-
Superhuman performance in urology board questions by an explainable large language model enabled for context integration of the European Association of Urology guidelines: the UroBot study 3 Jun 2024 · 0 repositories · arXiv:2406.01428
-
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding 3 Jun 2024 · 0 repositories · arXiv:2406.01306
-
A Theory for Token-Level Harmonization in Retrieval-Augmented Generation 3 Jun 2024 · 0 repositories · arXiv:2406.00944
-
Applying Fine-Tuned LLMs for Reducing Data Needs in Load Profile Analysis 2 Jun 2024 · 0 repositories · arXiv:2406.02479
-
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction 2 Jun 2024 · 1 repository · arXiv:2406.00755Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
FOCUS: Forging Originality through Contrastive Use in Self-Plagiarism for Language Models 2 Jun 2024 · 0 repositories · arXiv:2406.00839
-
Formality Style Transfer in Persian 2 Jun 2024 · 0 repositories · arXiv:2406.00867
-
Pretrained Hybrids with MAD Skills 2 Jun 2024 · 0 repositories · arXiv:2406.00894
-
Domain-specific ReAct for physics-integrated iterative modeling: A case study of LLM agents for gas path analysis of gas turbines 1 Jun 2024 · 0 repositories · arXiv:2406.07572
-
An Evaluation Benchmark for Autoformalization in Lean4 1 Jun 2024 · 0 repositories · arXiv:2406.06555
-
Beyond Metrics: Evaluating LLMs' Effectiveness in Culturally Nuanced, Low-Resource Real-World Scenarios 1 Jun 2024 · 0 repositories · arXiv:2406.00343
-
CASE: Efficient Curricular Data Pre-training for Building Assistive Psychology Expert Models 1 Jun 2024 · 1 repository · arXiv:2406.00314
-
Mix-of-Granularity: Optimize the Chunking Granularity for Retrieval-Augmented Generation 1 Jun 2024 · 1 repository · arXiv:2406.00456
-
Pseudo-label Based Domain Adaptation for Zero-Shot Text Steganalysis 1 Jun 2024 · 0 repositories · arXiv:2406.18565
-
RoBERTa-BiLSTM: A Context-Aware Hybrid Model for Sentiment Analysis 1 Jun 2024 · 1 repository · arXiv:2406.00367
-
A comparison of correspondence analysis with PMI-based word embedding methods 31 May 2024 · 1 repository · arXiv:2405.20895
-
Bi-Directional Transformers vs. word2vec: Discovering Vulnerabilities in Lifted Compiled Code 31 May 2024 · 0 repositories · arXiv:2405.20611
-
Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial Training 31 May 2024 · 1 repository · arXiv:2405.20978Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Generative AI Voting: Fair Collective Choice is Resilient to LLM Biases and Inconsistencies 31 May 2024 · 1 repository · arXiv:2406.11871
-
Hard Cases Detection in Motion Prediction by Vision-Language Foundation Models 31 May 2024 · 1 repository · arXiv:2405.20991
-
Large Language Models are Zero-Shot Next Location Predictors 31 May 2024 · 1 repository · arXiv:2405.20962Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
LOLAMEME: Logic, Language, Memory, Mechanistic Framework 31 May 2024 · 0 repositories · arXiv:2406.02592
-
Multilingual Text Style Transfer: Datasets & Models for Indian Languages 31 May 2024 · 2 repositories · arXiv:2405.20805
-
RAG Does Not Work for Enterprises 31 May 2024 · 0 repositories · arXiv:2406.04369
-
Retrieval Meets Reasoning: Even High-school Textbook Knowledge Benefits Multimodal Reasoning 31 May 2024 · 0 repositories · arXiv:2405.20834
-
The Point of View of a Sentiment: Towards Clinician Bias Detection in Psychiatric Notes 31 May 2024 · 0 repositories · arXiv:2405.20582
-
ANAH: Analytical Annotation of Hallucinations in Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20315Syntology official: no sample here; runs from other or unrecorded repositories · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
AutoBreach: Universal and Adaptive Jailbreaking with Efficient Wordplay-Guided Optimization 30 May 2024 · 0 repositories · arXiv:2405.19668
-
Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation 30 May 2024 · 0 repositories · arXiv:2405.20092
-
Ensemble Model With Bert,Roberta and Xlnet For Molecular property prediction 30 May 2024 · 0 repositories · arXiv:2406.06553
-
GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning 30 May 2024 · 1 repository · arXiv:2405.20139Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Heidelberg-Boston @ SIGTYP 2024 Shared Task: Enhancing Low-Resource Language Analysis With Character-Aware Hierarchical Transformers 30 May 2024 · 1 repository · arXiv:2405.20145
-
Is My Data in Your Retrieval Database? Membership Inference Attacks Against Retrieval Augmented Generation 30 May 2024 · 0 repositories · arXiv:2405.20446
-
KerasCV and KerasNLP: Vision and Language Power-Ups 30 May 2024 · 0 repositories · arXiv:2405.20247
-
Knowledge Graph Tuning: Real-time Large Language Model Personalization based on Human Feedback 30 May 2024 · 0 repositories · arXiv:2405.19686
-
LLaMEA: A Large Language Model Evolutionary Algorithm for Automatically Generating Metaheuristics 30 May 2024 · 2 repositories · arXiv:2405.20132Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
One Token Can Help! Learning Scalable and Pluggable Virtual Tokens for Retrieval-Augmented Large Language Models 30 May 2024 · 2 repositories · arXiv:2405.19670
-
Phantom: General Trigger Attacks on Retrieval Augmented Language Generation 30 May 2024 · 0 repositories · arXiv:2405.20485
-
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs 30 May 2024 · 0 repositories · arXiv:2405.20179
-
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation 30 May 2024 · 0 repositories · arXiv:2405.19701
-
Student Answer Forecasting: Transformer-Driven Answer Choice Prediction for Language Learning 30 May 2024 · 1 repository · arXiv:2405.20079
-
Towards Ontology-Enhanced Representation Learning for Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20527
-
A Multi-Source Retrieval Question Answering Framework Based on RAG 29 May 2024 · 0 repositories · arXiv:2405.19207
-
Beyond Agreement: Diagnosing the Rationale Alignment of Automated Essay Scoring Methods based on Linguistically-informed Counterfactuals 29 May 2024 · 1 repository · arXiv:2405.19433
-
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension 29 May 2024 · 0 repositories · arXiv:2405.18682
-
CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control 29 May 2024 · 1 repository · arXiv:2405.18727Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
Efficient Model-agnostic Alignment via Bayesian Persuasion 29 May 2024 · 0 repositories · arXiv:2405.18718
-
Faster Cascades via Speculative Decoding 29 May 2024 · 0 repositories · arXiv:2405.19261
-
LMO-DP: Optimizing the Randomization Mechanism for Differentially Private Fine-Tuning (Large) Language Models 29 May 2024 · 0 repositories · arXiv:2405.18776
-
MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series 29 May 2024 · 1 repository · arXiv:2405.19327Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Offline Regularised Reinforcement Learning for Large Language Models Alignment 29 May 2024 · 0 repositories · arXiv:2405.19107
-
STAT: Shrinking Transformers After Training 29 May 2024 · 0 repositories · arXiv:2406.00061
-
Toward Conversational Agents with Context and Time Sensitive Long-term Memory 29 May 2024 · 1 repository · arXiv:2406.00057
-
Two-Layer Retrieval-Augmented Generation Framework for Low-Resource Medical Question Answering Using Reddit Data: Proof-of-Concept Study 29 May 2024 · 0 repositories · arXiv:2405.19519
-
Aligning to Thousands of Preferences via System Message Generalization 28 May 2024 · 1 repository · arXiv:2405.17977Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
An Empirical Analysis on Large Language Models in Debate Evaluation 28 May 2024 · 1 repository · arXiv:2406.00050Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Are PPO-ed Language Models Hackable? 28 May 2024 · 0 repositories · arXiv:2406.02577
-
ATM: Adversarial Tuning Multi-agent System Makes a Robust Retrieval-Augmented Generator 28 May 2024 · 1 repository · arXiv:2405.18111Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Attention-based sequential recommendation system using multimodal data 28 May 2024 · 0 repositories · arXiv:2405.17959
-
Don't Forget to Connect! Improving RAG with Graph-based Reranking 28 May 2024 · 0 repositories · arXiv:2405.18414
-
Edinburgh Clinical NLP at MEDIQA-CORR 2024: Guiding Large Language Models with Hints 28 May 2024 · 0 repositories · arXiv:2405.18028
-
LLMs and Memorization: On Quality and Specificity of Copyright Compliance 28 May 2024 · 1 repository · arXiv:2405.18492
-
Understanding Intrinsic Socioeconomic Biases in Large Language Models 28 May 2024 · 0 repositories · arXiv:2405.18662
-
WIDIn: Wording Image for Domain-Invariant Representation in Single-Source Domain Generalization 28 May 2024 · 0 repositories · arXiv:2405.18405
-
Assessing LLMs Suitability for Knowledge Graph Completion 27 May 2024 · 1 repository · arXiv:2405.17249
-
Augmenting Textual Generation via Topology Aware Retrieval 27 May 2024 · 0 repositories · arXiv:2405.17602
-
DeeperImpact: Optimizing Sparse Learned Index Structures 27 May 2024 · 1 repository · arXiv:2405.17093
-
Detecting Deceptive Dark Patterns in E-commerce Platforms 27 May 2024 · 0 repositories · arXiv:2406.01608
-
Exploiting the Layered Intrinsic Dimensionality of Deep Models for Practical Adversarial Training 27 May 2024 · 0 repositories · arXiv:2405.17130
-
InversionView: A General-Purpose Method for Reading Information from Neural Activations 27 May 2024 · 1 repository · arXiv:2405.17653Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; the one sample that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
REVECA: Adaptive Planning and Trajectory-based Validation in Cooperative Language Agents using Information Relevance and Relative Proximity 27 May 2024 · 0 repositories · arXiv:2405.16751
-
NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models 27 May 2024 · 0 repositories · arXiv:2405.17428
-
PAE: LLM-based Product Attribute Extraction for E-Commerce Fashion Trends 27 May 2024 · 0 repositories · arXiv:2405.17533
-
Performance evaluation of Reddit Comments using Machine Learning and Natural Language Processing methods in Sentiment Analysis 27 May 2024 · 0 repositories · arXiv:2405.16810
-
QUB-Cirdan at "Discharge Me!": Zero shot discharge letter generation by open-source LLM 27 May 2024 · 0 repositories · arXiv:2406.00041
-
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation 27 May 2024 · 1 repository · arXiv:2405.17057Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
RTL-Repo: A Benchmark for Evaluating LLMs on Large-Scale RTL Design Projects 27 May 2024 · 1 repository · arXiv:2405.17378
-
The Scaling Law in Stellar Light Curves 27 May 2024 · 0 repositories · arXiv:2405.17156
-
THREAD: Thinking Deeper with Recursive Spawning 27 May 2024 · 1 repository · arXiv:2405.17402Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Video Enriched Retrieval Augmented Generation Using Aligned Video Captions 27 May 2024 · 1 repository · arXiv:2405.17706
-
AI-Generated Text Detection and Classification Based on BERT Deep Learning Algorithm 26 May 2024 · 0 repositories · arXiv:2405.16422
-
GRAG: Graph Retrieval-Augmented Generation 26 May 2024 · 1 repository · arXiv:2405.16506
-
M-RAG: Reinforcing Large Language Model Performance through Retrieval-Augmented Generation with Multiple Partitions 26 May 2024 · 0 repositories · arXiv:2405.16420
-
Accelerating Inference of Retrieval-Augmented Generation via Sparse Context Selection 25 May 2024 · 0 repositories · arXiv:2405.16178
-
Accelerating Transformers with Spectrum-Preserving Token Merging 25 May 2024 · 1 repository · arXiv:2405.16148Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
AutoManual: Constructing Instruction Manuals by LLM Agents via Interactive Environmental Learning 25 May 2024 · 1 repository · arXiv:2405.16247Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Incremental Comprehension of Garden-Path Sentences by Large Language Models: Semantic Interpretation, Syntactic Re-Analysis, and Attention 25 May 2024 · 0 repositories · arXiv:2405.16042
-
MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time 25 May 2024 · 0 repositories · arXiv:2405.16265
-
Towards Unlocking Insights from Logbooks Using AI 25 May 2024 · 0 repositories · arXiv:2406.12881
-
An Evaluation of Estimative Uncertainty in Large Language Models 24 May 2024 · 0 repositories · arXiv:2405.15185
-
Benchmarking the Performance of Pre-trained LLMs across Urdu NLP Tasks 24 May 2024 · 0 repositories · arXiv:2405.15453
-
CulturePark: Boosting Cross-cultural Understanding in Large Language Models 24 May 2024 · 1 repository · arXiv:2405.15145Syntology official: harvested, nothing ran · 0 ran · 7 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Enhancing Augmentative and Alternative Communication with Card Prediction and Colourful Semantics 24 May 2024 · 0 repositories · arXiv:2405.15896