Methods › General › Regularization › Attention Dropout › Papers, page 17
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 17 of 109: papers 1,601 to 1,700 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Can Large Language Model Predict Employee Attrition? 2 Nov 2024 · 0 repositories · arXiv:2411.01353
-
Enhancing Neural Network Interpretability with Feature-Aligned Sparse Autoencoders 2 Nov 2024 · 1 repository · arXiv:2411.01220Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AttackQA: Development and Adoption of a Dataset for Assisting Cybersecurity Operations using Fine-tuned and Open-Source LLMs 1 Nov 2024 · 0 repositories · arXiv:2411.01073
-
CORAG: A Cost-Constrained Retrieval Optimization System for Retrieval-Augmented Generation 1 Nov 2024 · 0 repositories · arXiv:2411.00744
-
Evaluating the Impact of Lab Test Results on Large Language Models Generated Differential Diagnoses from Clinical Case Vignettes 1 Nov 2024 · 0 repositories · arXiv:2411.02523
-
LLM-Ref: Enhancing Reference Handling in Technical Writing with Large Language Models 1 Nov 2024 · 0 repositories · arXiv:2411.00294
-
LLMs: A Game-Changer for Software Engineers? 1 Nov 2024 · 0 repositories · arXiv:2411.00932
-
Provenance: A Light-weight Fact-checker for Retrieval Augmented LLM Generation Output 1 Nov 2024 · 0 repositories · arXiv:2411.01022
-
Rationale-Guided Retrieval Augmented Generation for Medical Question Answering 1 Nov 2024 · 1 repository · arXiv:2411.00300Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Towards Multi-Source Retrieval-Augmented Generation via Synergizing Reasoning and Preference-Driven Retrieval 1 Nov 2024 · 0 repositories · arXiv:2411.00689
-
Analyzing & Reducing the Need for Learning Rate Warmup in GPT Training 31 Oct 2024 · 0 repositories · arXiv:2410.23922
-
Automating Quantum Software Maintenance: Flakiness Detection and Root Cause Analysis 31 Oct 2024 · 0 repositories · arXiv:2410.23578
-
JudgeRank: Leveraging Large Language Models for Reasoning-Intensive Reranking 31 Oct 2024 · 0 repositories · arXiv:2411.00142
-
LEAF: Learning and Evaluation Augmented by Fact-Checking to Improve Factualness in Large Language Models 31 Oct 2024 · 0 repositories · arXiv:2410.23526
-
Responsible Retrieval Augmented Generation for Climate Decision Making from Documents 31 Oct 2024 · 0 repositories · arXiv:2410.23902
-
SelfCodeAlign: Self-Alignment for Code Generation 31 Oct 2024 · 2 repositories · arXiv:2410.24198Syntology official (archive's flag): 9 ran · 30 ran (of which 3 constructed an object rather than computing a result; 22 with no instrument failure: 1 honoured, 0 violated, 21 with no contract checked; 8 where Syntology's instrument failed) · 7 unverified (of 37 harvested samples)
-
A Comprehensive Study on Quantization Techniques for Large Language Models 30 Oct 2024 · 0 repositories · arXiv:2411.02530
-
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation 30 Oct 2024 · 1 repository · arXiv:2410.23090
-
Eliciting Critical Reasoning in Retrieval-Augmented Language Models via Contrastive Explanations 30 Oct 2024 · 0 repositories · arXiv:2410.22874
-
Emotional RAG: Enhancing Role-Playing Agents through Emotional Retrieval 30 Oct 2024 · 1 repository · arXiv:2410.23041
-
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models 30 Oct 2024 · 0 repositories · arXiv:2410.22832
-
ProTransformer: Robustify Transformers via Plug-and-Play Paradigm 30 Oct 2024 · 1 repository · arXiv:2410.23182
-
Retrieval-Augmented Generation with Estimation of Source Reliability 30 Oct 2024 · 0 repositories · arXiv:2410.22954
-
Semantic Enrichment of the Quantum Cascade Laser Properties in Text- A Knowledge Graph Generation Approach 30 Oct 2024 · 1 repository · arXiv:2410.22996
-
Long²RAG: Evaluating Long-Context & Long-Form Retrieval-Augmented Generation with Key Point Recall 30 Oct 2024 · 0 repositories · arXiv:2410.23000
-
Abrupt Learning in Transformers: A Case Study on Matrix Completion 29 Oct 2024 · 0 repositories · arXiv:2410.22244Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Beyond Text: Optimizing RAG with Multimodal Inputs for Industrial Applications 29 Oct 2024 · 1 repository · arXiv:2410.21943Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
CFSafety: Comprehensive Fine-grained Safety Assessment for LLMs 29 Oct 2024 · 0 repositories · arXiv:2410.21695
-
Coupling quantum-like cognition with the neuronal networks within generalized probability theory 29 Oct 2024 · 0 repositories · arXiv:2411.00036
-
FactBench: A Dynamic Benchmark for In-the-Wild Language Model Factuality Evaluation 29 Oct 2024 · 0 repositories · arXiv:2410.22257
-
Meta-Learning Adaptable Foundation Models 29 Oct 2024 · 0 repositories · arXiv:2410.22264
-
Sequential choice in ordered bundles 29 Oct 2024 · 0 repositories · arXiv:2410.21670
-
A Simple Yet Effective Corpus Construction Framework for Indonesian Grammatical Error Correction 28 Oct 2024 · 1 repository · arXiv:2410.20838
-
AutoRAG: Automated Framework for optimization of Retrieval Augmented Generation Pipeline 28 Oct 2024 · 2 repositories · arXiv:2410.20878
-
BanditCAT and AutoIRT: Machine Learning Approaches to Computerized Adaptive Testing and Item Calibration 28 Oct 2024 · 0 repositories · arXiv:2410.21033
-
BLAST: Block-Level Adaptive Structured Matrices for Efficient Deep Neural Network Inference 28 Oct 2024 · 1 repository · arXiv:2410.21262
-
Calibrated Decision-Making through LLM-Assisted Retrieval 28 Oct 2024 · 0 repositories · arXiv:2411.08891
-
Causal Interventions on Causal Paths: Mapping GPT-2's Reasoning From Syntax to Semantics 28 Oct 2024 · 0 repositories · arXiv:2410.21353
-
Combining Domain-Specific Models and LLMs for Automated Disease Phenotyping from Survey Data 28 Oct 2024 · 0 repositories · arXiv:2410.20695
-
CRAT: A Multi-Agent Framework for Causality-Enhanced Reflective and Retrieval-Augmented Translation with Large Language Models 28 Oct 2024 · 0 repositories · arXiv:2410.21067
-
Deep Learning for Medical Text Processing: BERT Model Fine-Tuning and Comparative Study 28 Oct 2024 · 0 repositories · arXiv:2410.20792
-
Embedding with Large Language Models for Classification of HIPAA Safeguard Compliance Rules 28 Oct 2024 · 0 repositories · arXiv:2410.20664
-
Geo-FuB: A Method for Constructing an Operator-Function Knowledge Base for Geospatial Code Generation Tasks Using Large Language Models 28 Oct 2024 · 1 repository · arXiv:2410.20975
-
Is GPT-4 Less Politically Biased than GPT-3.5? A Renewed Investigation of ChatGPT's Political Biases 28 Oct 2024 · 0 repositories · arXiv:2410.21008
-
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation 28 Oct 2024 · 1 repository · arXiv:2410.20777
-
LinFormer: A Linear-based Lightweight Transformer Architecture For Time-Aware MIMO Channel Prediction 28 Oct 2024 · 0 repositories · arXiv:2410.21351
-
LLMs are Biased Evaluators But Not Biased for Retrieval Augmented Generation 28 Oct 2024 · 1 repository · arXiv:2410.20833
-
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression 28 Oct 2024 · 1 repository · arXiv:2410.21548
-
Plan×RAG: Planning-guided Retrieval Augmented Generation 28 Oct 2024 · 0 repositories · arXiv:2410.20753
-
Semantic Search Evaluation 28 Oct 2024 · 0 repositories · arXiv:2410.21549
-
Simple Is Effective: The Roles of Graphs and Large Language Models in Knowledge-Graph-Based Retrieval-Augmented Generation 28 Oct 2024 · 1 repository · arXiv:2410.20724Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
Stealthy Jailbreak Attacks on Large Language Models via Benign Data Mirroring 28 Oct 2024 · 0 repositories · arXiv:2410.21083
-
uOttawa at LegalLens-2024: Transformer-based Classification Experiments 28 Oct 2024 · 1 repository · arXiv:2410.21139
-
Deep Learning Based Dense Retrieval: A Comparative Study 27 Oct 2024 · 0 repositories · arXiv:2410.20315
-
LLM Robustness Against Misinformation in Biomedical Question Answering 27 Oct 2024 · 1 repository · arXiv:2410.21330
-
R^3AG: First Workshop on Refined and Reliable Retrieval Augmented Generation 27 Oct 2024 · 0 repositories · arXiv:2410.20598
-
Sequential Large Language Model-Based Hyper-parameter Optimization 27 Oct 2024 · 1 repository · arXiv:2410.20302
-
Mask-based Membership Inference Attacks for Retrieval-Augmented Generation 26 Oct 2024 · 0 repositories · arXiv:2410.20142
-
Think Carefully and Check Again! Meta-Generation Unlocking LLMs for Low-Resource Cross-Lingual Summarization 26 Oct 2024 · 0 repositories · arXiv:2410.20021
-
A Tutorial on Teaching Data Analytics with Generative AI 25 Oct 2024 · 0 repositories · arXiv:2411.07244
-
ChunkRAG: Novel LLM-Chunk Filtering Method for RAG Systems 25 Oct 2024 · 0 repositories · arXiv:2410.19572
-
FISHNET: Financial Intelligence from Sub-querying, Harmonizing, Neural-Conditioning, Expert Swarms, and Task Planning 25 Oct 2024 · 0 repositories · arXiv:2410.19727
-
GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing 25 Oct 2024 · 1 repository · arXiv:2410.19552Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Integrating Large Language Models with Internet of Things Applications 25 Oct 2024 · 0 repositories · arXiv:2410.19223
-
Bielik 7B v0.1: A Polish Language Model -- Development, Insights, and Evaluation 24 Oct 2024 · 0 repositories · arXiv:2410.18565
-
Difficult for Whom? A Study of Japanese Lexical Complexity 24 Oct 2024 · 1 repository · arXiv:2410.18567
-
Iterative Self-Tuning LLMs for Enhanced Jailbreaking Capabilities 24 Oct 2024 · 1 repository · arXiv:2410.18469
-
Little Giants: Synthesizing High-Quality Embedding Data at Scale 24 Oct 2024 · 1 repository · arXiv:2410.18634Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
PDL: A Declarative Prompt Programming Language 24 Oct 2024 · 1 repository · arXiv:2410.19135
-
Understanding Ranking LLMs: A Mechanistic Analysis for Information Retrieval 24 Oct 2024 · 0 repositories · arXiv:2410.18527
-
Scaling up Masked Diffusion Models on Text 24 Oct 2024 · 1 repository · arXiv:2410.18514Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Understanding Players as if They Are Talking to the Game in a Customized Language: A Pilot Study 24 Oct 2024 · 0 repositories · arXiv:2410.18605
-
An Adaptive Framework for Generating Systematic Explanatory Answer in Online Q&A Platforms 23 Oct 2024 · 1 repository · arXiv:2410.17694
-
Differentially Private Learning Needs Better Model Initialization and Self-Distillation 23 Oct 2024 · 1 repository · arXiv:2410.17566
-
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction 23 Oct 2024 · 0 repositories · arXiv:2410.18160
-
Leveraging the Domain Adaptation of Retrieval Augmented Generation Models for Question Answering and Reducing Hallucination 23 Oct 2024 · 0 repositories · arXiv:2410.17783
-
Small Singular Values Matter: A Random Matrix Analysis of Transformer Models 23 Oct 2024 · 0 repositories · arXiv:2410.17770
-
LongRAG: A Dual-Perspective Retrieval-Augmented Generation Paradigm for Long-Context Question Answering 23 Oct 2024 · 1 repository · arXiv:2410.18050Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
MCUBERT: Memory-Efficient BERT Inference on Commodity Microcontrollers 23 Oct 2024 · 0 repositories · arXiv:2410.17957
-
OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation 23 Oct 2024 · 1 repository · arXiv:2410.17799
-
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains 23 Oct 2024 · 0 repositories · arXiv:2410.17952
-
Dhoroni: Exploring Bengali Climate Change and Environmental Views with a Multi-Perspective News Dataset and Natural Language Processing 22 Oct 2024 · 1 repository · arXiv:2410.17225
-
Distill-SynthKG: Distilling Knowledge Graph Synthesis Workflow for Improved Coverage and Efficiency 22 Oct 2024 · 0 repositories · arXiv:2410.16597
-
DNAHLM -- DNA sequence and Human Language mixed large language Model 22 Oct 2024 · 1 repository · arXiv:2410.16917
-
Exploring Possibilities of AI-Powered Legal Assistance in Bangladesh through Large Language Modeling 22 Oct 2024 · 1 repository · arXiv:2410.17210
-
Scattered Forest Search: Smarter Code Space Exploration with LLMs 22 Oct 2024 · 0 repositories · arXiv:2411.05010
-
SmartRAG: Jointly Learn RAG-Related Tasks From the Environment Feedback 22 Oct 2024 · 0 repositories · arXiv:2410.18141
-
Tracing the Development of the Virtual Particle Concept Using Semantic Change Detection 22 Oct 2024 · 1 repository · arXiv:2410.16855
-
An Efficient System for Automatic Map Storytelling -- A Case Study on Historical Maps 21 Oct 2024 · 1 repository · arXiv:2410.15780
-
Building A Coding Assistant via the Retrieval-Augmented Language Model 21 Oct 2024 · 1 repository · arXiv:2410.16229
-
Deep Learning and Data Augmentation for Detecting Self-Admitted Technical Debt 21 Oct 2024 · 1 repository · arXiv:2410.15804
-
Developing Retrieval Augmented Generation (RAG) based LLM Systems from PDFs: An Experience Report 21 Oct 2024 · 1 repository · arXiv:2410.15944
-
Exploring Pretraining via Active Forgetting for Improving Cross Lingual Transfer for Decoder Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16168
-
Guardians of Discourse: Evaluating LLMs on Multilingual Offensive Language Detection 21 Oct 2024 · 0 repositories · arXiv:2410.15623
-
Improving Neuron-level Interpretability with White-box Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16443
-
Leveraging Retrieval-Augmented Generation for Culturally Inclusive Hakka Chatbots: Design Insights and User Perceptions 21 Oct 2024 · 0 repositories · arXiv:2410.15572
-
LightFusionRec: Lightweight Transformers-Based Cross-Domain Recommendation Model 21 Oct 2024 · 0 repositories · arXiv:2410.15656
-
Natural GaLore: Accelerating GaLore for memory-efficient LLM Training and Fine-tuning 21 Oct 2024 · 1 repository · arXiv:2410.16029
-
On Creating an English-Thai Code-switched Machine Translation in Medical Domain 21 Oct 2024 · 1 repository · arXiv:2410.16221
-
RAG4ITOps: A Supervised Fine-Tunable and Comprehensive RAG Framework for IT Operations and Maintenance 21 Oct 2024 · 0 repositories · arXiv:2410.15805