Methods › General › Regularization › Attention Dropout
Attention Dropout
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the softmax in the attention equation. For example, for scaled-dot product attention, we would drop elements from the first term:
Attention(Q, K, V) = softmax(QKᵀ/(√(dₖ)))V
Papers archive 2025-07-28
30 shown of 10,892, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Making Language Model a Hierarchical Classifier and Generator 17 Jul 2025 · 1 repository · arXiv:2507.12930
-
Generative Click-through Rate Prediction with Applications to Search Advertising 15 Jul 2025 · 0 repositories · arXiv:2507.11246
-
Chat-Ghosting: A Comparative Study of Methods for Auto-Completion in Dialog Systems 8 Jul 2025 · 0 repositories · arXiv:2507.05940
-
SARA: Selective and Adaptive Retrieval-augmented Generation with Context Compression 8 Jul 2025 · 0 repositories · arXiv:2507.05633
-
AI Generated Text Detection Using Instruction Fine-tuned Large Language and Transformer-Based Models 7 Jul 2025 · 0 repositories · arXiv:2507.05157
-
Behaviour Space Analysis of LLM-driven Meta-heuristic Discovery 4 Jul 2025 · 0 repositories · arXiv:2507.03605
-
CyberRAG: An agentic RAG cyber attack classification and reporting tool 3 Jul 2025 · 0 repositories · arXiv:2507.02424
-
Knowledge Protocol Engineering: A New Paradigm for AI in Domain-Specific Knowledge Work 3 Jul 2025 · 0 repositories · arXiv:2507.02760
-
Robustness of Misinformation Classification Systems to Adversarial Examples Through BeamAttack 30 Jun 2025 · 1 repository · arXiv:2506.23661
-
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models 28 Jun 2025 · 0 repositories · arXiv:2506.22957
-
Cat and Mouse -- Can Fake Text Generation Outpace Detector Systems? 26 Jun 2025 · 0 repositories · arXiv:2506.21274
-
EraRAG: Efficient and Incremental Retrieval Augmented Generation for Growing Corpora 26 Jun 2025 · 1 repository · arXiv:2506.20963
-
Large Language Models Acing Chartered Accountancy 26 Jun 2025 · 0 repositories · arXiv:2506.21031
-
Leveraging LLM-Assisted Query Understanding for Live Retrieval-Augmented Generation 26 Jun 2025 · 0 repositories · arXiv:2506.21384
-
PsyLite Technical Report 26 Jun 2025 · 1 repository · arXiv:2506.21536
-
Response Quality Assessment for Retrieval-Augmented Generation via Conditional Conformal Factuality 26 Jun 2025 · 1 repository · arXiv:2506.20978Syntology ran 0 of 3 samples · 3 unverified · 3 pointer-only (licence)
-
AI Assistants to Enhance and Exploit the PETSc Knowledge Base 25 Jun 2025 · 0 repositories · arXiv:2506.20608
-
CCRS: A Zero-Shot LLM-as-a-Judge Framework for Comprehensive RAG Evaluation 25 Jun 2025 · 0 repositories · arXiv:2506.20128
-
Engineering RAG Systems for Real-World Applications: Design, Development, and Evaluation 25 Jun 2025 · 0 repositories · arXiv:2506.20869
-
Knowledge-Aware Diverse Reranking for Cross-Source Question Answering 25 Jun 2025 · 0 repositories · arXiv:2506.20476
-
Large Language Model-Driven Code Compliance Checking in Building Information Modeling 25 Jun 2025 · 0 repositories · arXiv:2506.20551
-
Memento: Note-Taking for Your Future Self 25 Jun 2025 · 0 repositories · arXiv:2506.20642
-
Accurate and Energy Efficient: Local Retrieval-Augmented Generation Models Outperform Commercial Large Language Models in Medical Tasks 24 Jun 2025 · 0 repositories · arXiv:2506.20009
-
Controlled Retrieval-augmented Context Evaluation for Long-form RAG 24 Jun 2025 · 0 repositories · arXiv:2506.20051
-
Inference Scaled GraphRAG: Improving Multi Hop Question Answering on Knowledge Graphs 24 Jun 2025 · 0 repositories · arXiv:2506.19967
-
KunLunBaizeRAG: Reinforcement Learning Driven Inference Performance Leap for Large Language Models 24 Jun 2025 · 0 repositories · arXiv:2506.19466
-
Unlocking Insights Addressing Alcohol Inference Mismatch through Database-Narrative Alignment 24 Jun 2025 · 0 repositories · arXiv:2506.19342
-
An Audio-centric Multi-task Learning Framework for Streaming Ads Targeting on Spotify 23 Jun 2025 · 0 repositories · arXiv:2506.18735
-
Semantic similarity estimation for domain specific data using BERT and other techniques 23 Jun 2025 · 0 repositories · arXiv:2506.18602
-
T-CPDL: A Temporal Causal Probabilistic Description Logic for Developing Logic-RAG Agent 23 Jun 2025 · 0 repositories · arXiv:2506.18559
Tasks archive 2025-07-28
20 shown of 1,501 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
| Task | Papers |
|---|---|
| Language Modelling | 1,902 |
| Language Modeling | 1,489 |
| Retrieval | 1,433 |
| RAG | 1,308 |
| Retrieval-augmented Generation | 1,122 |
| Question Answering | 1,070 |
| Sentence | 920 |
| Large Language Model | 501 |
| Sentiment Analysis | 470 |
| Text Generation | 470 |
| Text Classification | 459 |
| text-classification | 409 |
| Transfer Learning | 393 |
| Natural Language Understanding | 311 |
| Information Retrieval | 302 |
| Classification | 299 |
| Decoder | 295 |
| Translation | 294 |
| Word Embeddings | 280 |
| Named Entity Recognition | 273 |
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections