Methods › General › Regularization › Attention Dropout › Papers, page 33
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 33 of 109: papers 3,201 to 3,300 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Do Large Language Models Rank Fairly? An Empirical Study on the Fairness of LLMs as Rankers 4 Apr 2024 · 0 repositories · arXiv:2404.03192
-
NLP at UC Santa Cruz at SemEval-2024 Task 5: Legal Answer Validation using Few-Shot Multi-Choice QA 4 Apr 2024 · 1 repository · arXiv:2404.03150
-
Outlier-Efficient Hopfield Layers for Large Transformer-Based Models 4 Apr 2024 · 1 repository · arXiv:2404.03828Syntology official (archive's flag): 10 ran · 10 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
The Death of Feature Engineering? BERT with Linguistic Features on SQuAD 2.0 4 Apr 2024 · 0 repositories · arXiv:2404.03184
-
Adaptive Cross-lingual Text Classification through In-Context One-Shot Demonstrations 3 Apr 2024 · 1 repository · arXiv:2404.02452Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AI-Tutoring in Software Engineering Education 3 Apr 2024 · 0 repositories · arXiv:2404.02548
-
An Incomplete Loop: Deductive, Inductive, and Abductive Learning in Large Language Models 3 Apr 2024 · 0 repositories · arXiv:2404.03028
-
BCAmirs at SemEval-2024 Task 4: Beyond Words: A Multimodal and Multilingual Exploration of Persuasion in Memes 3 Apr 2024 · 1 repository · arXiv:2404.03022
-
Benchmarking Large Language Models for Persian: A Preliminary Study Focusing on ChatGPT 3 Apr 2024 · 1 repository · arXiv:2404.02403
-
GPT-DETOX: An In-Context Learning-Based Paraphraser for Text Detoxification 3 Apr 2024 · 0 repositories · arXiv:2404.03052
-
On Linearizing Structured Data in Encoder-Decoder Language Models: Insights from Text-to-SQL 3 Apr 2024 · 0 repositories · arXiv:2404.02389
-
uTeBC-NLP at SemEval-2024 Task 9: Can LLMs be Lateral Thinkers? 3 Apr 2024 · 1 repository · arXiv:2404.02474Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Advancing LLM Reasoning Generalists with Preference Trees 2 Apr 2024 · 1 repository · arXiv:2404.02078Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 1 violated, 12 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 20 harvested samples) · 2 pointer-only (licence)
-
Stereotype Detection in LLMs: A Multiclass, Explainable, and Benchmark-Driven Approach 2 Apr 2024 · 0 repositories · arXiv:2404.01768
-
CLAPNQ: Cohesive Long-form Answers from Passages in Natural Questions for RAG systems 2 Apr 2024 · 1 repository · arXiv:2404.02103
-
CMAT: A Multi-Agent Collaboration Tuning Framework for Enhancing Small Language Models 2 Apr 2024 · 1 repository · arXiv:2404.01663
-
Collapse of Self-trained Language Models 2 Apr 2024 · 1 repository · arXiv:2404.02305Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Comparative Study of Domain Driven Terms Extraction Using Large Language Models 2 Apr 2024 · 0 repositories · arXiv:2404.02330
-
Deconstructing In-Context Learning: Understanding Prompts via Corruption 2 Apr 2024 · 1 repository · arXiv:2404.02054
-
GINopic: Topic Modeling with Graph Isomorphism Network 2 Apr 2024 · 1 repository · arXiv:2404.02115Syntology official: no sample here; runs from other or unrecorded repositories · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks 2 Apr 2024 · 1 repository · arXiv:2404.02151Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
METAL: Towards Multilingual Meta-Evaluation 2 Apr 2024 · 0 repositories · arXiv:2404.01667
-
Symbolic Prompt Program Search: A Structure-Aware Approach to Efficient Compile-Time Prompt Optimization 2 Apr 2024 · 1 repository · arXiv:2404.02319
-
Scene Adaptive Sparse Transformer for Event-based Object Detection 2 Apr 2024 · 1 repository · arXiv:2404.01882Syntology official (archive's flag): 19 ran · 19 ran (of which 0 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 23 harvested samples)
-
SGSH: Stimulate Large Language Models with Skeleton Heuristics for Knowledge Base Question Generation 2 Apr 2024 · 1 repository · arXiv:2404.01923
-
Toward Informal Language Processing: Knowledge of Slang in Large Language Models 2 Apr 2024 · 1 repository · arXiv:2404.02323
-
Advancing AI with Integrity: Ethical Challenges and Solutions in Neural Machine Translation 1 Apr 2024 · 0 repositories · arXiv:2404.01070
-
ARAGOG: Advanced RAG Output Grading 1 Apr 2024 · 1 repository · arXiv:2404.01037
-
Artificial Intelligence and the Spatial Documentation of Languages 1 Apr 2024 · 0 repositories · arXiv:2404.01263
-
BERT-Enhanced Retrieval Tool for Homework Plagiarism Detection System 1 Apr 2024 · 0 repositories · arXiv:2404.01582
-
FABLES: Evaluating faithfulness and content selection in book-length summarization 1 Apr 2024 · 3 repositories · arXiv:2404.01261Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Securing Social Spaces: Harnessing Deep Learning to Eradicate Cyberbullying 1 Apr 2024 · 0 repositories · arXiv:2404.03686
-
Unveiling Divergent Inductive Biases of LLMs on Temporal Data 1 Apr 2024 · 1 repository · arXiv:2404.01453
-
A General and Efficient Training for Transformer via Token Expansion 31 Mar 2024 · 1 repository · arXiv:2404.00672Syntology official (archive's flag): 7 ran · 7 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
CHOPS: CHat with custOmer Profile Systems for Customer Service with LLMs 31 Mar 2024 · 1 repository · arXiv:2404.01343
-
CoUDA: Coherence Evaluation via Unified Data Augmentation 31 Mar 2024 · 1 repository · arXiv:2404.00681
-
Observations on Building RAG Systems for Technical Documents 31 Mar 2024 · 0 repositories · arXiv:2404.00657
-
RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation 31 Mar 2024 · 1 repository · arXiv:2404.00610
-
Training-Free Semantic Segmentation via LLM-Supervision 31 Mar 2024 · 0 repositories · arXiv:2404.00701
-
A Comprehensive Study on NLP Data Augmentation for Hate Speech Detection: Legacy Methods, BERT, and LLMs 30 Mar 2024 · 0 repositories · arXiv:2404.00303
-
Leveraging Pre-trained and Transformer-derived Embeddings from EHRs to Characterize Heterogeneity Across Alzheimer's Disease and Related Dementias 30 Mar 2024 · 0 repositories · arXiv:2404.00464
-
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks 30 Mar 2024 · 0 repositories · arXiv:2404.00376
-
ChatGPT v.s. Media Bias: A Comparative Study of GPT-3.5 and Fine-tuned Language Models 29 Mar 2024 · 0 repositories · arXiv:2403.20158
-
DataAgent: Evaluating Large Language Models' Ability to Answer Zero-Shot, Natural Language Queries 29 Mar 2024 · 0 repositories · arXiv:2404.00188
-
Explainable Deep Learning: A Visual Analytics Approach with Transition Matrices 29 Mar 2024 · 1 repository
-
LayerNorm: A key component in parameter-efficient fine-tuning 29 Mar 2024 · 0 repositories · arXiv:2403.20284
-
ReALM: Reference Resolution As Language Modeling 29 Mar 2024 · 0 repositories · arXiv:2403.20329
-
Shallow Cross-Encoders for Low-Latency Retrieval 29 Mar 2024 · 1 repository · arXiv:2403.20222
-
A Review of Multi-Modal Large Language and Vision Models 28 Mar 2024 · 0 repositories · arXiv:2404.01322
-
AlloyBERT: Alloy Property Prediction with Large Language Models 28 Mar 2024 · 0 repositories · arXiv:2403.19783
-
Are Large Language Models Good at Utility Judgments? 28 Mar 2024 · 1 repository · arXiv:2403.19216Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
FACTOID: FACtual enTailment fOr hallucInation Detection 28 Mar 2024 · 0 repositories · arXiv:2403.19113
-
Generating Multi-Aspect Queries for Conversational Search 28 Mar 2024 · 0 repositories · arXiv:2403.19302
-
Intelligent Classification and Personalized Recommendation of E-commerce Products Based on Machine Learning 28 Mar 2024 · 0 repositories · arXiv:2403.19345
-
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models 28 Mar 2024 · 1 repository · arXiv:2403.19521Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Just-DNA-Seq, open-source personal genomics platform: longevity science for everyone 28 Mar 2024 · 0 repositories · arXiv:2403.19087
-
Mitigating Misleading Chain-of-Thought Reasoning with Selective Filtering 28 Mar 2024 · 1 repository · arXiv:2403.19167
-
Risk prediction of pathological gambling on social media 28 Mar 2024 · 0 repositories · arXiv:2403.19358
-
A Novel Corpus of Annotated Medical Imaging Reports and Information Extraction Results Using BERT-based Language Models 27 Mar 2024 · 1 repository · arXiv:2403.18975
-
A Survey on Large Language Models from Concept to Implementation 27 Mar 2024 · 0 repositories · arXiv:2403.18969
-
AcTED: Automatic Acquisition of Typical Event Duration for Semi-supervised Temporal Commonsense QA 27 Mar 2024 · 0 repositories · arXiv:2403.18504
-
Boosting Conversational Question Answering with Fine-Grained Retrieval-Augmentation and Self-Check 27 Mar 2024 · 0 repositories · arXiv:2403.18243
-
CPR: Retrieval Augmented Generation for Copyright Protection 27 Mar 2024 · 0 repositories · arXiv:2403.18920
-
Evaluating Large Language Models for Health-Related Text Classification Tasks with Public Social Media Data 27 Mar 2024 · 0 repositories · arXiv:2403.19031
-
Fusion approaches for emotion recognition from speech using acoustic and text-based features 27 Mar 2024 · 0 repositories · arXiv:2403.18635
-
LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research Practices 27 Mar 2024 · 0 repositories · arXiv:2403.18173
-
Long-form factuality in large language models 27 Mar 2024 · 3 repositories · arXiv:2403.18802Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
ParCo: Part-Coordinating Text-to-Motion Synthesis 27 Mar 2024 · 1 repository · arXiv:2403.18512
-
Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers 27 Mar 2024 · 1 repository · arXiv:2403.18938
-
SemRoDe: Macro Adversarial Training to Learn Representations That are Robust to Word-Level Attacks 27 Mar 2024 · 1 repository · arXiv:2403.18423Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Vulnerability Detection with Code Language Models: How Far Are We? 27 Mar 2024 · 1 repository · arXiv:2403.18624Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Are Compressed Language Models Less Subgroup Robust? 26 Mar 2024 · 1 repository · arXiv:2403.17811
-
Decoding Probing: Revealing Internal Linguistic Structures in Neural Language Models using Minimal Pairs 26 Mar 2024 · 0 repositories · arXiv:2403.17299
-
Disambiguate Entity Matching using Large Language Models through Relation Discovery 26 Mar 2024 · 0 repositories · arXiv:2403.17344
-
Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization 26 Mar 2024 · 1 repository · arXiv:2403.18120Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
ELLEN: Extremely Lightly Supervised Learning For Efficient Named Entity Recognition 26 Mar 2024 · 1 repository · arXiv:2403.17385
-
Enhancing Legal Document Retrieval: A Multi-Phase Approach with Large Language Models 26 Mar 2024 · 0 repositories · arXiv:2403.18093
-
Fingerprinting web servers through Transformer-encoded HTTP response headers 26 Mar 2024 · 1 repository · arXiv:2404.00056
-
Hierarchical Multi-label Classification for Fine-level Event Extraction from Aviation Accident Reports 26 Mar 2024 · 0 repositories · arXiv:2403.17914
-
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution 26 Mar 2024 · 0 repositories · arXiv:2403.17927
-
Multilingual Sentence-T5: Scalable Sentence Encoders for Multilingual Applications 26 Mar 2024 · 0 repositories · arXiv:2403.17528
-
Not All Similarities Are Created Equal: Leveraging Data-Driven Biases to Inform GenAI Copyright Disputes 26 Mar 2024 · 0 repositories · arXiv:2403.17691
-
OmniVid: A Generative Framework for Universal Video Understanding 26 Mar 2024 · 1 repository · arXiv:2403.17935Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Targeted Visualization of the Backbone of Encoder LLMs 26 Mar 2024 · 1 repository · arXiv:2403.18872
-
Transcribing Bengali Text with Regional Dialects to IPA using District Guided Tokens 26 Mar 2024 · 0 repositories · arXiv:2403.17407
-
Verbing Weirds Language (Models): Evaluation of English Zero-Derivation in Five LLMs 26 Mar 2024 · 0 repositories · arXiv:2403.17856
-
A comparison of Human, GPT-3.5, and GPT-4 Performance in a University-Level Coding Course 25 Mar 2024 · 1 repository · arXiv:2403.16977
-
A Study on How Attention Scores in the BERT Model are Aware of Lexical Categories in Syntactic and Semantic Tasks on the GLUE Benchmark 25 Mar 2024 · 0 repositories · arXiv:2403.16447
-
Concurrent Linguistic Error Detection (CLED) for Large Language Models 25 Mar 2024 · 0 repositories · arXiv:2403.16393
-
Iterative Refinement of Project-Level Code Context for Precise Code Generation with Compiler Feedback 25 Mar 2024 · 1 repository · arXiv:2403.16792
-
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair 25 Mar 2024 · 1 repository · arXiv:2403.17134Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LinkPrompt: Natural and Universal Adversarial Attacks on Prompt-based Language Models 25 Mar 2024 · 1 repository · arXiv:2403.16432Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data 25 Mar 2024 · 1 repository · arXiv:2403.16909
-
LexDrafter: Terminology Drafting for Legislative Documents using Retrieval Augmented Generation 24 Mar 2024 · 1 repository · arXiv:2403.16295
-
R3CD: Scene Graph to Image Generation with Relation-aware Compositional Contrastive Control Diffusion 24 Mar 2024 · 0 repositories
-
SMTF: Sparse transformer with multiscale contextual fusion for medical image segmentation 24 Mar 2024 · 1 repository
-
SQL-Encoder: Improving NL2SQL In-Context Learning Through a Context-Aware Encoder 24 Mar 2024 · 0 repositories · arXiv:2403.16204
-
CodeShell Technical Report 23 Mar 2024 · 0 repositories · arXiv:2403.15747
-
Contact-aware Human Motion Generation from Textual Descriptions 23 Mar 2024 · 0 repositories · arXiv:2403.15709
-
Fine Tuning LLM for Enterprise: Practical Guidelines and Recommendations 23 Mar 2024 · 0 repositories · arXiv:2404.10779