Methods › General › Regularization › Attention Dropout › Papers, page 37
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 37 of 109: papers 3,601 to 3,700 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
PAL: Proxy-Guided Black-Box Attack on Large Language Models 15 Feb 2024 · 1 repository · arXiv:2402.09674
-
The Butterfly Effect of Model Editing: Few Edits Can Trigger Large Language Models Collapse 15 Feb 2024 · 1 repository · arXiv:2402.09656
-
A Language Model for Particle Tracking 14 Feb 2024 · 0 repositories · arXiv:2402.10239
-
API Pack: A Massive Multi-Programming Language Dataset for API Call Generation 14 Feb 2024 · 1 repository · arXiv:2402.09615
-
Emerging Opportunities of Using Large Language Models for Translation Between Drug Molecules and Indications 14 Feb 2024 · 0 repositories · arXiv:2402.09588
-
FGeo-TP: A Language Model-Enhanced Solver for Geometry Problems 14 Feb 2024 · 0 repositories · arXiv:2402.09047
-
Leveraging Large Language Models for Enhanced NLP Task Performance through Knowledge Distillation and Optimized Training Strategies 14 Feb 2024 · 0 repositories · arXiv:2402.09282
-
MPIrigen: MPI Code Generation through Domain-Specific Language Models 14 Feb 2024 · 1 repository · arXiv:2402.09126
-
Regional inflation analysis using social network data 14 Feb 2024 · 0 repositories · arXiv:2403.00774
-
ScamSpot: Fighting Financial Fraud in Instagram Comments 14 Feb 2024 · 0 repositories · arXiv:2402.08869
-
Using Counterfactual Tasks to Evaluate the Generality of Analogical Reasoning in Large Language Models 14 Feb 2024 · 0 repositories · arXiv:2402.08955
-
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments 13 Feb 2024 · 0 repositories · arXiv:2402.08498
-
BERT4FCA: A Method for Bipartite Link Prediction using Formal Concept Analysis and BERT 13 Feb 2024 · 0 repositories · arXiv:2402.08236
-
COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability 13 Feb 2024 · 1 repository · arXiv:2402.08679Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Eliciting Personality Traits in Large Language Models 13 Feb 2024 · 0 repositories · arXiv:2402.08341
-
Improving Black-box Robustness with In-Context Rewriting 13 Feb 2024 · 1 repository · arXiv:2402.08225
-
Lying Blindly: Bypassing ChatGPT's Safeguards to Generate Hard-to-Detect Disinformation Claims 13 Feb 2024 · 0 repositories · arXiv:2402.08467
-
Measuring and Controlling Instruction (In)Stability in Language Model Dialogs 13 Feb 2024 · 1 repository · arXiv:2402.10962Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Mitigating Object Hallucination in Large Vision-Language Models via Classifier-Free Guidance 13 Feb 2024 · 0 repositories · arXiv:2402.08680
-
PRompt Optimization in Multi-Step Tasks (PROMST): Integrating Human Feedback and Heuristic-based Sampling 13 Feb 2024 · 1 repository · arXiv:2402.08702Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
Addressing cognitive bias in medical language models 12 Feb 2024 · 1 repository · arXiv:2402.08113Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
BreakGPT: A Large Language Model with Multi-stage Structure for Financial Breakout Detection 12 Feb 2024 · 1 repository · arXiv:2402.07536
-
CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge 12 Feb 2024 · 1 repository · arXiv:2402.07688
-
Developing a Multi-variate Prediction Model For COVID-19 From Crowd-sourced Respiratory Voice Data 12 Feb 2024 · 0 repositories · arXiv:2402.07619
-
G-Retriever: Retrieval-Augmented Generation for Textual Graph Understanding and Question Answering 12 Feb 2024 · 2 repositories · arXiv:2402.07630Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Investigating the Impact of Data Contamination of Large Language Models in Text-to-SQL Translation 12 Feb 2024 · 0 repositories · arXiv:2402.08100
-
Leveraging AI to Advance Science and Computing Education across Africa: Challenges, Progress and Opportunities 12 Feb 2024 · 0 repositories · arXiv:2402.07397
-
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models 12 Feb 2024 · 2 repositories · arXiv:2402.07867Syntology official (archive's flag): 10 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 17 harvested samples) · 7 pointer-only (licence)
-
T-RAG: Lessons from the LLM Trenches 12 Feb 2024 · 0 repositories · arXiv:2402.07483
-
HyperBERT: Mixing Hypergraph-Aware Layers with Language Models for Node Classification on Text-Attributed Hypergraphs 11 Feb 2024 · 1 repository · arXiv:2402.07309
-
Prompt Perturbation in Retrieval-Augmented Generation based Large Language Models 11 Feb 2024 · 0 repositories · arXiv:2402.07179
-
Can Graph Descriptive Order Affect Solving Graph Problems with LLMs? 11 Feb 2024 · 0 repositories · arXiv:2402.07140
-
ChemLLM: A Chemical Large Language Model 10 Feb 2024 · 1 repository · arXiv:2402.06852
-
Sentinels of the Stream: Unleashing Large Language Models for Dynamic Packet Classification in Software Defined Networks -- Position Paper 10 Feb 2024 · 0 repositories · arXiv:2402.07950
-
Understanding the Training Speedup from Sampling with Approximate Losses 10 Feb 2024 · 0 repositories · arXiv:2402.07052
-
CultureLLM: Incorporating Cultural Differences into Large Language Models 9 Feb 2024 · 2 repositories · arXiv:2402.10946Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
EntGPT: Linking Generative Large Language Models with Knowledge Bases 9 Feb 2024 · 1 repository · arXiv:2402.06738
-
ExaRanker-Open: Synthetic Explanation for IR using Open-Source LLMs 9 Feb 2024 · 1 repository · arXiv:2402.06334
-
FaBERT: Pre-training BERT on Persian Blogs 9 Feb 2024 · 0 repositories · arXiv:2402.06617
-
G-SciEdBERT: A Contextualized LLM for Science Assessment Tasks in German 9 Feb 2024 · 1 repository · arXiv:2402.06584
-
Learn To be Efficient: Build Structured Sparsity in Large Language Models 9 Feb 2024 · 0 repositories · arXiv:2402.06126
-
JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs 8 Feb 2024 · 2 repositories · arXiv:2402.05668Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Efficient Models for the Detection of Hate, Abuse and Profanity 8 Feb 2024 · 0 repositories · arXiv:2402.05624
-
Efficient Stagewise Pretraining via Progressive Subnetworks 8 Feb 2024 · 0 repositories · arXiv:2402.05913
-
In-Context Principle Learning from Mistakes 8 Feb 2024 · 1 repository · arXiv:2402.05403
-
InkSight: Offline-to-Online Handwriting Conversion by Learning to Read and Write 8 Feb 2024 · 1 repository · arXiv:2402.05804
-
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data 8 Feb 2024 · 0 repositories · arXiv:2402.05545
-
Neural Models for Source Code Synthesis and Completion 8 Feb 2024 · 0 repositories · arXiv:2402.06690
-
Zero-Shot Chain-of-Thought Reasoning Guided by Evolutionary Algorithms in Large Language Models 8 Feb 2024 · 0 repositories · arXiv:2402.05376
-
A Hypothesis-Driven Framework for the Analysis of Self-Rationalising Models 7 Feb 2024 · 1 repository · arXiv:2402.04787
-
Aspect-Based Sentiment Analysis for Open-Ended HR Survey Responses 7 Feb 2024 · 0 repositories · arXiv:2402.04812
-
Amortized Planning with Large-Scale Transformers: A Case Study on Chess 7 Feb 2024 · 1 repository · arXiv:2402.04494Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
How BERT Speaks Shakespearean English? Evaluating Historical Bias in Contextual Language Models 7 Feb 2024 · 0 repositories · arXiv:2402.05034
-
Improving Cross-Domain Low-Resource Text Generation through LLM Post-Editing: A Programmer-Interpreter Approach 7 Feb 2024 · 0 repositories · arXiv:2402.04609
-
Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning 7 Feb 2024 · 1 repository · arXiv:2402.04833Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Advancing Legal Reasoning: The Integration of AI to Navigate Complexities and Biases in Global Jurisprudence with Semi-Automated Arbitration Processes (SAAPs) 6 Feb 2024 · 0 repositories · arXiv:2402.04140
-
Behind the Screen: Investigating ChatGPT's Dark Personality Traits and Conspiracy Beliefs 6 Feb 2024 · 0 repositories · arXiv:2402.04110
-
CEHR-GPT: Generating Electronic Health Records with Chronological Patient Timelines 6 Feb 2024 · 0 repositories · arXiv:2402.04400
-
Detecting Mode Collapse in Language Models via Narration 6 Feb 2024 · 0 repositories · arXiv:2402.04477
-
Enhancing Retrieval Processes for Language Generation with Augmented Queries 6 Feb 2024 · 0 repositories · arXiv:2402.16874
-
Large Language Models as an Indirect Reasoner: Contrapositive and Contradiction for Automated Reasoning 6 Feb 2024 · 0 repositories · arXiv:2402.03667
-
Large Language Models As MOOCs Graders 6 Feb 2024 · 0 repositories · arXiv:2402.03776
-
Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs 6 Feb 2024 · 0 repositories · arXiv:2402.03927
-
LegalLens: Leveraging LLMs for Legal Violation Identification in Unstructured Text 6 Feb 2024 · 1 repository · arXiv:2402.04335
-
Lens: A Foundation Model for Network Traffic 6 Feb 2024 · 0 repositories · arXiv:2402.03646
-
Are Machines Better at Complex Reasoning? Unveiling Human-Machine Inference Gaps in Entailment Verification 6 Feb 2024 · 0 repositories · arXiv:2402.03686
-
Pard: Permutation-Invariant Autoregressive Diffusion for Graph Generation 6 Feb 2024 · 1 repository · arXiv:2402.03687
-
Stanceosaurus 2.0: Classifying Stance Towards Russian and Spanish Misinformation 6 Feb 2024 · 0 repositories · arXiv:2402.03642
-
The Hedgehog & the Porcupine: Expressive Linear Attentions with Softmax Mimicry 6 Feb 2024 · 1 repository · arXiv:2402.04347
-
The Use of a Large Language Model for Cyberbullying Detection 6 Feb 2024 · 0 repositories · arXiv:2402.04088
-
Training Language Models to Generate Text with Citations via Fine-grained Rewards 6 Feb 2024 · 1 repository · arXiv:2402.04315Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Accurate and Well-Calibrated ICD Code Assignment Through Attention Over Diverse Label Embeddings 5 Feb 2024 · 1 repository · arXiv:2402.03172Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Arabic Synonym BERT-based Adversarial Examples for Text Classification 5 Feb 2024 · 1 repository · arXiv:2402.03477
-
C-RAG: Certified Generation Risks for Retrieval-Augmented Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03181Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Reconstruct Your Previous Conversations! Comprehensively Investigating Privacy Leakage Risks in Conversations with GPT Models 5 Feb 2024 · 1 repository · arXiv:2402.02987
-
Enhancing textual textbook question answering with large language models and retrieval augmented generation 5 Feb 2024 · 1 repository · arXiv:2402.05128Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Financial Report Chunking for Effective Retrieval Augmented Generation 5 Feb 2024 · 1 repository · arXiv:2402.05131
-
Harnessing PubMed User Query Logs for Post Hoc Explanations of Recommended Similar Articles 5 Feb 2024 · 0 repositories · arXiv:2402.03484
-
LB-KBQA: Large-language-model and BERT based Knowledge-Based Question and Answering System 5 Feb 2024 · 0 repositories · arXiv:2402.05130
-
LLM Agents in Interaction: Measuring Personality Consistency and Linguistic Alignment in Interacting Populations of Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.02896
-
Multi-Lingual Malaysian Embedding: Leveraging Large Language Models for Semantic Representations 5 Feb 2024 · 0 repositories · arXiv:2402.03053
-
SWAG: Storytelling With Action Guidance 5 Feb 2024 · 1 repository · arXiv:2402.03483
-
UniMem: Towards a Unified View of Long-Context Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03009
-
A Graph is Worth K Words: Euclideanizing Graph using Pure Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02464Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models 4 Feb 2024 · 1 repository · arXiv:2402.02370
-
Breaking MLPerf Training: A Case Study on Optimizing BERT 4 Feb 2024 · 0 repositories · arXiv:2402.02447
-
GeReA: Question-Aware Prompt Captions for Knowledge-based Visual Question Answering 4 Feb 2024 · 1 repository · arXiv:2402.02503Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation 4 Feb 2024 · 0 repositories · arXiv:2402.14594
-
Data Quality Matters: Suicide Intention Detection on Social Media Posts Using RoBERTa-CNN 3 Feb 2024 · 0 repositories · arXiv:2402.02262
-
DE³-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks 3 Feb 2024 · 0 repositories · arXiv:2402.05948
-
EffiBench: Benchmarking the Efficiency of Automatically Generated Code 3 Feb 2024 · 1 repository · arXiv:2402.02037Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Hierarchical Structure Enhances the Convergence and Generalizability of Linear Molecular Representation 3 Feb 2024 · 1 repository · arXiv:2402.02164
-
Clarifying the Path to User Satisfaction: An Investigation into Clarification Usefulness 2 Feb 2024 · 1 repository · arXiv:2402.01934
-
COMET: Generating Commit Messages using Delta Graph Context Representation 2 Feb 2024 · 0 repositories · arXiv:2402.01841
-
Can LLMs perform structured graph reasoning? 2 Feb 2024 · 1 repository · arXiv:2402.01805
-
Improving Sequential Recommendations with LLMs 2 Feb 2024 · 1 repository · arXiv:2402.01339Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
LLM-Detector: Improving AI-Generated Chinese Text Detection with Open-Source LLM Instruction Tuning 2 Feb 2024 · 1 repository · arXiv:2402.01158
-
Predicting ATP binding sites in protein sequences using Deep Learning and Natural Language Processing 2 Feb 2024 · 0 repositories · arXiv:2402.01829
-
Retrieval Augmented End-to-End Spoken Dialog Models 2 Feb 2024 · 0 repositories · arXiv:2402.01828
-
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks 2 Feb 2024 · 0 repositories · arXiv:2402.01176