Methods › General › Regularization › Attention Dropout › Papers, page 25
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 25 of 109: papers 2,401 to 2,500 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models 18 Jul 2024 · 0 repositories · arXiv:2407.13757
-
Can Open-Source LLMs Compete with Commercial Models? Exploring the Few-Shot Performance of Current GPT Models in Biomedical Tasks 18 Jul 2024 · 1 repository · arXiv:2407.13511
-
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts 18 Jul 2024 · 1 repository · arXiv:2407.13228
-
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency 18 Jul 2024 · 0 repositories · arXiv:2407.13578
-
Learning-From-Mistakes Prompting for Indigenous Language Translation 18 Jul 2024 · 0 repositories · arXiv:2407.13343
-
PRAGyan -- Connecting the Dots in Tweets 18 Jul 2024 · 0 repositories · arXiv:2407.13909
-
Qalam : A Multimodal LLM for Arabic Optical Character and Handwriting Recognition 18 Jul 2024 · 0 repositories · arXiv:2407.13559
-
Reconstruct the Pruned Model without Any Retraining 18 Jul 2024 · 0 repositories · arXiv:2407.13331
-
Retrieval-Augmented Generation for Natural Language Processing: A Survey 18 Jul 2024 · 0 repositories · arXiv:2407.13193
-
Retrieve, Summarize, Plan: Advancing Multi-hop Question Answering with an Iterative Approach 18 Jul 2024 · 0 repositories · arXiv:2407.13101
-
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction 18 Jul 2024 · 1 repository · arXiv:2407.13943Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases 17 Jul 2024 · 1 repository · arXiv:2407.12784Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 1 honoured, 0 violated, 14 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 1 pointer-only (licence)
-
Deep Learning-based Sentiment Analysis of Olympics Tweets 17 Jul 2024 · 0 repositories · arXiv:2407.12376
-
On Initializing Transformers with Pre-trained Embeddings 17 Jul 2024 · 0 repositories · arXiv:2407.12514
-
Optimizing Query Generation for Enhanced Document Retrieval in RAG 17 Jul 2024 · 0 repositories · arXiv:2407.12325
-
Evaluating Search Engines and Large Language Models for Answering Health Questions 17 Jul 2024 · 1 repository · arXiv:2407.12468
-
Sharif-STR at SemEval-2024 Task 1: Transformer as a Regression Model for Fine-Grained Scoring of Textual Semantic Relations 17 Jul 2024 · 1 repository · arXiv:2407.12426
-
Textualized and Feature-based Models for Compound Multimodal Emotion Recognition in the Wild 17 Jul 2024 · 2 repositories · arXiv:2407.12927Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Better RAG using Relevant Information Gain 16 Jul 2024 · 1 repository · arXiv:2407.12101
-
Beyond Binary: Multiclass Paraphasia Detection with Generative Pretrained Transformers and End-to-End Models 16 Jul 2024 · 0 repositories · arXiv:2407.11345
-
ChatBCG: Can AI Read Your Slide Deck? 16 Jul 2024 · 0 repositories · arXiv:2407.12875
-
Does Refusal Training in LLMs Generalize to the Past Tense? 16 Jul 2024 · 1 repository · arXiv:2407.11969Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text 16 Jul 2024 · 0 repositories · arXiv:2407.11827
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Large Language Models as Misleading Assistants in Conversation 16 Jul 2024 · 0 repositories · arXiv:2407.11789
-
Large Visual-Language Models Are Also Good Classifiers: A Study of In-Context Multimodal Fake News Detection 16 Jul 2024 · 0 repositories · arXiv:2407.12879
-
LLMs-in-the-loop Part-1: Expert Small AI Models for Bio-Medical Text Translation 16 Jul 2024 · 0 repositories · arXiv:2407.12126
-
Mindful-RAG: A Study of Points of Failure in Retrieval Augmented Generation 16 Jul 2024 · 0 repositories · arXiv:2407.12216
-
R-SFLLM: Jamming Resilient Framework for Split Federated Learning with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11654
-
Representation Bias in Political Sample Simulations with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11409
-
ReFeR: Improving Evaluation and Reasoning through Hierarchy of Models 16 Jul 2024 · 0 repositories · arXiv:2407.12877
-
Scientific QA System with Verifiable Answers 16 Jul 2024 · 1 repository · arXiv:2407.11485
-
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild 16 Jul 2024 · 1 repository · arXiv:2407.11438
-
Communication- and Computation-Efficient Distributed Submodular Optimization in Robot Mesh Networks 15 Jul 2024 · 1 repository · arXiv:2407.10382
-
Deep Learning-Based Operators for Evolutionary Algorithms 15 Jul 2024 · 0 repositories · arXiv:2407.10477
-
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization 15 Jul 2024 · 0 repositories · arXiv:2407.10424
-
Enhancing Retrieval and Managing Retrieval: A Four-Module Synergy for Improved Quality and Efficiency in RAG Systems 15 Jul 2024 · 1 repository · arXiv:2407.10670Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Evaluation of RAG Metrics for Question Answering in the Telecom Domain 15 Jul 2024 · 0 repositories · arXiv:2407.12873
-
Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis 15 Jul 2024 · 0 repositories · arXiv:2407.10899
-
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game 15 Jul 2024 · 0 repositories · arXiv:2407.11240
-
Mechanistic interpretability of large language models with applications to the financial services industry 15 Jul 2024 · 0 repositories · arXiv:2407.11215
-
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs 15 Jul 2024 · 1 repository · arXiv:2407.10834Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
MixGR: Enhancing Retriever Generalization for Scientific Domain through Complementary Granularity 15 Jul 2024 · 1 repository · arXiv:2407.10691
-
Think-on-Graph 2.0: Deep and Faithful Large Language Model Reasoning with Knowledge-guided Retrieval Augmented Generation 15 Jul 2024 · 1 repository · arXiv:2407.10805Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Weighted Grouped Query Attention in Transformers 15 Jul 2024 · 0 repositories · arXiv:2407.10855
-
YouTube-SL-25: A Large-Scale, Open-Domain Multilingual Sign Language Parallel Corpus 15 Jul 2024 · 1 repository · arXiv:2407.11144
-
Curriculum Learning for Small Code Language Models 14 Jul 2024 · 0 repositories · arXiv:2407.10194
-
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation 14 Jul 2024 · 0 repositories · arXiv:2407.10091
-
Causality extraction from medical text using Large Language Models (LLMs) 13 Jul 2024 · 0 repositories · arXiv:2407.10020
-
Document-level Clinical Entity and Relation Extraction via Knowledge Base-Guided Generation 13 Jul 2024 · 0 repositories · arXiv:2407.10021
-
Generating In-store Customer Journeys from Scratch with GPT Architectures 13 Jul 2024 · 0 repositories · arXiv:2407.11081
-
Hydra: Bidirectional State Space Models Through Generalized Matrix Mixers 13 Jul 2024 · 1 repository · arXiv:2407.09941Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Resource Management for Low-latency Cooperative Fine-tuning of Foundation Models at the Network Edge 13 Jul 2024 · 0 repositories · arXiv:2407.09873
-
ASTPrompter: Weakly Supervised Automated Language Model Red-Teaming to Identify Low-Perplexity Toxic Prompts 12 Jul 2024 · 1 repository · arXiv:2407.09447
-
Deep Bag-of-Words Model: An Efficient and Interpretable Relevance Architecture for Chinese E-Commerce 12 Jul 2024 · 0 repositories · arXiv:2407.09395
-
Enhancing Depressive Post Detection in Bangla: A Comparative Study of TF-IDF, BERT and FastText Embeddings 12 Jul 2024 · 0 repositories · arXiv:2407.09187
-
Human-like Episodic Memory for Infinite Context LLMs 12 Jul 2024 · 1 repository · arXiv:2407.09450Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Movie Recommendation with Poster Attention via Multi-modal Transformer Feature Fusion 12 Jul 2024 · 0 repositories · arXiv:2407.09157
-
Robustness of LLMs to Perturbations in Text 12 Jul 2024 · 0 repositories · arXiv:2407.08989
-
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner 12 Jul 2024 · 0 repositories · arXiv:2407.08937
-
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay 12 Jul 2024 · 1 repository · arXiv:2407.11068
-
Surgical Text-to-Image Generation 12 Jul 2024 · 0 repositories · arXiv:2407.09230
-
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs 12 Jul 2024 · 0 repositories · arXiv:2407.09152
-
Toward Automatic Group Membership Annotation for Group Fairness Evaluation 12 Jul 2024 · 0 repositories · arXiv:2407.08926
-
Beyond Benchmarks: Evaluating Embedding Model Similarity for Retrieval Augmented Generation Systems 11 Jul 2024 · 1 repository · arXiv:2407.08275
-
fairBERTs: Erasing Sensitive Information Through Semantic and Fairness-aware Perturbations 11 Jul 2024 · 0 repositories · arXiv:2407.08189
-
GPT-4 is judged more human than humans in displaced and inverted Turing tests 11 Jul 2024 · 0 repositories · arXiv:2407.08853
-
HDT: Hierarchical Document Transformer 11 Jul 2024 · 0 repositories · arXiv:2407.08330
-
Investigating LLMs as Voting Assistants via Contextual Augmentation: A Case Study on the European Parliament Elections 2024 11 Jul 2024 · 0 repositories · arXiv:2407.08495
-
Knowledge distillation to effectively attain both region-of-interest and global semantics from an image where multiple objects appear 11 Jul 2024 · 1 repository · arXiv:2407.08257
-
LLMs' morphological analyses of complex FST-generated Finnish words 11 Jul 2024 · 1 repository · arXiv:2407.08269
-
MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine 11 Jul 2024 · 3 repositories · arXiv:2407.08739
-
On the (In)Security of LLM App Stores 11 Jul 2024 · 0 repositories · arXiv:2407.08422
-
Real-Time Anomaly Detection and Reactive Planning with Large Language Models 11 Jul 2024 · 0 repositories · arXiv:2407.08735
-
Speculative RAG: Enhancing Retrieval Augmented Generation through Drafting 11 Jul 2024 · 0 repositories · arXiv:2407.08223
-
Vox Populi, Vox AI? Using Language Models to Estimate German Public Opinion 11 Jul 2024 · 1 repository · arXiv:2407.08563
-
Attribute or Abstain: Large Language Models as Long Document Assistants 10 Jul 2024 · 1 repository · arXiv:2407.07799Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
DS@GT eRisk 2024: Sentence Transformers for Social Media Risk Assessment 10 Jul 2024 · 1 repository · arXiv:2407.08008
-
FACTS About Building Retrieval Augmented Generation-based Chatbots 10 Jul 2024 · 0 repositories · arXiv:2407.07858
-
FsPONER: Few-shot Prompt Optimization for Named Entity Recognition in Domain-specific Scenarios 10 Jul 2024 · 1 repository · arXiv:2407.08035
-
KpopMT: Translation Dataset with Terminology for Kpop Fandom 10 Jul 2024 · 1 repository · arXiv:2407.07413
-
A Guide To Effectively Leveraging LLMs for Low-Resource Text Summarization: Data Augmentation and Semi-supervised Approaches 10 Jul 2024 · 0 repositories · arXiv:2407.07341
-
Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture 10 Jul 2024 · 0 repositories · arXiv:2407.07342
-
Examining Long-Context Large Language Models for Environmental Review Document Comprehension 10 Jul 2024 · 0 repositories · arXiv:2407.07321
-
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning 10 Jul 2024 · 1 repository · arXiv:2407.07802
-
A Simple Architecture for Enterprise Large Language Model Applications based on Role based security and Clearance Levels using Retrieval-Augmented Generation or Mixture of Experts 9 Jul 2024 · 0 repositories · arXiv:2407.06718
-
AI AI Bias: Large Language Models Favor Their Own Generated Content 9 Jul 2024 · 1 repository · arXiv:2407.12856
-
ChatGPT Doesn't Trust Chargers Fans: Guardrail Sensitivity in Context 9 Jul 2024 · 1 repository · arXiv:2407.06866Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Empirical analysis of Binding Precedent efficiency in the Brazilian Supreme Court via Similar Case Retrieval 9 Jul 2024 · 0 repositories · arXiv:2407.07004
-
Identification of emotions on Twitter during the 2022 electoral process in Colombia 9 Jul 2024 · 0 repositories · arXiv:2407.07258
-
Measuring Sustainability Intention of ESG Fund Disclosure using Few-Shot Learning 9 Jul 2024 · 0 repositories · arXiv:2407.06893
-
Prompting Techniques for Secure Code Generation: A Systematic Investigation 9 Jul 2024 · 0 repositories · arXiv:2407.07064
-
Raply: A profanity-mitigated rap generator 9 Jul 2024 · 0 repositories · arXiv:2407.06941
-
Segment-Based Interactive Machine Translation for Pre-trained Models 9 Jul 2024 · 0 repositories · arXiv:2407.06990
-
Solving General Natural-Language-Description Optimization Problems with Large Language Models 9 Jul 2024 · 0 repositories · arXiv:2407.07924
-
Using Large Language Models for Generating Smart Contracts for Health Insurance from Textual Policies 9 Jul 2024 · 0 repositories · arXiv:2407.07019
-
Using Pretrained Large Language Model with Prompt Engineering to Answer Biomedical Questions 9 Jul 2024 · 0 repositories · arXiv:2407.06779
-
An Empirical Comparison of Vocabulary Expansion and Initialization Approaches for Language Models 8 Jul 2024 · 1 repository · arXiv:2407.05841Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct 8 Jul 2024 · 1 repository · arXiv:2407.05700Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Large Language Models Understand Layout 8 Jul 2024 · 1 repository · arXiv:2407.05750