Methods › General › Regularization › Attention Dropout › Papers, page 32
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 32 of 109: papers 3,101 to 3,200 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Models 23 Apr 2024 · 0 repositories · arXiv:2404.15549
-
From Complexity to Clarity: How AI Enhances Perceptions of Scientists and the Public's Understanding of Science 23 Apr 2024 · 0 repositories · arXiv:2405.00706
-
Watch Out for Your Guidance on Generation! Exploring Conditional Backdoor Attacks against Large Language Models 23 Apr 2024 · 0 repositories · arXiv:2404.14795
-
Unsupervised End-to-End Task-Oriented Dialogue with LLMs: The Power of the Noisy Channel 23 Apr 2024 · 1 repository · arXiv:2404.15219
-
Automated Long Answer Grading with RiceChem Dataset 22 Apr 2024 · 1 repository · arXiv:2404.14316
-
Pre-Calc: Learning to Use the Calculator Improves Numeracy in Language Models 22 Apr 2024 · 1 repository · arXiv:2404.14355Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Generating Attractive and Authentic Copywriting from Customer Reviews 22 Apr 2024 · 0 repositories · arXiv:2404.13906
-
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO 22 Apr 2024 · 1 repository · arXiv:2404.13957Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Information Re-Organization Improves Reasoning in Large Language Models 22 Apr 2024 · 0 repositories · arXiv:2404.13985
-
LLMs Know What They Need: Leveraging a Missing Information Guided Framework to Empower Retrieval-Augmented Generation 22 Apr 2024 · 1 repository · arXiv:2404.14043
-
Marking: Visual Grading with Highlighting Errors and Annotating Missing Bits 22 Apr 2024 · 0 repositories · arXiv:2404.14301
-
Navigating the Path of Writing: Outline-guided Text Generation with Large Language Models 22 Apr 2024 · 0 repositories · arXiv:2404.13919
-
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone 22 Apr 2024 · 0 repositories · arXiv:2404.14219
-
Typos that Broke the RAG's Back: Genetic Attack on RAG Pipeline by Simulating Documents in the Wild via Low-level Perturbations 22 Apr 2024 · 1 repository · arXiv:2404.13948Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
What do Transformers Know about Government? 22 Apr 2024 · 1 repository · arXiv:2404.14270
-
Zero-shot Cross-lingual Stance Detection via Adversarial Language Adaptation 22 Apr 2024 · 1 repository · arXiv:2404.14339
-
Automated Text Mining of Experimental Methodologies from Biomedical Literature 21 Apr 2024 · 0 repositories · arXiv:2404.13779
-
Evaluating Retrieval Quality in Retrieval-Augmented Generation 21 Apr 2024 · 1 repository · arXiv:2404.13781Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
SVGEditBench: A Benchmark Dataset for Quantitative Assessment of LLM's SVG Editing Capabilities 21 Apr 2024 · 1 repository · arXiv:2404.13710Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
BERT: Accelerating Vital Signs Measurement for Bioradar with An Efficient Recursive Technique 20 Apr 2024 · 0 repositories · arXiv:2404.13315
-
Do "English" Named Entity Recognizers Work Well on Global Englishes? 20 Apr 2024 · 2 repositories · arXiv:2404.13465
-
Evaluating Subword Tokenization: Alien Subword Composition and OOV Generalization Challenge 20 Apr 2024 · 1 repository · arXiv:2404.13292
-
Retrieval-Augmented Generation-based Relation Extraction 20 Apr 2024 · 1 repository · arXiv:2404.13397
-
Data Alignment for Zero-Shot Concept Generation in Dermatology AI 19 Apr 2024 · 0 repositories · arXiv:2404.13043
-
Dubo-SQL: Diverse Retrieval-Augmented Generation and Fine Tuning for Text-to-SQL 19 Apr 2024 · 1 repository · arXiv:2404.12560
-
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning 19 Apr 2024 · 0 repositories · arXiv:2404.12897
-
Multi Class Depression Detection Through Tweets using Artificial Intelligence 19 Apr 2024 · 1 repository · arXiv:2404.13104
-
The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions 19 Apr 2024 · 1 repository · arXiv:2404.13208
-
Unlocking Multi-View Insights in Knowledge-Dense Retrieval-Augmented Generation 19 Apr 2024 · 0 repositories · arXiv:2404.12879
-
Augmenting emotion features in irony detection with Large language modeling 18 Apr 2024 · 0 repositories · arXiv:2404.12291
-
emrQA-msquad: A Medical Dataset Structured with the SQuAD V2.0 Framework, Enriched with emrQA Medical Information 18 Apr 2024 · 0 repositories · arXiv:2404.12050
-
From Form(s) to Meaning: Probing the Semantic Depths of Language Models Using Multisense Consistency 18 Apr 2024 · 1 repository · arXiv:2404.12145Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
iRAG: Advancing RAG for Videos with an Incremental Approach 18 Apr 2024 · 0 repositories · arXiv:2404.12309
-
LongEmbed: Extending Embedding Models for Long Context Retrieval 18 Apr 2024 · 1 repository · arXiv:2404.12096Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
MIDGET: Music Conditioned 3D Dance Generation 18 Apr 2024 · 0 repositories · arXiv:2404.12062
-
RAGAR, Your Falsehood Radar: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language Models 18 Apr 2024 · 0 repositories · arXiv:2404.12065
-
RAGCache: Efficient Knowledge Caching for Retrieval-Augmented Generation 18 Apr 2024 · 0 repositories · arXiv:2404.12457
-
RAM: Towards an Ever-Improving Memory System by Learning from Communications 18 Apr 2024 · 0 repositories · arXiv:2404.12045
-
Stance Detection on Social Media with Fine-Tuned Large Language Models 18 Apr 2024 · 0 repositories · arXiv:2404.12171
-
A Survey on Retrieval-Augmented Text Generation for Large Language Models 17 Apr 2024 · 0 repositories · arXiv:2404.10981
-
Comparative Analysis of Deep Natural Networks and Large Language Models for Aspect-Based Sentiment Analysis 17 Apr 2024 · 1 repository
-
Demystifying Legalese: An Automated Approach for Summarizing and Analyzing Overlaps in Privacy Policies and Terms of Service 17 Apr 2024 · 0 repositories · arXiv:2404.13087
-
Enhancing Q&A with Domain-Specific Fine-Tuning and Iterative Reasoning: A Comparative Study 17 Apr 2024 · 0 repositories · arXiv:2404.11792
-
Improvement in Semantic Address Matching using Natural Language Processing 17 Apr 2024 · 0 repositories · arXiv:2404.11691
-
A Sentiment Analysis of Medical Text Based on Deep Learning 16 Apr 2024 · 0 repositories · arXiv:2404.10503
-
BayesJudge: Bayesian Kernel Language Modelling with Confidence Uncertainty in Legal Judgment Prediction 16 Apr 2024 · 0 repositories · arXiv:2404.10481
-
Empowering Interdisciplinary Research with BERT-Based Models: An Approach Through SciBERT-CNN with Topic Modeling 16 Apr 2024 · 0 repositories · arXiv:2404.13078
-
LaDiC: Are Diffusion Models Really Inferior to Autoregressive Counterparts for Image-to-Text Generation? 16 Apr 2024 · 1 repository · arXiv:2404.10763
-
Relational Graph Convolutional Networks for Sentiment Analysis 16 Apr 2024 · 0 repositories · arXiv:2404.13079
-
Spiral of Silence: How is Large Language Model Killing Information Retrieval? -- A Case Study on Open Domain Question Answering 16 Apr 2024 · 1 repository · arXiv:2404.10496
-
AesExpert: Towards Multi-modality Foundation Model for Image Aesthetics Perception 15 Apr 2024 · 1 repository · arXiv:2404.09624
-
Reinforcement Learning from Multi-role Debates as Feedback for Bias Mitigation in LLMs 15 Apr 2024 · 0 repositories · arXiv:2404.10160
-
Detecting AI Generated Text Based on NLP and Machine Learning Approaches 15 Apr 2024 · 0 repositories · arXiv:2404.10032
-
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks 15 Apr 2024 · 0 repositories · arXiv:2404.10155
-
σ-GPTs: A New Approach to Autoregressive Models 15 Apr 2024 · 1 repository · arXiv:2404.09562
-
BERT-LSH: Reducing Absolute Compute For Attention 12 Apr 2024 · 1 repository · arXiv:2404.08836
-
CreativEval: Evaluating Creativity of LLM-Based Hardware Code Generation 12 Apr 2024 · 0 repositories · arXiv:2404.08806
-
FastLogAD: Log Anomaly Detection with Mask-Guided Pseudo Anomaly Generation and Discrimination 12 Apr 2024 · 1 repository · arXiv:2404.08750
-
Is ChatGPT Transforming Academics' Writing Style? 12 Apr 2024 · 0 repositories · arXiv:2404.08627
-
Inheritune: Training Smaller Yet More Attentive Language Models 12 Apr 2024 · 1 repository · arXiv:2404.08634Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Reducing hallucination in structured outputs via Retrieval-Augmented Generation 12 Apr 2024 · 0 repositories · arXiv:2404.08189
-
Revisiting Code Similarity Evaluation with Abstract Syntax Tree Edit Distance 12 Apr 2024 · 0 repositories · arXiv:2404.08817
-
Small Models Are (Still) Effective Cross-Domain Argument Extractors 12 Apr 2024 · 1 repository · arXiv:2404.08579
-
AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs 11 Apr 2024 · 1 repository · arXiv:2404.07921
-
Generative Information Retrieval Evaluation 11 Apr 2024 · 0 repositories · arXiv:2404.08137
-
HLTCOE at TREC 2023 NeuCLIR Track 11 Apr 2024 · 0 repositories · arXiv:2404.08118
-
Medical mT5: An Open-Source Multilingual Text-to-Text LLM for The Medical Domain 11 Apr 2024 · 0 repositories · arXiv:2404.07613
-
On Training Data Influence of GPT Models 11 Apr 2024 · 2 repositories · arXiv:2404.07840Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Rumour Evaluation with Very Large Language Models 11 Apr 2024 · 1 repository · arXiv:2404.16859
-
Control-DAG: Constrained Decoding for Non-Autoregressive Directed Acyclic T5 using Weighted Finite State Automata 10 Apr 2024 · 1 repository · arXiv:2404.06854
-
Emotion-cause pair extraction method based on multi-granularity information and multi-module interaction 10 Apr 2024 · 0 repositories · arXiv:2404.06812
-
Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness 10 Apr 2024 · 1 repository · arXiv:2404.06714Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation 10 Apr 2024 · 1 repository · arXiv:2404.06809
-
Simpler becomes Harder: Do LLMs Exhibit a Coherent Behavior on Simplified Corpora? 10 Apr 2024 · 1 repository · arXiv:2404.06838
-
Superposition Prompting: Improving and Accelerating Retrieval-Augmented Generation 10 Apr 2024 · 1 repository · arXiv:2404.06910
-
Heuristic-enhanced Candidates Selection strategy for GPTs tackle Few-Shot Aspect-Based Sentiment Analysis 9 Apr 2024 · 1 repository · arXiv:2404.06063
-
Characterizing Multimodal Long-form Summarization: A Case Study on Financial Reports 9 Apr 2024 · 0 repositories · arXiv:2404.06162
-
Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection 9 Apr 2024 · 1 repository · arXiv:2404.06194Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Generative Pre-Trained Transformer for Symbolic Regression Base In-Context Reinforcement Learning 9 Apr 2024 · 0 repositories · arXiv:2404.06330
-
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs 9 Apr 2024 · 0 repositories · arXiv:2404.07242
-
Comprehensive Study on German Language Models for Clinical and Biomedical Text Understanding 8 Apr 2024 · 0 repositories · arXiv:2404.05694
-
Guiding Large Language Models to Generate Computer-Parsable Content 8 Apr 2024 · 0 repositories · arXiv:2404.05499
-
Evaluating Interventional Reasoning Capabilities of Large Language Models 8 Apr 2024 · 0 repositories · arXiv:2404.05545
-
LTNER: Large Language Model Tagging for Named Entity Recognition with Contextualized Entity Marking 8 Apr 2024 · 0 repositories · arXiv:2404.05624
-
MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering 8 Apr 2024 · 0 repositories · arXiv:2404.05590
-
PetKaz at SemEval-2024 Task 3: Advancing Emotion Classification with an LLM for Emotion-Cause Pair Extraction in Conversations 8 Apr 2024 · 1 repository · arXiv:2404.05502
-
Physics of Language Models: Part 3.3, Knowledge Capacity Scaling Laws 8 Apr 2024 · 0 repositories · arXiv:2404.05405
-
Relation Extraction Using Large Language Models: A Case Study on Acupuncture Point Locations 8 Apr 2024 · 0 repositories · arXiv:2404.05415
-
Semantic Stealth: Adversarial Text Attacks on NLP Using Several Methods 8 Apr 2024 · 0 repositories · arXiv:2404.05159
-
A Multi-Level Framework for Accelerating Training Transformer Models 7 Apr 2024 · 1 repository · arXiv:2404.07999Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Data Bias According to Bipol: Men are Naturally Right and It is the Role of Women to Follow Their Lead 7 Apr 2024 · 1 repository · arXiv:2404.04838
-
A Morphology-Based Investigation of Positional Encodings 6 Apr 2024 · 0 repositories · arXiv:2404.04530
-
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials 6 Apr 2024 · 1 repository · arXiv:2404.04510
-
RecGPT: Generative Personalized Prompts for Sequential Recommendation via ChatGPT Training Paradigm 6 Apr 2024 · 0 repositories · arXiv:2404.08675
-
Deciphering Political Entity Sentiment in News with Large Language Models: Zero-Shot and Few-Shot Strategies 5 Apr 2024 · 1 repository · arXiv:2404.04361
-
Investigating the Robustness of Modelling Decisions for Few-Shot Cross-Topic Stance Detection: A Preregistered Study 5 Apr 2024 · 1 repository · arXiv:2404.03987
-
Scope Ambiguities in Large Language Models 5 Apr 2024 · 1 repository · arXiv:2404.04332
-
BanglaAutoKG: Automatic Bangla Knowledge Graph Construction with Semantic Neural Graph Filtering 4 Apr 2024 · 1 repository · arXiv:2404.03528Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CBR-RAG: Case-Based Reasoning for Retrieval Augmented Generation in LLMs for Legal Question Answering 4 Apr 2024 · 1 repository · arXiv:2404.04302
-
CONFLARE: CONFormal LArge language model REtrieval 4 Apr 2024 · 1 repository · arXiv:2404.04287