Methods › General › Regularization › Attention Dropout › Papers, page 24
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 24 of 109: papers 2,301 to 2,400 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Building Trust in Mental Health Chatbots: Safety Metrics and LLM-Based Evaluation Tools 3 Aug 2024 · 0 repositories · arXiv:2408.04650
-
Cross-domain Named Entity Recognition via Graph Matching 2 Aug 2024 · 0 repositories · arXiv:2408.00981
-
LLM as Runtime Error Handler: A Promising Pathway to Adaptive Self-Healing of Software Systems 2 Aug 2024 · 0 repositories · arXiv:2408.01055
-
BioRAG: A RAG-LLM Framework for Biological Question Reasoning 2 Aug 2024 · 0 repositories · arXiv:2408.01107
-
High-Throughput Phenotyping of Clinical Text Using Large Language Models 2 Aug 2024 · 0 repositories · arXiv:2408.01214
-
RAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework 2 Aug 2024 · 1 repository · arXiv:2408.01262
-
Evaluating the Impact of Advanced LLM Techniques on AI-Lecture Tutors for a Robotics Course 2 Aug 2024 · 0 repositories · arXiv:2408.04645
-
Improving Retrieval-Augmented Generation in Medicine with Iterative Follow-up Questions 1 Aug 2024 · 1 repository · arXiv:2408.00727
-
AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation 1 Aug 2024 · 1 repository · arXiv:2408.00764Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Automatic Pull Request Description Generation Using LLMs: A T5 Model Approach 1 Aug 2024 · 0 repositories · arXiv:2408.00921
-
Guiding Sentiment Analysis with Hierarchical Text Clustering: Analyzing the German X/Twitter Discourse on Face Masks in the 2020 COVID-19 Pandemic 1 Aug 2024 · 1 repository
-
Ontological Relations from Word Embeddings 1 Aug 2024 · 0 repositories · arXiv:2408.00444
-
What comes after transformers? -- A selective survey connecting ideas in deep learning 1 Aug 2024 · 0 repositories · arXiv:2408.00386
-
Multi-Level Querying using A Knowledge Pyramid 31 Jul 2024 · 0 repositories · arXiv:2407.21276
-
SAKR: Enhancing Retrieval-Augmented Generation via Streaming Algorithm and K-Means Clustering 31 Jul 2024 · 0 repositories · arXiv:2407.21300
-
MetaOpenFOAM: an LLM-based multi-agent framework for CFD 31 Jul 2024 · 1 repository · arXiv:2407.21320Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Performance of Recent Large Language Models for a Low-Resourced Language 31 Jul 2024 · 0 repositories · arXiv:2407.21330
-
Improving Faithfulness of Large Language Models in Summarization via Sliding Generation and Self-Consistency 31 Jul 2024 · 0 repositories · arXiv:2407.21443
-
Generative Expressive Conversational Speech Synthesis 31 Jul 2024 · 1 repository · arXiv:2407.21491
-
Adaptive Retrieval-Augmented Generation for Conversational Systems 31 Jul 2024 · 0 repositories · arXiv:2407.21712
-
Automated Software Vulnerability Static Code Analysis Using Generative Pre-Trained Transformer Models 31 Jul 2024 · 0 repositories · arXiv:2408.00197
-
Recording First-person Experiences to Build a New Type of Foundation Model 31 Jul 2024 · 0 repositories · arXiv:2408.02680
-
A New Type of Foundation Model Based on Recordings of People's Emotions and Physiology 31 Jul 2024 · 0 repositories · arXiv:2408.00030
-
Event-Arguments Extraction Corpus and Modeling using BERT for Arabic 30 Jul 2024 · 0 repositories · arXiv:2407.21153
-
Decomposed Prompting to Answer Questions on a Course Discussion Board 30 Jul 2024 · 1 repository · arXiv:2407.21170
-
BERT and LLMs-Based avGFP Brightness Prediction and Mutation Design 30 Jul 2024 · 0 repositories · arXiv:2407.20534
-
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification 30 Jul 2024 · 0 repositories · arXiv:2407.20859
-
CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning 30 Jul 2024 · 1 repository · arXiv:2407.21011
-
Comparison of Large Language Models for Generating Contextually Relevant Questions 30 Jul 2024 · 1 repository · arXiv:2407.20578
-
A Study on the Implementation Method of an Agent-Based Advanced RAG System Using Graph 29 Jul 2024 · 0 repositories · arXiv:2407.19994
-
AgEval: A Benchmark for Zero-Shot and Few-Shot Plant Stress Phenotyping with Multimodal LLMs 29 Jul 2024 · 0 repositories · arXiv:2407.19617
-
AutoScale: Scale-Aware Data Mixing for Pre-Training LLMs 29 Jul 2024 · 1 repository · arXiv:2407.20177Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Detecting and Understanding Vulnerabilities in Language Models via Mechanistic Interpretability 29 Jul 2024 · 1 repository · arXiv:2407.19842
-
Enhancing Code Translation in Language Models with Few-Shot Learning via Retrieval-Augmented Generation 29 Jul 2024 · 0 repositories · arXiv:2407.19619
-
Introducing a new hyper-parameter for RAG: Context Window Utilization 29 Jul 2024 · 0 repositories · arXiv:2407.19794
-
Sentiment Analysis of Lithuanian Online Reviews Using Large Language Models 29 Jul 2024 · 0 repositories · arXiv:2407.19914
-
Faculty Perspectives on the Potential of RAG in Computer Science Higher Education 28 Jul 2024 · 0 repositories · arXiv:2408.01462
-
AdaCoder: Adaptive Prompt Compression for Programmatic Visual Question Answering 28 Jul 2024 · 0 repositories · arXiv:2407.19410
-
Are LLMs Good Annotators for Discourse-level Event Relation Extraction? 28 Jul 2024 · 1 repository · arXiv:2407.19568
-
Exploring Genre and Success Classification through Song Lyrics using DistilBERT: A Fun NLP Venture 28 Jul 2024 · 0 repositories · arXiv:2407.21068
-
Is Generative AI an Existential Threat to Human Creatives? Insights from Financial Economics 28 Jul 2024 · 0 repositories · arXiv:2407.19586
-
Motamot: A Dataset for Revealing the Supremacy of Large Language Models over Transformer Models in Bengali Political Sentiment Analysis 28 Jul 2024 · 1 repository · arXiv:2407.19528
-
FarSSiBERT: A Novel Transformer-based Model for Semantic Similarity Measurement of Persian Social Networks Informal Texts 27 Jul 2024 · 0 repositories · arXiv:2407.19173
-
Modular RAG: Transforming RAG Systems into LEGO-like Reconfigurable Frameworks 26 Jul 2024 · 0 repositories · arXiv:2407.21059
-
A Reliable Common-Sense Reasoning Socialbot Built Using LLMs and Goal-Directed ASP 26 Jul 2024 · 0 repositories · arXiv:2407.18498
-
Human-artificial intelligence teaming for scientific information extraction from data-driven additive manufacturing research using large language models 26 Jul 2024 · 0 repositories · arXiv:2407.18827
-
ClinicRealm: Re-evaluating Large Language Models with Conventional Machine Learning for Non-Generative Clinical Prediction Tasks 26 Jul 2024 · 1 repository · arXiv:2407.18525Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
MistralBSM: Leveraging Mistral-7B for Vehicular Networks Misbehavior Detection 26 Jul 2024 · 0 repositories · arXiv:2407.18462
-
Mixed Non-linear Quantization for Vision Transformers 26 Jul 2024 · 1 repository · arXiv:2407.18437
-
REAPER: Reasoning based Retrieval Planning for Complex RAG Systems 26 Jul 2024 · 0 repositories · arXiv:2407.18553
-
TAGIFY: LLM-powered Tagging Interface for Improved Data Findability on OGD portals 26 Jul 2024 · 0 repositories · arXiv:2407.18764
-
Using Large Language Models for the Interpretation of Building Regulations 26 Jul 2024 · 0 repositories · arXiv:2407.21060
-
Banyan: Improved Representation Learning with Explicit Structure 25 Jul 2024 · 0 repositories · arXiv:2407.17771
-
Closing the gap between open-source and commercial large language models for medical evidence summarization 25 Jul 2024 · 0 repositories · arXiv:2408.00588
-
Cost-effective Instruction Learning for Pathology Vision and Language Analysis 25 Jul 2024 · 1 repository · arXiv:2407.17734
-
PEFT-U: Parameter-Efficient Fine-Tuning for User Personalization 25 Jul 2024 · 1 repository · arXiv:2407.18078Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
PersonaGym: Evaluating Persona Agents and LLMs 25 Jul 2024 · 1 repository · arXiv:2407.18416
-
Positive Text Reframing under Multi-strategy Optimization 25 Jul 2024 · 1 repository · arXiv:2407.17940
-
RoBERTa, ResNeXt and BiLSTM with self-attention: The ultimate trio for customer sentiment analysis 25 Jul 2024 · 0 repositories
-
The Geometry of Queries: Query-Based Innovations in Retrieval-Augmented Generation 25 Jul 2024 · 0 repositories · arXiv:2407.18044
-
Understanding the Interplay of Scale, Data, and Bias in Language Models: A Case Study with BERT 25 Jul 2024 · 0 repositories · arXiv:2407.21058
-
What Matters in Explanations: Towards Explainable Fake Review Detection Focusing on Transformers 24 Jul 2024 · 0 repositories · arXiv:2407.21056
-
A Comprehensive Approach to Misspelling Correction with BERT and Levenshtein Distance 24 Jul 2024 · 0 repositories · arXiv:2407.17383
-
A Novel Two-Step Fine-Tuning Pipeline for Cold-Start Active Learning in Text Classification Tasks 24 Jul 2024 · 0 repositories · arXiv:2407.17284
-
Bailicai: A Domain-Optimized Retrieval-Augmented Generation Framework for Medical Applications 24 Jul 2024 · 0 repositories · arXiv:2407.21055
-
Reporting and Analysing the Environmental Impact of Language Models on the Example of Commonsense Question Answering with External Knowledge 24 Jul 2024 · 0 repositories · arXiv:2408.01453
-
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles 24 Jul 2024 · 0 repositories · arXiv:2407.17211
-
Analyzing Polysemy Evolution Using Semantic Cells 23 Jul 2024 · 0 repositories · arXiv:2407.16110
-
Artificial Intelligence in Extracting Diagnostic Data from Dental Records 23 Jul 2024 · 0 repositories · arXiv:2407.21050
-
Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data? 23 Jul 2024 · 1 repository · arXiv:2407.16607Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing LLM's Cognition via Structurization 23 Jul 2024 · 1 repository · arXiv:2407.16434Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
LawLuo: A Multi-Agent Collaborative Framework for Multi-Round Chinese Legal Consultation 23 Jul 2024 · 0 repositories · arXiv:2407.16252
-
Patched RTC: evaluating LLMs for diverse software development tasks 23 Jul 2024 · 1 repository · arXiv:2407.16557
-
Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach 23 Jul 2024 · 0 repositories · arXiv:2407.16833
-
Robust Privacy Amidst Innovation with Large Language Models Through a Critical Assessment of the Risks 23 Jul 2024 · 1 repository · arXiv:2407.16166
-
TookaBERT: A Step Forward for Persian NLU 23 Jul 2024 · 0 repositories · arXiv:2407.16382
-
An Empirical Comparison of Video Frame Sampling Methods for Multi-Modal RAG Retrieval 22 Jul 2024 · 0 repositories · arXiv:2408.03340
-
Customized Retrieval Augmented Generation and Benchmarking for EDA Tool Documentation QA 22 Jul 2024 · 0 repositories · arXiv:2407.15353
-
Impacts of Anthropomorphizing Large Language Models in Learning Environments 22 Jul 2024 · 0 repositories · arXiv:2408.03945
-
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models 22 Jul 2024 · 0 repositories · arXiv:2407.15399
-
Inverted Activations: Reducing Memory Footprint in Neural Network Training 22 Jul 2024 · 1 repository · arXiv:2407.15545
-
LLMmap: Fingerprinting For Large Language Models 22 Jul 2024 · 1 repository · arXiv:2407.15847Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples)
-
MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity 22 Jul 2024 · 1 repository · arXiv:2407.15838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
MoRSE: Bridging the Gap in Cybersecurity Expertise with Retrieval Augmented Generation 22 Jul 2024 · 0 repositories · arXiv:2407.15748
-
Promises and Pitfalls of Generative Masked Language Modeling: Theoretical Framework and Practical Guidelines 22 Jul 2024 · 1 repository · arXiv:2407.21046
-
RadioRAG: Factual large language models for enhanced diagnostics in radiology using online retrieval augmented generation 22 Jul 2024 · 1 repository · arXiv:2407.15621
-
Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget 22 Jul 2024 · 1 repository · arXiv:2407.15811Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research 22 Jul 2024 · 0 repositories · arXiv:2407.21045
-
ZZU-NLP at SIGHAN-2024 dimABSA Task: Aspect-Based Sentiment Analysis with Coarse-to-Fine In-context Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15341
-
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts 21 Jul 2024 · 0 repositories · arXiv:2407.15136
-
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment 21 Jul 2024 · 1 repository · arXiv:2407.15184
-
Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval 21 Jul 2024 · 1 repository · arXiv:2407.15051Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
Golden-Retriever: High-Fidelity Agentic Retrieval Augmented Generation for Industrial Knowledge Base 20 Jul 2024 · 0 repositories · arXiv:2408.00798
-
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models 20 Jul 2024 · 1 repository · arXiv:2407.14944
-
Differential Privacy of Cross-Attention with Provable Guarantee 20 Jul 2024 · 0 repositories · arXiv:2407.14717
-
Adversarial Databases Improve Success in Retrieval-based Large Language Models 19 Jul 2024 · 0 repositories · arXiv:2407.14609
-
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities 19 Jul 2024 · 0 repositories · arXiv:2407.14482
-
Unipa-GPT: Large Language Models for university-oriented QA in Italian 19 Jul 2024 · 1 repository · arXiv:2407.14246
-
SQLfuse: Enhancing Text-to-SQL Performance through Comprehensive LLM Synergy 19 Jul 2024 · 0 repositories · arXiv:2407.14568
-
Scalable Exploration via Ensemble++ 18 Jul 2024 · 2 repositories · arXiv:2407.13195Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)