Methods › General › Regularization › Attention Dropout › Papers, page 23
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 23 of 109: papers 2,201 to 2,300 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Acquiring Bidirectionality via Large and Small Language Models 19 Aug 2024 · 1 repository · arXiv:2408.09640
-
Active Learning for Identifying Disaster-Related Tweets: A Comparison with Keyword Filtering and Generic Fine-Tuning 19 Aug 2024 · 0 repositories · arXiv:2408.09914
-
Carbon Footprint Accounting Driven by Large Language Models and Retrieval-augmented Generation 19 Aug 2024 · 0 repositories · arXiv:2408.09713
-
ELDER: Enhancing Lifelong Model Editing with Mixture-of-LoRA 19 Aug 2024 · 1 repository · arXiv:2408.11869
-
Enhanced document retrieval with topic embeddings 19 Aug 2024 · 0 repositories · arXiv:2408.10435
-
Factorized-Dreamer: Training A High-Quality Video Generator with Limited and Low-Quality Data 19 Aug 2024 · 0 repositories · arXiv:2408.10119
-
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching 19 Aug 2024 · 0 repositories · arXiv:2408.10286
-
LegalBench-RAG: A Benchmark for Retrieval-Augmented Generation in the Legal Domain 19 Aug 2024 · 1 repository · arXiv:2408.10343Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Rhyme-aware Chinese lyric generator based on GPT 19 Aug 2024 · 0 repositories · arXiv:2408.10130
-
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training 19 Aug 2024 · 0 repositories · arXiv:2408.10013
-
Agentic Retrieval-Augmented Generation for Time Series Analysis 18 Aug 2024 · 0 repositories · arXiv:2408.14484
-
ConVerSum: A Contrastive Learning-based Approach for Data-Scarce Solution of Cross-Lingual Summarization Beyond Direct Equivalents 17 Aug 2024 · 0 repositories · arXiv:2408.09273
-
Sentiment analysis of preservice teachers' reflections using a large language model 17 Aug 2024 · 0 repositories · arXiv:2408.11862
-
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering 17 Aug 2024 · 0 repositories · arXiv:2408.09174
-
TC-RAG:Turing-Complete RAG's Case study on Medical LLM Systems 17 Aug 2024 · 2 repositories · arXiv:2408.09199
-
A Mean Field Ansatz for Zero-Shot Weight Transfer 16 Aug 2024 · 0 repositories · arXiv:2408.08681
-
CIKMar: A Dual-Encoder Approach to Prompt-Based Reranking in Educational Dialogue Systems 16 Aug 2024 · 0 repositories · arXiv:2408.08805
-
CommunityKG-RAG: Leveraging Community Structures in Knowledge Graphs for Advanced Retrieval-Augmented Generation in Fact-Checking 16 Aug 2024 · 1 repository · arXiv:2408.08535
-
Fine-tuning LLMs for Autonomous Spacecraft Control: A Case Study Using Kerbal Space Program 16 Aug 2024 · 1 repository · arXiv:2408.08676
-
Improving VTE Identification through Language Models from Radiology Reports: A Comparative Study of Mamba, Phi-3 Mini, and BERT 16 Aug 2024 · 0 repositories · arXiv:2408.09043
-
Meta Knowledge for Retrieval Augmented Large Language Models 16 Aug 2024 · 0 repositories · arXiv:2408.09017
-
Quantifying the Effectiveness of Student Organization Activities using Natural Language Processing 16 Aug 2024 · 0 repositories · arXiv:2408.08694
-
The Fellowship of the LLMs: Multi-Agent Workflows for Synthetic Preference Optimization Dataset Generation 16 Aug 2024 · 1 repository · arXiv:2408.08688
-
VERA: Validation and Evaluation of Retrieval-Augmented Systems 16 Aug 2024 · 0 repositories · arXiv:2409.03759
-
FuseChat: Knowledge Fusion of Chat Models 15 Aug 2024 · 3 repositories · arXiv:2408.07990
-
Graph Retrieval-Augmented Generation: A Survey 15 Aug 2024 · 1 repository · arXiv:2408.08921
-
Leveraging Web-Crawled Data for High-Quality Fine-Tuning 15 Aug 2024 · 1 repository · arXiv:2408.08003
-
Plan with Code: Comparing approaches for robust NL to DSL generation 15 Aug 2024 · 0 repositories · arXiv:2408.08335
-
PQV-Mobile: A Combined Pruning and Quantization Toolkit to Optimize Vision Transformers for Mobile Applications 15 Aug 2024 · 1 repository · arXiv:2408.08437
-
Predicting Lung Cancer Patient Prognosis with Large Language Models 15 Aug 2024 · 0 repositories · arXiv:2408.07971
-
RAGChecker: A Fine-grained Framework for Diagnosing Retrieval-Augmented Generation 15 Aug 2024 · 1 repository · arXiv:2408.08067Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
CodeMirage: Hallucinations in Code Generated by Large Language Models 14 Aug 2024 · 0 repositories · arXiv:2408.08333
-
DataVisT5: A Pre-trained Language Model for Jointly Understanding Text and Data Visualization 14 Aug 2024 · 1 repository · arXiv:2408.07401
-
Enhancing Visual Question Answering through Ranking-Based Hybrid Training and Multimodal Fusion 14 Aug 2024 · 0 repositories · arXiv:2408.07303
-
Exploring Retrieval Augmented Generation in Arabic 14 Aug 2024 · 1 repository · arXiv:2408.07425
-
LiPCoT: Linear Predictive Coding based Tokenizer for Self-supervised Learning of Time Series Data via Language Models 14 Aug 2024 · 1 repository · arXiv:2408.07292
-
SAGE-RT: Synthetic Alignment data Generation for Safety Evaluation and Red Teaming 14 Aug 2024 · 0 repositories · arXiv:2408.11851
-
Transformers and Large Language Models for Efficient Intrusion Detection Systems: A Comprehensive Survey 14 Aug 2024 · 0 repositories · arXiv:2408.07583
-
BERT's Conceptual Cartography: Mapping the Landscapes of Meaning 13 Aug 2024 · 0 repositories · arXiv:2408.07190
-
Evaluating Cultural Adaptability of a Large Language Model via Simulation of Synthetic Personas 13 Aug 2024 · 1 repository · arXiv:2408.06929
-
Generative AI for automatic topic labelling 13 Aug 2024 · 0 repositories · arXiv:2408.07003
-
Pragmatic inference of scalar implicature by LLMs 13 Aug 2024 · 0 repositories · arXiv:2408.06673
-
TableGuard -- Securing Structured & Unstructured Data 13 Aug 2024 · 0 repositories · arXiv:2408.07045
-
VulCatch: Enhancing Binary Vulnerability Detection through CodeT5 Decompilation and KAN Advanced Feature Extraction 13 Aug 2024 · 0 repositories · arXiv:2408.07181
-
Bayesian inference to improve quality of Retrieval Augmented Generation 12 Aug 2024 · 0 repositories · arXiv:2408.08901
-
Improving Structural Diversity of Blackbox LLMs via Chain-of-Specification Prompting 12 Aug 2024 · 0 repositories · arXiv:2408.06186
-
LOLgorithm: Integrating Semantic,Syntactic and Contextual Elements for Humor Classification 12 Aug 2024 · 0 repositories · arXiv:2408.06335
-
Optimizing RAG Techniques for Automotive Industry PDF Chatbots: A Case Study with Locally Deployed Ollama Models 12 Aug 2024 · 0 repositories · arXiv:2408.05933
-
The Language of Trauma: Modeling Traumatic Event Descriptions Across Domains with Explainable AI 12 Aug 2024 · 0 repositories · arXiv:2408.05977
-
Kov: Transferable and Naturalistic Black-Box LLM Attacks using Markov Decision Processes and Tree Search 11 Aug 2024 · 1 repository · arXiv:2408.08899
-
PhishLang: A Real-Time, Fully Client-Side Phishing Detection Framework Using MobileBERT 11 Aug 2024 · 2 repositories · arXiv:2408.05667
-
Chain of Condition: Construct, Verify and Solve Conditions for Conditional Question Answering 10 Aug 2024 · 0 repositories · arXiv:2408.05442
-
Improving Whisper's Recognition Performance for Under-Represented Language Kazakh Leveraging Unpaired Speech and Text 10 Aug 2024 · 0 repositories · arXiv:2408.05554
-
A Hybrid RAG System with Comprehensive Enhancement on Complex Reasoning 9 Aug 2024 · 0 repositories · arXiv:2408.05141
-
ConfusedPilot: Confused Deputy Risks in RAG-based LLMs 9 Aug 2024 · 0 repositories · arXiv:2408.04870
-
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis 9 Aug 2024 · 1 repository · arXiv:2408.05006
-
Ensemble BERT: A student social network text sentiment classification model based on ensemble learning and BERT architecture 9 Aug 2024 · 0 repositories · arXiv:2408.04849
-
Examining the Behavior of LLM Architectures Within the Framework of Standardized National Exams in Brazil 9 Aug 2024 · 0 repositories · arXiv:2408.05035
-
From Text to Insight: Leveraging Large Language Models for Performance Evaluation in Management 9 Aug 2024 · 0 repositories · arXiv:2408.05328
-
HybridRAG: Integrating Knowledge Graphs and Vector Retrieval Augmented Generation for Efficient Information Extraction 9 Aug 2024 · 0 repositories · arXiv:2408.04948
-
Rag and Roll: An End-to-End Evaluation of Indirect Prompt Manipulations in LLM-based Application Frameworks 9 Aug 2024 · 0 repositories · arXiv:2408.05025
-
Retrieval-augmented code completion for local projects using large language models 9 Aug 2024 · 0 repositories · arXiv:2408.05026
-
Analysis of Argument Structure Constructions in the Large Language Model BERT 8 Aug 2024 · 0 repositories · arXiv:2408.04270
-
EfficientRAG: Efficient Retriever for Multi-Hop Question Answering 8 Aug 2024 · 1 repository · arXiv:2408.04259Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Hybrid Student-Teacher Large Language Model Refinement for Cancer Toxicity Symptom Extraction 8 Aug 2024 · 0 repositories · arXiv:2408.04775
-
Medical Graph RAG: Towards Safe Medical Large Language Model via Graph Retrieval-Augmented Generation 8 Aug 2024 · 1 repository · arXiv:2408.04187Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
SCENE: Evaluating Explainable AI Techniques Using Soft Counterfactuals 8 Aug 2024 · 0 repositories · arXiv:2408.04575
-
Transformer Explainer: Interactive Learning of Text-Generative Models 8 Aug 2024 · 1 repository · arXiv:2408.04619
-
A Comparison of LLM Finetuning Methods & Evaluation Metrics with Travel Chatbot Use Case 7 Aug 2024 · 0 repositories · arXiv:2408.03562
-
Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants 7 Aug 2024 · 0 repositories · arXiv:2408.11841
-
VulScribeR: Exploring RAG-based Vulnerability Augmentation with LLMs 7 Aug 2024 · 1 repository · arXiv:2408.04125
-
Image-to-LaTeX Converter for Mathematical Formulas and Text 7 Aug 2024 · 1 repository · arXiv:2408.04015
-
Is Child-Directed Speech Effective Training Data for Language Models? 7 Aug 2024 · 1 repository · arXiv:2408.03617Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Query3D: LLM-Powered Open-Vocabulary Scene Segmentation with Language Embedded 3D Gaussian 7 Aug 2024 · 1 repository · arXiv:2408.03516
-
MaxMind: A Memory Loop Network to Enhance Software Productivity based on Large Language Models 7 Aug 2024 · 0 repositories · arXiv:2408.03841
-
SocFedGPT: Federated GPT-based Adaptive Content Filtering System Leveraging User Interactions in Social Networks 7 Aug 2024 · 0 repositories · arXiv:2408.05243
-
Data Poisoning in LLMs: Jailbreak-Tuning and Scaling Laws 6 Aug 2024 · 2 repositories · arXiv:2408.02946Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Empathy Level Alignment via Reinforcement Learning for Empathetic Response Generation 6 Aug 2024 · 1 repository · arXiv:2408.02976
-
Topic Modeling with Fine-tuning LLMs and Bag of Sentences 6 Aug 2024 · 1 repository · arXiv:2408.03099Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating the Translation Performance of Large Language Models Based on Euas-20 6 Aug 2024 · 0 repositories · arXiv:2408.03119
-
Leveraging Parameter Efficient Training Methods for Low Resource Text Classification: A Case Study in Marathi 6 Aug 2024 · 0 repositories · arXiv:2408.03172
-
FLASH: Federated Learning-Based LLMs for Advanced Query Processing in Social Networks through RAG 6 Aug 2024 · 0 repositories · arXiv:2408.05242
-
The Use of Large Language Models (LLM) for Cyber Threat Intelligence (CTI) in Cybercrime Forums 6 Aug 2024 · 0 repositories · arXiv:2408.03354
-
TrafficGPT: An LLM Approach for Open-Set Encrypted Traffic Classification 6 Aug 2024 · 1 repository
-
Training LLMs to Recognize Hedges in Spontaneous Narratives 6 Aug 2024 · 1 repository · arXiv:2408.03319
-
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02416
-
RAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented Generation 5 Aug 2024 · 2 repositories · arXiv:2408.02545
-
Wiping out the limitations of Large Language Models -- A Taxonomy for Retrieval Augmented Generation 5 Aug 2024 · 0 repositories · arXiv:2408.02854
-
AppAgent v2: Advanced Agent for Flexible Mobile Interactions 5 Aug 2024 · 0 repositories · arXiv:2408.11824
-
LLM Agents Improve Semantic Code Search 5 Aug 2024 · 0 repositories · arXiv:2408.11058
-
XMainframe: A Large Language Model for Mainframe Modernization 5 Aug 2024 · 1 repository · arXiv:2408.04660
-
Defining and Evaluating Decision and Composite Risk in Language Models Applied to Natural Language Inference 4 Aug 2024 · 0 repositories · arXiv:2408.01935
-
ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science 4 Aug 2024 · 1 repository · arXiv:2408.01966
-
AdaCBM: An Adaptive Concept Bottleneck Model for Explainable and Accurate Diagnosis 4 Aug 2024 · 1 repository · arXiv:2408.02001
-
Effective Demonstration Annotation for In-Context Learning via Language Model-Based Determinantal Point Process 4 Aug 2024 · 0 repositories · arXiv:2408.02103
-
Leveraging Large Language Models with Chain-of-Thought and Prompt Engineering for Traffic Crash Severity Analysis and Inference 4 Aug 2024 · 0 repositories · arXiv:2408.04652
-
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment 3 Aug 2024 · 0 repositories · arXiv:2408.01614
-
Indexing and Visualization of Climate Change Narratives Using BERT and Causal Extraction 3 Aug 2024 · 0 repositories · arXiv:2408.01745
-
Tracking Emotional Dynamics in Chat Conversations: A Hybrid Approach using DistilBERT and Emoji Sentiment Analysis 3 Aug 2024 · 0 repositories · arXiv:2408.01838
-
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly 3 Aug 2024 · 0 repositories · arXiv:2408.01866