Methods › General › Regularization › Attention Dropout › Papers, page 15
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 15 of 109: papers 1,401 to 1,500 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Know Your RAG: Dataset Taxonomy and Generation Strategies for Evaluating RAG Systems 29 Nov 2024 · 0 repositories · arXiv:2411.19710
-
Knowledge Management for Automobile Failure Analysis Using Graph RAG 29 Nov 2024 · 0 repositories · arXiv:2411.19539
-
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation 29 Nov 2024 · 0 repositories · arXiv:2411.19528
-
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation 29 Nov 2024 · 0 repositories · arXiv:2411.19921
-
Towards Santali Linguistic Inclusion: Building the First Santali-to-English Translation Model using mT5 Transformer and Data Augmentation 29 Nov 2024 · 0 repositories · arXiv:2411.19726
-
Towards Understanding Retrieval Accuracy and Prompt Quality in RAG Systems 29 Nov 2024 · 0 repositories · arXiv:2411.19463
-
An Extensive Evaluation of Factual Consistency in Large Language Models for Data-to-Text Generation 28 Nov 2024 · 0 repositories · arXiv:2411.19203
-
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection 28 Nov 2024 · 0 repositories · arXiv:2411.19220
-
Beautimeter: Harnessing GPT for Assessing Architectural and Urban Beauty based on the 15 Properties of Living Structure 28 Nov 2024 · 0 repositories · arXiv:2411.19094
-
DENIAHL: In-Context Features Influence LLM Needle-In-A-Haystack Abilities 28 Nov 2024 · 1 repository · arXiv:2411.19360
-
Efficient Learning Content Retrieval with Knowledge Injection 28 Nov 2024 · 0 repositories · arXiv:2412.00125
-
Habit Coach: Customising RAG-based chatbots to support behavior change 28 Nov 2024 · 0 repositories · arXiv:2411.19229
-
RevPRAG: Revealing Poisoning Attacks in Retrieval-Augmented Generation through LLM Activation Analysis 28 Nov 2024 · 0 repositories · arXiv:2411.18948
-
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation 28 Nov 2024 · 1 repository · arXiv:2411.19067
-
SmartLLMSentry: A Comprehensive LLM Based Smart Contract Vulnerability Detection Framework 28 Nov 2024 · 0 repositories · arXiv:2411.19234
-
The Impact of Example Selection in Few-Shot Prompting on Automated Essay Scoring Using GPT Models 28 Nov 2024 · 0 repositories · arXiv:2411.18924
-
A survey on cutting-edge relation extraction techniques based on language models 27 Nov 2024 · 0 repositories · arXiv:2411.18157
-
Automated Literature Review Using NLP Techniques and LLM-Based Retrieval-Augmented Generation 27 Nov 2024 · 0 repositories · arXiv:2411.18583
-
Can bidirectional encoder become the ultimate winner for downstream applications of foundation models? 27 Nov 2024 · 0 repositories · arXiv:2411.18021
-
ChatGPT as speechwriter for the French presidents 27 Nov 2024 · 0 repositories · arXiv:2411.18382
-
DRS: Deep Question Reformulation With Structured Output 27 Nov 2024 · 1 repository · arXiv:2411.17993
-
Evaluating and Improving the Robustness of Security Attack Detectors Generated by LLMs 27 Nov 2024 · 1 repository · arXiv:2411.18216
-
Fine-Tuning Large Language Models for Scientific Text Classification: A Comparative Study 27 Nov 2024 · 0 repositories · arXiv:2412.00098
-
Fine-Tuning Small Embeddings for Elevated Performance 27 Nov 2024 · 0 repositories · arXiv:2411.18099
-
On Importance of Code-Mixed Embeddings for Hate Speech Identification 27 Nov 2024 · 0 repositories · arXiv:2411.18577
-
Streamlining Prediction in Bayesian Deep Learning 27 Nov 2024 · 1 repository · arXiv:2411.18425Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Training and Evaluating Language Models with Template-based Data Generation 27 Nov 2024 · 1 repository · arXiv:2411.18104Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Training Noise Token Pruning 27 Nov 2024 · 1 repository · arXiv:2411.18092
-
Advancing Content Moderation: Evaluating Large Language Models for Detecting Sensitive Content Across Text, Images, and Videos 26 Nov 2024 · 0 repositories · arXiv:2411.17123
-
BERT or FastText? A Comparative Analysis of Contextual as well as Non-Contextual Embeddings 26 Nov 2024 · 1 repository · arXiv:2411.17661
-
Can artificial intelligence predict clinical trial outcomes? 26 Nov 2024 · 0 repositories · arXiv:2411.17595
-
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning 26 Nov 2024 · 1 repository · arXiv:2411.17426
-
Distributed Sign Momentum with Local Steps for Training Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17866
-
Fairness And Performance In Harmony: Data Debiasing Is All You Need 26 Nov 2024 · 0 repositories · arXiv:2411.17374
-
"Give me the code" -- Log Analysis of First-Year CS Students' Interactions With GPT 26 Nov 2024 · 0 repositories · arXiv:2411.17855
-
On Limitations of LLM as Annotator for Low Resource Languages 26 Nov 2024 · 0 repositories · arXiv:2411.17637
-
Pretrained LLM Adapted with LoRA as a Decision Transformer for Offline RL in Quantitative Trading 26 Nov 2024 · 1 repository · arXiv:2411.17900
-
Scalable iterative pruning of large language and vision models using block coordinate descent 26 Nov 2024 · 0 repositories · arXiv:2411.17796
-
What Differentiates Educational Literature? A Multimodal Fusion Approach of Transformers and Computational Linguistics 26 Nov 2024 · 0 repositories · arXiv:2411.17593
-
Adaptive Circuit Behavior and Generalization in Mechanistic Interpretability 25 Nov 2024 · 0 repositories · arXiv:2411.16105
-
Are Transformers Truly Foundational for Robotics? 25 Nov 2024 · 0 repositories · arXiv:2411.16917
-
AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning 25 Nov 2024 · 1 repository · arXiv:2411.16495Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16991
-
Fine-Tuning LLMs with Noisy Data for Political Argument Generation and Post Guidance 25 Nov 2024 · 0 repositories · arXiv:2411.16813
-
Human-Calibrated Automated Testing and Validation of Generative Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16391
-
LaB-RAG: Label Boosted Retrieval Augmented Generation for Radiology Report Generation 25 Nov 2024 · 1 repository · arXiv:2411.16523
-
MarketGPT: Developing a Pre-trained transformer (GPT) for Modeling Financial Time Series 25 Nov 2024 · 1 repository · arXiv:2411.16585Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Predictive Power of LLMs in Financial Markets 25 Nov 2024 · 0 repositories · arXiv:2411.16569
-
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training 25 Nov 2024 · 0 repositories · arXiv:2411.16618
-
Development of Pre-Trained Transformer-based Models for the Nepali Language 24 Nov 2024 · 0 repositories · arXiv:2411.15734
-
RAMIE: Retrieval-Augmented Multi-task Information Extraction with Large Language Models on Dietary Supplements 24 Nov 2024 · 0 repositories · arXiv:2411.15700
-
A Comparative Analysis of Transformer and LSTM Models for Detecting Suicidal Ideation on Reddit 23 Nov 2024 · 1 repository · arXiv:2411.15404
-
"All that Glitters": Approaches to Evaluations with Unreliable Model and Human Annotations 23 Nov 2024 · 1 repository · arXiv:2411.15634
-
ChatBCI: A P300 Speller BCI Leveraging Large Language Models for Improved Sentence Composition in Realistic Scenarios 23 Nov 2024 · 0 repositories · arXiv:2411.15395
-
Improving Next Tokens via Second-Last Predictions with Generate and Refine 23 Nov 2024 · 0 repositories · arXiv:2411.15661
-
Inducing Human-like Biases in Moral Reasoning Language Models 23 Nov 2024 · 0 repositories · arXiv:2411.15386
-
Traditional Chinese Medicine Case Analysis System for High-Level Semantic Abstraction: Optimized with Prompt and RAG 23 Nov 2024 · 0 repositories · arXiv:2411.15491
-
Astro-HEP-BERT: A bidirectional language model for studying the meanings of concepts in astrophysics and high energy physics 22 Nov 2024 · 0 repositories · arXiv:2411.14877
-
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective 22 Nov 2024 · 0 repositories · arXiv:2411.14654
-
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases 22 Nov 2024 · 1 repository · arXiv:2411.14790
-
An Experimental Study on Data Augmentation Techniques for Named Entity Recognition on Low-Resource Domains 21 Nov 2024 · 0 repositories · arXiv:2411.14551
-
Assessment of LLM Responses to End-user Security Questions 21 Nov 2024 · 0 repositories · arXiv:2411.14571
-
BERT-Based Approach for Automating Course Articulation Matrix Construction with Explainable AI 21 Nov 2024 · 1 repository · arXiv:2411.14254
-
Evaluating the Robustness of Analogical Reasoning in Large Language Models 21 Nov 2024 · 1 repository · arXiv:2411.14215Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
FastRAG: Retrieval Augmented Generation for Semi-structured Data 21 Nov 2024 · 0 repositories · arXiv:2411.13773
-
G-RAG: Knowledge Expansion in Material Science 21 Nov 2024 · 1 repository · arXiv:2411.14592
-
Generative Fuzzy System for Sequence Generation 21 Nov 2024 · 0 repositories · arXiv:2411.13867
-
POS-tagging to highlight the skeletal structure of sentences 21 Nov 2024 · 2 repositories · arXiv:2411.14393
-
Towards Knowledge Checking in Retrieval-augmented Generation: A Representation Perspective 21 Nov 2024 · 0 repositories · arXiv:2411.14572
-
AI-Driven Agents with Prompts Designed for High Agreeableness Increase the Likelihood of Being Mistaken for a Human in the Turing Test 20 Nov 2024 · 0 repositories · arXiv:2411.13749
-
Combining Autoregressive and Autoencoder Language Models for Text Classification 20 Nov 2024 · 1 repository · arXiv:2411.13282
-
DMQR-RAG: Diverse Multi-Query Rewriting for RAG 20 Nov 2024 · 0 repositories · arXiv:2411.13154
-
Exploring Large Language Models for Climate Forecasting 20 Nov 2024 · 0 repositories · arXiv:2411.13724
-
Multimodal large language model for wheat breeding: a new exploration of smart breeding 20 Nov 2024 · 0 repositories · arXiv:2411.15203
-
On the Way to LLM Personalization: Learning to Remember User Conversations 20 Nov 2024 · 0 repositories · arXiv:2411.13405
-
Retrieval-Augmented Generation for Domain-Specific Question Answering: A Case Study on Pittsburgh and CMU 20 Nov 2024 · 0 repositories · arXiv:2411.13691
-
Unlocking Historical Clinical Trial Data with ALIGN: A Compositional Large Language Model System for Medical Coding 20 Nov 2024 · 0 repositories · arXiv:2411.13163
-
A Combined Encoder and Transformer Approach for Coherent and High-Quality Text Generation 19 Nov 2024 · 0 repositories · arXiv:2411.12157
-
DLBacktrace: A Model Agnostic Explainability for any Deep Learning Models 19 Nov 2024 · 1 repository · arXiv:2411.12643
-
Enhancing Multi-Class Disease Classification: Neoplasms, Cardiovascular, Nervous System, and Digestive Disorders Using Advanced LLMs 19 Nov 2024 · 0 repositories · arXiv:2411.12712
-
Leveraging Virtual Reality and AI Tutoring for Language Learning: A Case Study of a Virtual Campus Environment with OpenAI GPT Integration with Unity 3D 19 Nov 2024 · 0 repositories · arXiv:2411.12619
-
Strengthening Fake News Detection: Leveraging SVM and Sophisticated Text Vectorization Techniques. Defying BERT? 19 Nov 2024 · 0 repositories · arXiv:2411.12703
-
Can Open-source LLMs Enhance Data Synthesis for Toxic Detection?: An Experimental Study 18 Nov 2024 · 0 repositories · arXiv:2411.15175
-
Chapter 7 Review of Data-Driven Generative AI Models for Knowledge Extraction from Scientific Literature in Healthcare 18 Nov 2024 · 0 repositories · arXiv:2411.11635
-
CNMBERT: A Model for Converting Hanyu Pinyin Abbreviations to Chinese Characters 18 Nov 2024 · 1 repository · arXiv:2411.11770
-
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback 18 Nov 2024 · 1 repository · arXiv:2412.03578Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Suicide Risk Assessment on Social Media with Semi-Supervised Learning 18 Nov 2024 · 0 repositories · arXiv:2411.12767
-
Understanding Student Sentiment on Mental Health Support in Colleges Using Large Language Models 18 Nov 2024 · 0 repositories · arXiv:2412.04326
-
VersaTune: An Efficient Data Composition Framework for Training Multi-Capability LLMs 18 Nov 2024 · 1 repository · arXiv:2411.11266
-
A Novel Approach to Eliminating Hallucinations in Large Language Model-Assisted Causal Discovery 16 Nov 2024 · 0 repositories · arXiv:2411.12759
-
Debias-CLR: A Contrastive Learning Based Debiasing Method for Algorithmic Fairness in Healthcare Applications 15 Nov 2024 · 0 repositories · arXiv:2411.10544
-
Does Prompt Formatting Have Any Impact on LLM Performance? 15 Nov 2024 · 0 repositories · arXiv:2411.10541
-
Hysteresis Activation Function for Efficient Inference 15 Nov 2024 · 1 repository · arXiv:2411.10573Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Information Extraction from Clinical Notes: Are We Ready to Switch to Large Language Models? 15 Nov 2024 · 1 repository · arXiv:2411.10020
-
MARS: Unleashing the Power of Variance Reduction for Training Large Models 15 Nov 2024 · 2 repositories · arXiv:2411.10438Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation 15 Nov 2024 · 0 repositories · arXiv:2411.10129
-
SoftLMs: Efficient Adaptive Low-Rank Approximation of Language Models using Soft-Thresholding Mechanism 15 Nov 2024 · 0 repositories · arXiv:2411.10543
-
Take Package as Language: Anomaly Detection Using Transformer 15 Nov 2024 · 0 repositories · arXiv:2412.04473
-
Adopting RAG for LLM-Aided Future Vehicle Design 14 Nov 2024 · 0 repositories · arXiv:2411.09590