Methods › General › Regularization › Attention Dropout › Papers, page 16
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 16 of 109: papers 1,501 to 1,600 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency 14 Nov 2024 · 0 repositories · arXiv:2411.09587
-
Beyond Static Tools: Evaluating Large Language Models for Cryptographic Misuse Detection 14 Nov 2024 · 0 repositories · arXiv:2411.09772
-
Comprehensive and Practical Evaluation of Retrieval-Augmented Generation Systems for Medical Question Answering 14 Nov 2024 · 0 repositories · arXiv:2411.09213
-
HateGPT: Unleashing GPT-3.5 Turbo to Combat Hate Speech on X 14 Nov 2024 · 0 repositories · arXiv:2411.09214
-
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework 14 Nov 2024 · 2 repositories · arXiv:2411.09607
-
A Large-Scale Study of Relevance Assessments with Large Language Models: An Initial Look 13 Nov 2024 · 1 repository · arXiv:2411.08275
-
Analyst Reports and Stock Performance: Evidence from the Chinese Market 13 Nov 2024 · 0 repositories · arXiv:2411.08726
-
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection 13 Nov 2024 · 0 repositories · arXiv:2411.08868
-
LLMStinger: Jailbreaking LLMs using RL fine-tuned LLMs 13 Nov 2024 · 0 repositories · arXiv:2411.08862
-
LogLLM: Log-based Anomaly Detection Using Large Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08561
-
Responsible AI in Construction Safety: Systematic Evaluation of Large Language Models and Prompt Engineering 13 Nov 2024 · 0 repositories · arXiv:2411.08320
-
Towards Optimizing a Retrieval Augmented Generation using Large Language Model on Academic Data 13 Nov 2024 · 0 repositories · arXiv:2411.08438
-
VALTEST: Automated Validation of Language Model Generated Test Cases 13 Nov 2024 · 0 repositories · arXiv:2411.08254
-
Controlled Evaluation of Syntactic Knowledge in Multilingual Language Models 12 Nov 2024 · 1 repository · arXiv:2411.07474
-
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis 12 Nov 2024 · 1 repository · arXiv:2411.07529
-
Fair Summarization: Bridging Quality and Diversity in Extractive Summaries 12 Nov 2024 · 1 repository · arXiv:2411.07521Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Leveraging Multimodal Models for Enhanced Neuroimaging Diagnostics in Alzheimer's Disease 12 Nov 2024 · 0 repositories · arXiv:2411.07871
-
LLM App Squatting and Cloning 12 Nov 2024 · 0 repositories · arXiv:2411.07518
-
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models 12 Nov 2024 · 0 repositories · arXiv:2411.07820
-
Retrieval Augmented Time Series Forecasting 12 Nov 2024 · 1 repository · arXiv:2411.08249Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders 12 Nov 2024 · 0 repositories · arXiv:2411.07870
-
A Unified Multi-Task Learning Architecture for Hate Detection Leveraging User-Based Information 11 Nov 2024 · 0 repositories · arXiv:2411.06855
-
Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models 11 Nov 2024 · 0 repositories · arXiv:2411.06713
-
AssistRAG: Boosting the Potential of Large Language Models with an Intelligent Information Assistant 11 Nov 2024 · 1 repository · arXiv:2411.06805Syntology official (archive's flag): 10 ran · 10 ran (of which 1 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Autonomous Droplet Microfluidic Design Framework with Large Language Models 11 Nov 2024 · 1 repository · arXiv:2411.06691
-
Cancer-Answer: Empowering Cancer Care with Advanced Large Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.06946
-
Evaluating Large Language Models on Financial Report Summarization: An Empirical Study 11 Nov 2024 · 0 repositories · arXiv:2411.06852
-
Explore the Reasoning Capability of LLMs in the Chess Testbed 11 Nov 2024 · 0 repositories · arXiv:2411.06655
-
Invar-RAG: Invariant LLM-aligned Retrieval for Better Generation 11 Nov 2024 · 0 repositories · arXiv:2411.07021
-
LA4SR: illuminating the dark proteome with generative AI 11 Nov 2024 · 0 repositories · arXiv:2411.06798
-
On Active Privacy Auditing in Supervised Fine-tuning for White-Box Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.07070
-
SPARTAN: A Sparse Transformer Learning Local Causation 11 Nov 2024 · 0 repositories · arXiv:2411.06890
-
TempCharBERT: Keystroke Dynamics for Continuous Access Control Based on Pre-trained Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.07224
-
The Backpropagation of the Wave Network 11 Nov 2024 · 0 repositories · arXiv:2411.06989
-
Toward Optimal Search and Retrieval for RAG 11 Nov 2024 · 1 repository · arXiv:2411.07396
-
Prompt-Efficient Fine-Tuning for GPT-like Deep Models to Reduce Hallucination and to Improve Reproducibility in Scientific Text Generation Using Stochastic Optimisation Techniques 10 Nov 2024 · 0 repositories · arXiv:2411.06445
-
Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement 10 Nov 2024 · 1 repository · arXiv:2411.06558Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Clustering Algorithms and RAG Enhancing Semi-Supervised Text Classification with Large LLMs 9 Nov 2024 · 0 repositories · arXiv:2411.06175
-
Detecting Reference Errors in Scientific Literature with Large Language Models 9 Nov 2024 · 1 repository · arXiv:2411.06101
-
Exploring Knowledge Boundaries in Large Language Models for Retrieval Judgment 9 Nov 2024 · 0 repositories · arXiv:2411.06207
-
Improved intent classification based on context information using a windows-based approach 9 Nov 2024 · 0 repositories · arXiv:2411.06022
-
Leveraging Retrieval-Augmented Generation for Persian University Knowledge Retrieval 9 Nov 2024 · 0 repositories · arXiv:2411.06237
-
Robust Detection of LLM-Generated Text: A Comparative Analysis 9 Nov 2024 · 0 repositories · arXiv:2411.06248
-
Sufficient Context: A New Lens on Retrieval Augmented Generation Systems 9 Nov 2024 · 0 repositories · arXiv:2411.06037
-
AgentOps: Enabling Observability of LLM Agents 8 Nov 2024 · 1 repository · arXiv:2411.05285
-
Enhancing Visual Classification using Comparative Descriptors 8 Nov 2024 · 1 repository · arXiv:2411.05357
-
FinDVer: Explainable Claim Verification over Long and Hybrid-Content Financial Documents 8 Nov 2024 · 1 repository · arXiv:2411.05764
-
GPT Semantic Cache: Reducing LLM Costs and Latency via Semantic Embedding Caching 8 Nov 2024 · 0 repositories · arXiv:2411.05276
-
HeartBERT: A Self-Supervised ECG Embedding Model for Efficient and Effective Medical Signal Analysis 8 Nov 2024 · 1 repository · arXiv:2411.11896
-
IntellBot: Retrieval Augmented LLM Chatbot for Cyber Threat Knowledge Delivery 8 Nov 2024 · 1 repository · arXiv:2411.05442
-
Learning the rules of peptide self-assembly through data mining with large language models 8 Nov 2024 · 1 repository · arXiv:2411.05421
-
Multi-Document Financial Question Answering using LLMs 8 Nov 2024 · 0 repositories · arXiv:2411.07264
-
NeKo: Toward Post Recognition Generative Correction Large Language Models with Task-Oriented Experts 8 Nov 2024 · 0 repositories · arXiv:2411.05945
-
Qwen2.5-32B: Leveraging Self-Consistent Tool-Integrated Reasoning for Bengali Mathematical Olympiad Problem Solving 8 Nov 2024 · 0 repositories · arXiv:2411.05934
-
Sentiment Analysis of Cyberbullying Data in Social Media 8 Nov 2024 · 1 repository · arXiv:2411.05958
-
Adversarial Robustness of In-Context Learning in Transformers for Linear Regression 7 Nov 2024 · 0 repositories · arXiv:2411.05189
-
Best Practices for Distilling Large Language Models into BERT for Web Search Ranking 7 Nov 2024 · 0 repositories · arXiv:2411.04539
-
Deploying Large Language Models With Retrieval Augmented Generation 7 Nov 2024 · 1 repository · arXiv:2411.11895
-
Enhancing classroom teaching with LLMs and RAG 7 Nov 2024 · 0 repositories · arXiv:2411.04341
-
FineTuneBench: How well do commercial fine-tuning APIs infuse knowledge into LLMs? 7 Nov 2024 · 1 repository · arXiv:2411.05059
-
GPT-Guided Monte Carlo Tree Search for Symbolic Regression in Financial Fraud Detection 7 Nov 2024 · 0 repositories · arXiv:2411.04459
-
M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding 7 Nov 2024 · 0 repositories · arXiv:2411.04952
-
RetrieveGPT: Merging Prompts and Mathematical Models for Enhanced Code-Mixed Information Retrieval 7 Nov 2024 · 0 repositories · arXiv:2411.04752
-
Selecting Between BERT and GPT for Text Classification in Political Science Research 7 Nov 2024 · 0 repositories · arXiv:2411.05050
-
STAND-Guard: A Small Task-Adaptive Content Moderation Model 7 Nov 2024 · 0 repositories · arXiv:2411.05214
-
Words that Move Markets- Quantifying the Impact of RBI's Monetary Policy Communications on Indian Financial Market 7 Nov 2024 · 0 repositories · arXiv:2411.04808
-
A Comparative Study of Recent Large Language Models on Generating Hospital Discharge Summaries for Lung Cancer Patients 6 Nov 2024 · 0 repositories · arXiv:2411.03805
-
A Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AI 6 Nov 2024 · 0 repositories · arXiv:2411.04316
-
Advanced RAG Models with Graph Structures: Optimizing Complex Knowledge Reasoning and Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03572
-
Can Custom Models Learn In-Context? An Exploration of Hybrid Architecture Performance on In-Context Learning Tasks 6 Nov 2024 · 1 repository · arXiv:2411.03945
-
Fine-Grained Guidance for Retrievers: Leveraging LLMs' Feedback in Retrieval-Augmented Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03957
-
From Word Vectors to Multimodal Embeddings: Techniques, Applications, and Future Directions For Large Language Models 6 Nov 2024 · 0 repositories · arXiv:2411.05036
-
On-Device Emoji Classifier Trained with GPT-based Data Augmentation for a Mobile Keyboard 6 Nov 2024 · 0 repositories · arXiv:2411.05031
-
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology 6 Nov 2024 · 0 repositories · arXiv:2411.10473
-
Prompt Engineering Using GPT for Word-Level Code-Mixed Language Identification in Low-Resource Dravidian Languages 6 Nov 2024 · 0 repositories · arXiv:2411.04025
-
RAGulator: Lightweight Out-of-Context Detectors for Grounded Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03920
-
Towards Interpreting Language Models: A Case Study in Multi-Hop Reasoning 6 Nov 2024 · 1 repository · arXiv:2411.05037Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Understanding the Effects of Human-written Paraphrases in LLM-generated Text Detection 6 Nov 2024 · 1 repository · arXiv:2411.03806
-
YouTube Comments Decoded: Leveraging LLMs for Low Resource Language Classification 6 Nov 2024 · 0 repositories · arXiv:2411.05039
-
Automatic Generation of Question Hints for Mathematics Problems using Large Language Models in Educational Technology 5 Nov 2024 · 0 repositories · arXiv:2411.03495
-
Enhancing Transformer Training Efficiency with Dynamic Dropout 5 Nov 2024 · 0 repositories · arXiv:2411.03236
-
Exploring the Benefits of Domain-Pretraining of Generative Large Language Models for Chemistry 5 Nov 2024 · 0 repositories · arXiv:2411.03542
-
HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems 5 Nov 2024 · 1 repository · arXiv:2411.02959
-
LASER: Attention with Exponential Transformation 5 Nov 2024 · 0 repositories · arXiv:2411.03493
-
Long Context RAG Performance of Large Language Models 5 Nov 2024 · 0 repositories · arXiv:2411.03538
-
PersianRAG: A Retrieval-Augmented Generation System for Persian Language 5 Nov 2024 · 0 repositories · arXiv:2411.02832
-
A Comparative Analysis of Counterfactual Explanation Methods for Text Classifiers 4 Nov 2024 · 0 repositories · arXiv:2411.02643
-
Advancements and limitations of LLMs in replicating human color-word associations 4 Nov 2024 · 0 repositories · arXiv:2411.02116
-
Ask, and it shall be given: On the Turing completeness of prompting 4 Nov 2024 · 1 repository · arXiv:2411.01992
-
Can Language Models Enable In-Context Database? 4 Nov 2024 · 0 repositories · arXiv:2411.01807
-
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast 4 Nov 2024 · 0 repositories · arXiv:2411.02318
-
Grounding Emotional Descriptions to Electrovibration Haptic Signals 4 Nov 2024 · 0 repositories · arXiv:2411.02118
-
MdEval: Massively Multilingual Code Debugging 4 Nov 2024 · 0 repositories · arXiv:2411.02310
-
RAGViz: Diagnose and Visualize Retrieval-Augmented Generation 4 Nov 2024 · 1 repository · arXiv:2411.01751
-
Regress, Don't Guess -- A Regression-like Loss on Number Tokens for Language Models 4 Nov 2024 · 1 repository · arXiv:2411.02083Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
TeleOracle: Fine-Tuned Retrieval-Augmented Generation with Long-Context Support for Network 4 Nov 2024 · 1 repository · arXiv:2411.02617
-
Training-free Regional Prompting for Diffusion Transformers 4 Nov 2024 · 1 repository · arXiv:2411.02395Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Wave Network: An Ultra-Small Language Model 4 Nov 2024 · 0 repositories · arXiv:2411.02674
-
Data Extraction Attacks in Retrieval-Augmented Generation via Backdoors 3 Nov 2024 · 0 repositories · arXiv:2411.01705
-
Enriching Tabular Data with Contextual LLM Embeddings: A Comprehensive Ablation Study for Ensemble Classifiers 3 Nov 2024 · 0 repositories · arXiv:2411.01645