Methods › General › Regularization › Attention Dropout › Papers, page 35
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 35 of 109: papers 3,401 to 3,500 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena 11 Mar 2024 · 0 repositories · arXiv:2403.06965
-
Narrating Causal Graphs with Large Language Models 11 Mar 2024 · 0 repositories · arXiv:2403.07118
-
An Audio-textual Diffusion Model For Converting Speech Signals Into Ultrasound Tongue Imaging Data 9 Mar 2024 · 0 repositories · arXiv:2403.05820
-
Enhancing Multi-Hop Knowledge Graph Reasoning through Reward Shaping Techniques 9 Mar 2024 · 0 repositories · arXiv:2403.05801
-
TokenMark: A Modality-Agnostic Watermark for Pre-trained Transformers 9 Mar 2024 · 0 repositories · arXiv:2403.05842
-
Segmentation Guided Sparse Transformer for Under-Display Camera Image Restoration 9 Mar 2024 · 0 repositories · arXiv:2403.05906
-
A Novel Nuanced Conversation Evaluation Framework for Large Language Models in Mental Health 8 Mar 2024 · 0 repositories · arXiv:2403.09705
-
Bias-Augmented Consistency Training Reduces Biased Reasoning in Chain-of-Thought 8 Mar 2024 · 1 repository · arXiv:2403.05518Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
PipeRAG: Fast Retrieval-Augmented Generation via Algorithm-System Co-design 8 Mar 2024 · 0 repositories · arXiv:2403.05676
-
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation 8 Mar 2024 · 1 repository · arXiv:2403.05313
-
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers 8 Mar 2024 · 0 repositories · arXiv:2403.05365
-
Automating the Information Extraction from Semi-Structured Interview Transcripts 7 Mar 2024 · 1 repository · arXiv:2403.04819
-
Federated Recommendation via Hybrid Retrieval Augmented Generation 7 Mar 2024 · 1 repository · arXiv:2403.04256
-
Feedback-Generation for Programming Exercises With GPT-4 7 Mar 2024 · 0 repositories · arXiv:2403.04449
-
Telecom Language Models: Must They Be Large? 7 Mar 2024 · 0 repositories · arXiv:2403.04666
-
Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem 6 Mar 2024 · 1 repository · arXiv:2403.03558
-
Can Large Language Models do Analytical Reasoning? 6 Mar 2024 · 0 repositories · arXiv:2403.04031
-
Enhancing ASD detection accuracy: a combined approach of machine learning and deep learning models with natural language processing 6 Mar 2024 · 1 repository · arXiv:2403.03581
-
FaaF: Facts as a Function for the evaluation of generated text 6 Mar 2024 · 1 repository · arXiv:2403.03888
-
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection 6 Mar 2024 · 3 repositories · arXiv:2403.03507
-
General2Specialized LLMs Translation for E-commerce 6 Mar 2024 · 0 repositories · arXiv:2403.03689
-
Guiding Enumerative Program Synthesis with Large Language Models 6 Mar 2024 · 0 repositories · arXiv:2403.03997
-
Japanese-English Sentence Translation Exercises Dataset for Automatic Grading 6 Mar 2024 · 0 repositories · arXiv:2403.03396
-
Rapidly Developing High-quality Instruction Data and Evaluation Benchmark for Large Language Models with Minimal Human Effort: A Case Study on Japanese 6 Mar 2024 · 2 repositories · arXiv:2403.03690
-
Evaluating and Optimizing Educational Content with Large Language Model Judgments 5 Mar 2024 · 1 repository · arXiv:2403.02795Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring Naive Approaches to Tell Apart LLMs Productions from Human-written Text 5 Mar 2024 · 1 repository
-
Improving Event Definition Following For Zero-Shot Event Detection 5 Mar 2024 · 0 repositories · arXiv:2403.02586
-
JMI at SemEval 2024 Task 3: Two-step approach for multimodal ECAC using in-context learning with GPT and instruction-tuned Llama models 5 Mar 2024 · 1 repository · arXiv:2403.04798Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Knowledge Graphs as Context Sources for LLM-Based Explanations of Learning Recommendations 5 Mar 2024 · 0 repositories · arXiv:2403.03008
-
MathScale: Scaling Instruction Tuning for Mathematical Reasoning 5 Mar 2024 · 1 repository · arXiv:2403.02884
-
MeanCache: User-Centric Semantic Caching for LLM Web Services 5 Mar 2024 · 0 repositories · arXiv:2403.02694
-
Towards Democratized Flood Risk Management: An Advanced AI Assistant Enabled by GPT-4 for Enhanced Interpretability and Public Engagement 5 Mar 2024 · 2 repositories · arXiv:2403.03188
-
Zero-Shot Cross-Lingual Document-Level Event Causality Identification with Heterogeneous Graph Contrastive Transfer Learning 5 Mar 2024 · 0 repositories · arXiv:2403.02893
-
Automated Generation of Multiple-Choice Cloze Questions for Assessing English Vocabulary Using GPT-turbo 3.5 4 Mar 2024 · 0 repositories · arXiv:2403.02078
-
Can LLMs Generate Architectural Design Decisions? -An Exploratory Empirical study 4 Mar 2024 · 0 repositories · arXiv:2403.01709
-
Differentially Private Synthetic Data via Foundation Model APIs 2: Text 4 Mar 2024 · 2 repositories · arXiv:2403.01749Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
EEE-QA: Exploring Effective and Efficient Question-Answer Representations 4 Mar 2024 · 1 repository · arXiv:2403.02176
-
Hypertext Entity Extraction in Webpage 4 Mar 2024 · 0 repositories · arXiv:2403.01698
-
NoteLLM: A Retrievable Large Language Model for Note Recommendation 4 Mar 2024 · 0 repositories · arXiv:2403.01744
-
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models 4 Mar 2024 · 0 repositories · arXiv:2403.02246
-
ProTrix: Building Models for Planning and Reasoning over Tables with Sentence Context 4 Mar 2024 · 1 repository · arXiv:2403.02177
-
SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis 4 Mar 2024 · 1 repository · arXiv:2403.01976Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Using LLMs for the Extraction and Normalization of Product Attribute Values 4 Mar 2024 · 1 repository · arXiv:2403.02130
-
Vanilla Transformers are Transfer Capability Teachers 4 Mar 2024 · 0 repositories · arXiv:2403.01994
-
Fine Tuning vs. Retrieval Augmented Generation for Less Popular Knowledge 3 Mar 2024 · 1 repository · arXiv:2403.01432Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Multi-level Product Category Prediction through Text Classification 3 Mar 2024 · 1 repository · arXiv:2403.01638
-
SERVAL: Synergy Learning between Vertical Models and LLMs towards Oracle-Level Zero-shot Medical Prediction 3 Mar 2024 · 0 repositories · arXiv:2403.01570Syntology 6 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Analysis of Privacy Leakage in Federated Large Language Models 2 Mar 2024 · 1 repository · arXiv:2403.04784Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks 2 Mar 2024 · 1 repository · arXiv:2403.04783Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LM4OPT: Unveiling the Potential of Large Language Models in Formulating Mathematical Optimization Problems 2 Mar 2024 · 0 repositories · arXiv:2403.01342
-
RAGged Edges: The Double-Edged Sword of Retrieval-Augmented Chatbots 2 Mar 2024 · 0 repositories · arXiv:2403.01193
-
ATP: Enabling Fast LLM Serving via Attention on Top Principal Keys 1 Mar 2024 · 0 repositories · arXiv:2403.02352
-
Gender Bias in Large Language Models across Multiple Languages 1 Mar 2024 · 0 repositories · arXiv:2403.00277
-
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction 1 Mar 2024 · 0 repositories · arXiv:2403.00528
-
SoftTiger: A Clinical Foundation Model for Healthcare Workflows 1 Mar 2024 · 1 repository · arXiv:2403.00868
-
ARTiST: Automated Text Simplification for Task Guidance in Augmented Reality 29 Feb 2024 · 1 repository · arXiv:2402.18797
-
Crafting Knowledge: Exploring the Creative Mechanisms of Chat-Based Search Engines 29 Feb 2024 · 0 repositories · arXiv:2402.19421
-
Leveraging pre-trained language models for code generation 29 Feb 2024 · 1 repository
-
LLM-Ensemble: Optimal Large Language Model Ensemble Method for E-commerce Product Attribute Value Extraction 29 Feb 2024 · 0 repositories · arXiv:2403.00863
-
PaECTER: Patent-level Representation Learning using Citation-informed Transformers 29 Feb 2024 · 0 repositories · arXiv:2402.19411
-
PeLLE: Encoder-based language models for Brazilian Portuguese based on open data 29 Feb 2024 · 0 repositories · arXiv:2402.19204
-
PROC2PDDL: Open-Domain Planning Representations from Texts 29 Feb 2024 · 0 repositories · arXiv:2403.00092
-
Prompting ChatGPT for Translation: A Comparative Analysis of Translation Brief and Persona Prompts 29 Feb 2024 · 0 repositories · arXiv:2403.00127
-
Retrieval-Augmented Generation for AI-Generated Content: A Survey 29 Feb 2024 · 3 repositories · arXiv:2402.19473
-
RL-GPT: Integrating Reinforcement Learning and Code-as-policy 29 Feb 2024 · 0 repositories · arXiv:2402.19299
-
VIXEN: Visual Text Comparison Network for Image Difference Captioning 29 Feb 2024 · 0 repositories · arXiv:2402.19119
-
Can GPT Improve the State of Prior Authorization via Guideline Based Automated Question Answering? 28 Feb 2024 · 0 repositories · arXiv:2402.18419
-
Clustering and Ranking: Diversity-preserved Instruction Selection through Expert-aligned Quality Estimation 28 Feb 2024 · 1 repository · arXiv:2402.18191Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Decomposed Prompting: Unveiling Multilingual Linguistic Structure Knowledge in English-Centric Large Language Models 28 Feb 2024 · 0 repositories · arXiv:2402.18397
-
Few-Shot Fairness: Unveiling LLM's Potential for Fairness-Aware Classification 28 Feb 2024 · 0 repositories · arXiv:2402.18502
-
SparseLLM: Towards Global Pruning for Pre-trained Language Models 28 Feb 2024 · 2 repositories · arXiv:2402.17946Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Keeping LLMs Aligned After Fine-tuning: The Crucial Role of Prompt Templates 28 Feb 2024 · 1 repository · arXiv:2402.18540Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Orchid: Flexible and Data-Dependent Convolution for Sequence Modeling 28 Feb 2024 · 0 repositories · arXiv:2402.18508
-
WIKIGENBENCH: Exploring Full-length Wikipedia Generation under Real-World Scenario 28 Feb 2024 · 1 repository · arXiv:2402.18264
-
Unsupervised Information Refinement Training of Large Language Models for Retrieval-Augmented Generation 28 Feb 2024 · 1 repository · arXiv:2402.18150Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
A Language Model based Framework for New Concept Placement in Ontologies 27 Feb 2024 · 1 repository · arXiv:2402.17897
-
COCOA: CBT-based Conversational Counseling Agent using Memory Specialized in Cognitive Distortions and Dynamic Prompt 27 Feb 2024 · 0 repositories · arXiv:2402.17546
-
Deep Learning Detection Method for Large Language Models-Generated Scientific Content 27 Feb 2024 · 0 repositories · arXiv:2403.00828
-
Emotional Voice Messages (EMOVOME) database: emotion recognition in spontaneous voice messages 27 Feb 2024 · 0 repositories · arXiv:2402.17496
-
Evaluating Very Long-Term Conversational Memory of LLM Agents 27 Feb 2024 · 1 repository · arXiv:2402.17753Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems 27 Feb 2024 · 1 repository · arXiv:2402.17840Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
JMLR: Joint Medical LLM and Retrieval Training for Enhancing Reasoning and Professional Question Answering Capability 27 Feb 2024 · 1 repository · arXiv:2402.17887
-
Linguistic Knowledge Can Enhance Encoder-Decoder Models (If You Let It) 27 Feb 2024 · 1 repository · arXiv:2402.17608
-
Measuring Vision-Language STEM Skills of Neural Models 27 Feb 2024 · 1 repository · arXiv:2402.17205Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MAGPIE: Multi-Task Media-Bias Analysis Generalization for Pre-Trained Identification of Expressions 27 Feb 2024 · 1 repository · arXiv:2403.07910
-
REAR: A Relevance-Aware Retrieval-Augmented Framework for Open-Domain Question Answering 27 Feb 2024 · 1 repository · arXiv:2402.17497Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
SKT5SciSumm -- Revisiting Extractive-Generative Approach for Multi-Document Scientific Summarization 27 Feb 2024 · 0 repositories · arXiv:2402.17311
-
Variational Learning is Effective for Large Deep Networks 27 Feb 2024 · 1 repository · arXiv:2402.17641Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Adaptation of Biomedical and Clinical Pretrained Models to French Long Documents: A Comparative Study 26 Feb 2024 · 1 repository · arXiv:2402.16689
-
An Integrated Data Processing Framework for Pretraining Foundation Models 26 Feb 2024 · 2 repositories · arXiv:2402.16358
-
Asymmetry in Low-Rank Adapters of Foundation Models 26 Feb 2024 · 1 repository · arXiv:2402.16842Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
ESG Sentiment Analysis: comparing human and language model performance including GPT 26 Feb 2024 · 0 repositories · arXiv:2402.16650
-
From RAGs to riches: Using large language models to write documents for clinical trials 26 Feb 2024 · 0 repositories · arXiv:2402.16406
-
Retrieval Augmented Generation Systems: Automatic Dataset Creation, Evaluation and Boolean Agent Setup 26 Feb 2024 · 1 repository · arXiv:2403.00820
-
ChatMusician: Understanding and Generating Music Intrinsically with LLM 25 Feb 2024 · 1 repository · arXiv:2402.16153Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Deep Learning Approaches for Improving Question Answering Systems in Hepatocellular Carcinoma Research 25 Feb 2024 · 0 repositories · arXiv:2402.16038
-
Emotion Classification in Short English Texts using Deep Learning Techniques 25 Feb 2024 · 0 repositories · arXiv:2402.16034
-
Knowledge Fusion of Chat LLMs: A Preliminary Technical Report 25 Feb 2024 · 2 repositories · arXiv:2402.16107
-
Hitting "Probe"rty with Non-Linearity, and More 25 Feb 2024 · 0 repositories · arXiv:2402.16168
-
LLMs Can Defend Themselves Against Jailbreaking in a Practical Manner: A Vision Paper 24 Feb 2024 · 0 repositories · arXiv:2402.15727