Methods › General › Regularization › Attention Dropout › Papers, page 28
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 28 of 109: papers 2,701 to 2,800 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ViD-GPT: Introducing GPT-style Autoregressive Generation in Video Diffusion Models 16 Jun 2024 · 1 repository · arXiv:2406.10981
-
A Comprehensive Survey of Foundation Models in Medicine 15 Jun 2024 · 0 repositories · arXiv:2406.10729
-
Beyond Raw Videos: Understanding Edited Videos with Large Multimodal Model 15 Jun 2024 · 1 repository · arXiv:2406.10484
-
MINT: a Multi-modal Image and Narrative Text Dubbing Dataset for Foley Audio Content Planning and Generation 15 Jun 2024 · 1 repository · arXiv:2406.10591
-
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation 15 Jun 2024 · 0 repositories · arXiv:2406.10561
-
Bag of Lies: Robustness in Continuous Pre-training BERT 14 Jun 2024 · 0 repositories · arXiv:2406.09967
-
Exploring the Correlation between Human and Machine Evaluation of Simultaneous Speech Translation 14 Jun 2024 · 0 repositories · arXiv:2406.10091
-
HIRO: Hierarchical Information Retrieval Optimization 14 Jun 2024 · 1 repository · arXiv:2406.09979
-
LieRE: Generalizing Rotary Position Encodings 14 Jun 2024 · 1 repository · arXiv:2406.10322Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models 14 Jun 2024 · 1 repository · arXiv:2406.10130
-
Towards Efficient Pareto Set Approximation via Mixture of Experts Based Model Fusion 14 Jun 2024 · 1 repository · arXiv:2406.09770
-
A More Practical Approach to Machine Unlearning 13 Jun 2024 · 0 repositories · arXiv:2406.09391
-
Analyzing Gender Polarity in Short Social Media Texts with BERT: The Role of Emojis and Emoticons 13 Jun 2024 · 0 repositories · arXiv:2406.09573
-
BPE-knockout: Pruning Pre-existing BPE Tokenisers with Backwards-compatible Morphological Semi-supervision 13 Jun 2024 · 1 repository
-
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination 13 Jun 2024 · 0 repositories · arXiv:2406.08818
-
Optimizing Large Model Training through Overlapped Activation Recomputation 13 Jun 2024 · 0 repositories · arXiv:2406.08756
-
PC-LoRA: Low-Rank Adaptation for Progressive Model Compression with Knowledge Distillation 13 Jun 2024 · 0 repositories · arXiv:2406.09117
-
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models 13 Jun 2024 · 0 repositories · arXiv:2406.09519
-
Ad Auctions for LLMs via Retrieval Augmented Generation 12 Jun 2024 · 0 repositories · arXiv:2406.09459
-
Exploring Fact Memorization and Style Imitation in LLMs Using QLoRA: An Experimental Study and Quality Assessment Methods 12 Jun 2024 · 0 repositories · arXiv:2406.08582
-
FaithFill: Faithful Inpainting for Object Completion Using a Single Reference Image 12 Jun 2024 · 0 repositories · arXiv:2406.07865
-
Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification 12 Jun 2024 · 1 repository · arXiv:2406.08660Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
How well it works: Benchmarking performance of GPT models on medical natural language processing tasks 12 Jun 2024 · 0 repositories
-
Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests 12 Jun 2024 · 0 repositories · arXiv:2406.07794
-
Label-aware Hard Negative Sampling Strategies with Momentum Contrastive Learning for Implicit Hate Speech Detection 12 Jun 2024 · 1 repository · arXiv:2406.07886
-
Leveraging Large Language Models for Web Scraping 12 Jun 2024 · 0 repositories · arXiv:2406.08246
-
Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation 12 Jun 2024 · 0 repositories · arXiv:2406.08328
-
Tailoring Generative AI Chatbots for Multiethnic Communities in Disaster Preparedness Communication: Extending the CASA Paradigm 12 Jun 2024 · 1 repository · arXiv:2406.08411
-
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis 11 Jun 2024 · 0 repositories · arXiv:2406.10273
-
Bilingual Sexism Classification: Fine-Tuned XLM-RoBERTa and GPT-3.5 Few-Shot Learning 11 Jun 2024 · 0 repositories · arXiv:2406.07287
-
COVID-19 Twitter Sentiment Classification Using Hybrid Deep Learning Model Based on Grid Search Methodology 11 Jun 2024 · 0 repositories · arXiv:2406.10266
-
DR-RAG: Applying Dynamic Document Relevance to Retrieval-Augmented Generation for Question-Answering 11 Jun 2024 · 0 repositories · arXiv:2406.07348
-
Flextron: Many-in-One Flexible Large Language Model 11 Jun 2024 · 0 repositories · arXiv:2406.10260
-
Multi-objective Reinforcement learning from AI Feedback 11 Jun 2024 · 1 repository · arXiv:2406.07295
-
Multimodal Belief Prediction 11 Jun 2024 · 1 repository · arXiv:2406.07466
-
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language 11 Jun 2024 · 0 repositories · arXiv:2406.08519
-
Unused information in token probability distribution of generative LLM: improving LLM reading comprehension through calculation of expected values 11 Jun 2024 · 1 repository · arXiv:2406.10267
-
AGB-DE: A Corpus for the Automated Legal Assessment of Clauses in German Consumer Contracts 10 Jun 2024 · 1 repository · arXiv:2406.06809
-
Compute Better Spent: Replacing Dense Layers with Structured Matrices 10 Jun 2024 · 1 repository · arXiv:2406.06248Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
In-Context Learning and Fine-Tuning GPT for Argument Mining 10 Jun 2024 · 1 repository · arXiv:2406.06699
-
Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing 10 Jun 2024 · 0 repositories · arXiv:2406.06723
-
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching 10 Jun 2024 · 0 repositories · arXiv:2406.06799
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection 10 Jun 2024 · 0 repositories · arXiv:2406.06663
-
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs 10 Jun 2024 · 0 repositories · arXiv:2406.10251
-
UMBRELA: UMbrela is the (Open-Source Reproduction of the) Bing RELevance Assessor 10 Jun 2024 · 1 repository · arXiv:2406.06519
-
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation 9 Jun 2024 · 2 repositories · arXiv:2406.05654Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Hidden Holes: topological aspects of language models 9 Jun 2024 · 0 repositories · arXiv:2406.05798
-
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents 9 Jun 2024 · 0 repositories · arXiv:2406.05870
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering 9 Jun 2024 · 0 repositories · arXiv:2406.05845
-
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation 9 Jun 2024 · 1 repository · arXiv:2406.05794
-
Text2VP: Generative AI for Visual Programming and Parametric Modeling 9 Jun 2024 · 0 repositories · arXiv:2407.07732
-
Advancing Semantic Textual Similarity Modeling: A Regression Framework with Translated ReLU and Smooth K2 Loss 8 Jun 2024 · 2 repositories · arXiv:2406.05326
-
Concept Formation and Alignment in Language Models: Bridging Statistical Patterns in Latent Space to Concept Taxonomy 8 Jun 2024 · 0 repositories · arXiv:2406.05315
-
Critical Phase Transition in Large Language Models 8 Jun 2024 · 0 repositories · arXiv:2406.05335
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts 8 Jun 2024 · 0 repositories · arXiv:2406.05569
-
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature 8 Jun 2024 · 0 repositories · arXiv:2406.05431
-
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner 8 Jun 2024 · 0 repositories · arXiv:2406.05498
-
VP-LLM: Text-Driven 3D Volume Completion with Large Language Models through Patchification 8 Jun 2024 · 0 repositories · arXiv:2406.05543
-
BAMO at SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense 7 Jun 2024 · 1 repository · arXiv:2406.04947
-
BERTs are Generative In-Context Learners 7 Jun 2024 · 1 repository · arXiv:2406.04823Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 26 harvested samples)
-
Corpus Poisoning via Approximate Greedy Gradient Descent 7 Jun 2024 · 1 repository · arXiv:2406.05087
-
CRAG -- Comprehensive RAG Benchmark 7 Jun 2024 · 2 repositories · arXiv:2406.04744Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
DiNeR: a Large Realistic Dataset for Evaluating Compositional Generalization 7 Jun 2024 · 1 repository · arXiv:2406.04669
-
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents 7 Jun 2024 · 2 repositories · arXiv:2406.06613Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Large Generative Graph Models 7 Jun 2024 · 0 repositories · arXiv:2406.05109
-
Low-Resource Cross-Lingual Summarization through Few-Shot Learning with Large Language Models 7 Jun 2024 · 0 repositories · arXiv:2406.04630
-
Multi-Head RAG: Solving Multi-Aspect Problems with LLMs 7 Jun 2024 · 2 repositories · arXiv:2406.05085
-
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation 7 Jun 2024 · 1 repository · arXiv:2406.05213Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
VTrans: Accelerating Transformer Compression with Variational Information Bottleneck based Pruning 7 Jun 2024 · 0 repositories · arXiv:2406.05276
-
A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions 6 Jun 2024 · 0 repositories · arXiv:2406.03712
-
Do Language Models Understand Morality? Towards a Robust Detection of Moral Content 6 Jun 2024 · 1 repository · arXiv:2406.04143
-
Empirical Guidelines for Deploying LLMs onto Resource-constrained Edge Devices 6 Jun 2024 · 0 repositories · arXiv:2406.03777
-
HORAE: A Domain-Agnostic Language for Automated Service Regulation 6 Jun 2024 · 1 repository · arXiv:2406.06600
-
LLMEmbed: Rethinking Lightweight LLM's Genuine Function in Text Classification 6 Jun 2024 · 1 repository · arXiv:2406.03725
-
Simplified and Generalized Masked Diffusion for Discrete Data 6 Jun 2024 · 1 repository · arXiv:2406.04329Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 15 unverified (of 21 harvested samples)
-
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech 6 Jun 2024 · 1 repository · arXiv:2406.03953
-
Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data 6 Jun 2024 · 2 repositories · arXiv:2406.03736Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Automating Turkish Educational Quiz Generation Using Large Language Models 5 Jun 2024 · 4 repositories · arXiv:2406.03397
-
Exact Conversion of In-Context Learning to Model Weights in Linearized-Attention Transformers 5 Jun 2024 · 0 repositories · arXiv:2406.02847
-
Exploring Multilingual Large Language Models for Enhanced TNM classification of Radiology Report in lung cancer staging 5 Jun 2024 · 0 repositories · arXiv:2406.06591
-
Missci: Reconstructing Fallacies in Misrepresented Science 5 Jun 2024 · 2 repositories · arXiv:2406.03181
-
PoLYTC: a novel BERT-based classifier to detect political leaning of YouTube videos based on their titles 5 Jun 2024 · 1 repository
-
RICo: Reddit ideological communities 5 Jun 2024 · 1 repository
-
StatBot.Swiss: Bilingual Open Data Exploration in Natural Language 5 Jun 2024 · 0 repositories · arXiv:2406.03170
-
The Good, the Bad, and the Hulk-like GPT: Analyzing Emotional Decisions of Large Language Models in Cooperation and Bargaining Games 5 Jun 2024 · 0 repositories · arXiv:2406.03299
-
Too Big to Fail: Larger Language Models are Disproportionately Resilient to Induction of Dementia-Related Linguistic Anomalies 5 Jun 2024 · 1 repository · arXiv:2406.02830
-
Chain of Agents: Large Language Models Collaborating on Long-Context Tasks 4 Jun 2024 · 0 repositories · arXiv:2406.02818
-
CheckEmbed: Effective Verification of LLM Solutions to Open-Ended Tasks 4 Jun 2024 · 1 repository · arXiv:2406.02524
-
Large Language Model-Enabled Multi-Agent Manufacturing Systems 4 Jun 2024 · 0 repositories · arXiv:2406.01893
-
OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step 4 Jun 2024 · 0 repositories · arXiv:2406.06576
-
Probing the Category of Verbal Aspect in Transformer Language Models 4 Jun 2024 · 0 repositories · arXiv:2406.02335
-
Randomized Geometric Algebra Methods for Convex Neural Networks 4 Jun 2024 · 1 repository · arXiv:2406.02806Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SMS Spam Detection and Classification to Combat Abuse in Telephone Networks Using Natural Language Processing 4 Jun 2024 · 0 repositories · arXiv:2406.06578
-
Synergetic Event Understanding: A Collaborative Approach to Cross-Document Event Coreference Resolution with Large Language Models 4 Jun 2024 · 1 repository · arXiv:2406.02148
-
Towards Effective Time-Aware Language Representation: Exploring Enhanced Temporal Understanding in Language Models 4 Jun 2024 · 0 repositories · arXiv:2406.01863
-
Annotation Guidelines-Based Knowledge Augmentation: Towards Enhancing Large Language Models for Educational Text Classification 3 Jun 2024 · 0 repositories · arXiv:2406.00954
-
Ask-EDA: A Design Assistant Empowered by LLM, Hybrid RAG and Abbreviation De-hallucination 3 Jun 2024 · 0 repositories · arXiv:2406.06575
-
BadRAG: Identifying Vulnerabilities in Retrieval Augmented Generation of Large Language Models 3 Jun 2024 · 0 repositories · arXiv:2406.00083
-
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP 3 Jun 2024 · 1 repository · arXiv:2406.01583Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
FactGenius: Combining Zero-Shot Prompting and Fuzzy Relation Mining to Improve Fact Verification with Knowledge Graphs 3 Jun 2024 · 1 repository · arXiv:2406.01311