Methods › General › Regularization › Attention Dropout › Papers, page 40
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 40 of 109: papers 3,901 to 4,000 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Adapt or Perish: Adaptive Sparse Transformer with Attentive Feature Refinement for Image Restoration 1 Jan 2024 · 1 repository
-
Large Language Models aren't all that you need 1 Jan 2024 · 0 repositories · arXiv:2401.00698
-
Large Language Models in Mental Health Care: a Scoping Review 1 Jan 2024 · 0 repositories · arXiv:2401.02984
-
Revisiting Counterfactual Problems in Referring Expression Comprehension 1 Jan 2024 · 1 repository
-
SEED-Bench: Benchmarking Multimodal Large Language Models 1 Jan 2024 · 1 repository
-
An Analysis of Embedding Layers and Similarity Scores using Siamese Neural Networks 31 Dec 2023 · 0 repositories · arXiv:2401.00582
-
RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models 31 Dec 2023 · 3 repositories · arXiv:2401.00396Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation 30 Dec 2023 · 1 repository · arXiv:2401.00280
-
Trace and Edit Relation Associations in GPT 30 Dec 2023 · 0 repositories · arXiv:2401.02976
-
Why is the User Interface a Dark Pattern? : Explainable Auto-Detection and its Analysis 30 Dec 2023 · 1 repository · arXiv:2401.04119
-
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models 29 Dec 2023 · 1 repository · arXiv:2312.17661
-
Jatmo: Prompt Injection Defense by Task-Specific Finetuning 29 Dec 2023 · 1 repository · arXiv:2312.17673Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
MosaicBERT: A Bidirectional Encoder Optimized for Fast Pretraining 29 Dec 2023 · 1 repository · arXiv:2312.17482
-
TuPy-E: detecting hate speech in Brazilian Portuguese social media with a novel dataset and comprehensive analysis of models 29 Dec 2023 · 1 repository · arXiv:2312.17704
-
Evaluating the Performance of Large Language Models for Spanish Language in Undergraduate Admissions Exams 28 Dec 2023 · 0 repositories · arXiv:2312.16845
-
Language Model as an Annotator: Unsupervised Context-aware Quality Phrase Generation 28 Dec 2023 · 0 repositories · arXiv:2312.17349
-
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model 28 Dec 2023 · 0 repositories · arXiv:2312.17122
-
ROI-Aware Multiscale Cross-Attention Vision Transformer for Pest Image Identification 28 Dec 2023 · 0 repositories · arXiv:2312.16914
-
SentinelLMs: Encrypted Input Adaptation and Fine-tuning of Language Models for Private and Secure Inference 28 Dec 2023 · 1 repository · arXiv:2312.17342
-
PanGu-π: Enhancing Language Model Architectures via Nonlinearity Compensation 27 Dec 2023 · 0 repositories · arXiv:2312.17276
-
Relationship between auditory and semantic entrainment using Deep Neural Networks (DNN) 27 Dec 2023 · 0 repositories · arXiv:2312.16599
-
ChartBench: A Benchmark for Complex Visual Reasoning in Charts 26 Dec 2023 · 0 repositories · arXiv:2312.15915
-
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4 26 Dec 2023 · 2 repositories · arXiv:2312.16171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Scaling Down, LiTting Up: Efficient Zero-Shot Listwise Reranking with Seq2seq Encoder-Decoder Models 26 Dec 2023 · 2 repositories · arXiv:2312.16098
-
SecQA: A Concise Question-Answering Dataset for Evaluating Large Language Models in Computer Security 26 Dec 2023 · 1 repository · arXiv:2312.15838Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Task Contamination: Language Models May Not Be Few-Shot Anymore 26 Dec 2023 · 0 repositories · arXiv:2312.16337
-
Compositional Generalization in Spoken Language Understanding 25 Dec 2023 · 0 repositories · arXiv:2312.15815
-
Fairness-Aware Structured Pruning in Transformers 24 Dec 2023 · 1 repository · arXiv:2312.15398Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Multi-level biomedical NER through multi-granularity embeddings and enhanced labeling 24 Dec 2023 · 0 repositories · arXiv:2312.15550
-
Do LLM Agents Exhibit Social Behavior? 23 Dec 2023 · 0 repositories · arXiv:2312.15198
-
Understanding the Potential of FPGA-Based Spatial Acceleration for Large Language Model Inference 23 Dec 2023 · 1 repository · arXiv:2312.15159
-
Efficacy of Machine-Generated Instructions 22 Dec 2023 · 0 repositories · arXiv:2312.14423
-
FM-OV3D: Foundation Model-based Cross-modal Knowledge Blending for Open-Vocabulary 3D Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14465
-
Numerical Reasoning for Financial Reports 22 Dec 2023 · 1 repository · arXiv:2312.14870
-
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection 22 Dec 2023 · 1 repository · arXiv:2312.15068
-
Theory of Hallucinations based on Equivariance 22 Dec 2023 · 0 repositories · arXiv:2312.14504
-
Towards Detecting Cascades of Biased Medical Claims on Twitter 22 Dec 2023 · 0 repositories · arXiv:2312.15040
-
How Smooth Is Attention? 22 Dec 2023 · 0 repositories · arXiv:2312.14820
-
Unsupervised Auditory and Semantic Entrainment Models with Deep Neural Networks 22 Dec 2023 · 0 repositories · arXiv:2312.15098
-
Argue with Me Tersely: Towards Sentence-Level Counter-Argument Generation 21 Dec 2023 · 1 repository · arXiv:2312.13608
-
ChatGPT as a commenter to the news: can LLMs generate human-like opinions? 21 Dec 2023 · 1 repository · arXiv:2312.13961
-
De novo Drug Design using Reinforcement Learning with Multiple GPT Agents 21 Dec 2023 · 2 repositories · arXiv:2401.06155Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
How to Prune Your Language Model: Recovering Accuracy on the "Sparsity May Cry'' Benchmark 21 Dec 2023 · 0 repositories · arXiv:2312.13547
-
InfoVisDial: An Informative Visual Dialogue Dataset by Bridging Large Multimodal and Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13503
-
TraceFL: Interpretability-Driven Debugging in Federated Learning via Neuron Provenance 21 Dec 2023 · 2 repositories · arXiv:2312.13632
-
Team Irisapu Project Description for DRC2023 21 Dec 2023 · 0 repositories · arXiv:2312.13765
-
Typhoon: Thai Large Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13951
-
AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation 20 Dec 2023 · 1 repository · arXiv:2312.13010Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Benchmarking and Analyzing In-context Learning, Fine-tuning and Supervised Learning for Biomedical Knowledge Curation: a focused study on chemical entities of biological interest 20 Dec 2023 · 0 repositories · arXiv:2312.12989
-
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks 20 Dec 2023 · 3 repositories · arXiv:2312.13322
-
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy 20 Dec 2023 · 1 repository · arXiv:2312.12728Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
An Empirical study of Unsupervised Neural Machine Translation: analyzing NMT output, model's behavior and sentences' contribution 19 Dec 2023 · 0 repositories · arXiv:2312.12588
-
Can ChatGPT be Your Personal Medical Assistant? 19 Dec 2023 · 0 repositories · arXiv:2312.12006
-
Can Transformers Learn Sequential Function Classes In Context? 19 Dec 2023 · 0 repositories · arXiv:2312.12655
-
Self-Admitted Technical Debt Detection Approaches: A Decade Systematic Review 19 Dec 2023 · 1 repository · arXiv:2312.15020
-
Efficient Title Reranker for Fast and Improved Knowledge-Intense NLP 19 Dec 2023 · 0 repositories · arXiv:2312.12430
-
Large Language Models in Medical Term Classification and Unexpected Misalignment Between Response and Reasoning 19 Dec 2023 · 0 repositories · arXiv:2312.14184
-
An In-depth Look at Gemini's Language Abilities 18 Dec 2023 · 1 repository · arXiv:2312.11444Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Aspect-Based Sentiment Analysis with Explicit Sentiment Augmentations 18 Dec 2023 · 0 repositories · arXiv:2312.10961
-
Generative linguistic representation for spoken language identification 18 Dec 2023 · 0 repositories · arXiv:2312.10964
-
"Knowing When You Don't Know": A Multilingual Relevance Assessment Dataset for Robust Retrieval-Augmented Generation 18 Dec 2023 · 1 repository · arXiv:2312.11361Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Retrieval-Augmented Generation for Large Language Models: A Survey 18 Dec 2023 · 4 repositories · arXiv:2312.10997
-
Artificial intelligence optical hardware empowers high-resolution hyperspectral video understanding at 1.2 Tb/s 17 Dec 2023 · 0 repositories · arXiv:2312.10639
-
Bengali Intent Classification with Generative Adversarial BERT 17 Dec 2023 · 1 repository · arXiv:2312.10679
-
Can persistent homology whiten Transformer-based black-box models? A case study on BERT compression 17 Dec 2023 · 0 repositories · arXiv:2312.10702
-
Decoding Concerns: Multi-label Classification of Vaccine Sentiments in Social Media 17 Dec 2023 · 1 repository · arXiv:2312.10626
-
Evaluating AI Vocational Skills Through Professional Testing 17 Dec 2023 · 0 repositories · arXiv:2312.10603
-
HyperPIE: Hyperparameter Information Extraction from Scientific Publications 17 Dec 2023 · 1 repository · arXiv:2312.10638
-
Investigating salient representations and label Variance in Dimensional Speech Emotion Analysis 17 Dec 2023 · 0 repositories · arXiv:2312.16180
-
Mixed Distillation Helps Smaller Language Model Better Reasoning 17 Dec 2023 · 0 repositories · arXiv:2312.10730
-
Multi-Label Classification of COVID-Tweets Using Large Language Models 17 Dec 2023 · 1 repository · arXiv:2312.10748
-
T2M-HiFiGPT: Generating High Quality Human Motion from Textual Descriptions with Residual Discrete Representations 17 Dec 2023 · 0 repositories · arXiv:2312.10628
-
A Comparative Analysis of Large Language Models for Code Documentation Generation 16 Dec 2023 · 0 repositories · arXiv:2312.10349
-
Cross-Linguistic Offensive Language Detection: BERT-Based Analysis of Bengali, Assamese, & Bodo Conversational Hateful Content from Social Media 16 Dec 2023 · 0 repositories · arXiv:2312.10528
-
Investigating Shallow and Deep Learning Techniques for Emotion Classification in Short Persian Texts 16 Dec 2023 · 1 repository
-
SPT: Fine-Tuning Transformer-based Language Models Efficiently with Sparsification 16 Dec 2023 · 1 repository · arXiv:2312.10365
-
A Novel Dataset for Financial Education Text Simplification in Spanish 15 Dec 2023 · 0 repositories · arXiv:2312.09897
-
Algorithms for automatic intents extraction and utterances classification for goal-oriented dialogue systems 15 Dec 2023 · 0 repositories · arXiv:2312.09658
-
Distilling Large Language Models for Matching Patients to Clinical Trials 15 Dec 2023 · 0 repositories · arXiv:2312.09958
-
Exploring Automatic Text Simplification of German Narrative Documents 15 Dec 2023 · 1 repository · arXiv:2312.09907
-
Exploring Multi-Level Threats in Telegram Data with AI-Human Annotation: A Preliminary Study 15 Dec 2023 · 0 repositories
-
No-Skim: Towards Efficiency Robustness Evaluation on Skimming-based Language Models 15 Dec 2023 · 0 repositories · arXiv:2312.09494
-
Red AI? Inconsistent Responses from GPT3.5 Models on Political Issues in the US and China 15 Dec 2023 · 0 repositories · arXiv:2312.09917
-
Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning 14 Dec 2023 · 0 repositories · arXiv:2312.08901
-
Dissecting vocabulary biases datasets through statistical testing and automated data augmentation for artifact mitigation in Natural Language Inference 14 Dec 2023 · 1 repository · arXiv:2312.08747
-
Dynamic Retrieval-Augmented Generation 14 Dec 2023 · 0 repositories · arXiv:2312.08976
-
Inter-Layer Scheduling Space Exploration for Multi-model Inference on Heterogeneous Chiplets 14 Dec 2023 · 0 repositories · arXiv:2312.09401
-
Motion Flow Matching for Human Motion Synthesis and Editing 14 Dec 2023 · 0 repositories · arXiv:2312.08895
-
Self-Evaluation Improves Selective Generation in Large Language Models 14 Dec 2023 · 0 repositories · arXiv:2312.09300
-
Successor Heads: Recurring, Interpretable Attention Heads In The Wild 14 Dec 2023 · 0 repositories · arXiv:2312.09230
-
TinyGSM: achieving >80% on GSM8k with small language models 14 Dec 2023 · 0 repositories · arXiv:2312.09241
-
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision 14 Dec 2023 · 0 repositories · arXiv:2312.09390
-
Weaving Pathways for Justice with GPT: LLM-driven automated drafting of interactive legal applications 14 Dec 2023 · 1 repository · arXiv:2312.09198
-
Causality Analysis for Evaluating the Security of Large Language Models 13 Dec 2023 · 1 repository · arXiv:2312.07876Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Large Language Models are Complex Table Parsers 13 Dec 2023 · 0 repositories · arXiv:2312.11521
-
Mono3DVG: 3D Visual Grounding in Monocular Images 13 Dec 2023 · 1 repository · arXiv:2312.08022Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Native Language Identification with Large Language Models 13 Dec 2023 · 0 repositories · arXiv:2312.07819
-
Prompt Engineering-assisted Malware Dynamic Analysis Using GPT-4 13 Dec 2023 · 1 repository · arXiv:2312.08317
-
AI Control: Improving Safety Despite Intentional Subversion 12 Dec 2023 · 1 repository · arXiv:2312.06942Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Exploring Large Language Models to Facilitate Variable Autonomy for Human-Robot Teaming 12 Dec 2023 · 0 repositories · arXiv:2312.07214