Methods › General › Regularization › Attention Dropout › Papers, page 50
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 50 of 109: papers 4,901 to 5,000 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
How is ChatGPT's behavior changing over time? 18 Jul 2023 · 4 repositories · arXiv:2307.09009Syntology official (archive's flag): 1 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 5 pointer-only (licence)
-
KATIE: A System for Key Attributes Identification in Product Knowledge Graph Construction 18 Jul 2023 · 0 repositories
-
Unveiling Gender Bias in Terms of Profession Across LLMs: Analyzing and Addressing Sociological Implications 18 Jul 2023 · 0 repositories · arXiv:2307.09162
-
A mixed policy to improve performance of language models on math problems 17 Jul 2023 · 1 repository · arXiv:2307.08767
-
A Study on the Performance of Generative Pre-trained Transformer (GPT) in Simulating Depressed Individuals on the Standardized Depressive Symptom Scale 17 Jul 2023 · 0 repositories · arXiv:2307.08576
-
ChatGPT is Good but Bing Chat is Better for Vietnamese Students 17 Jul 2023 · 0 repositories · arXiv:2307.08272
-
GEAR: Augmenting Language Models with Generalizable and Efficient Tool Resolution 17 Jul 2023 · 1 repository · arXiv:2307.08775Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Using an LLM to Help With Code Understanding 17 Jul 2023 · 0 repositories · arXiv:2307.08177
-
Legal Syllogism Prompting: Teaching Large Language Models for Legal Judgment Prediction 17 Jul 2023 · 1 repository · arXiv:2307.08321
-
Cross-Lingual NER for Financial Transaction Data in Low-Resource Languages 16 Jul 2023 · 0 repositories · arXiv:2307.08714
-
Domain Generalisation with Bidirectional Encoder Representations from Vision Transformers 16 Jul 2023 · 0 repositories · arXiv:2307.08117
-
Recognition of Mental Adjectives in An Efficient and Automatic Style 16 Jul 2023 · 0 repositories · arXiv:2307.11767
-
SentimentGPT: Exploiting GPT for Advanced Sentiment Analysis and its Departure from Current Machine Learning 16 Jul 2023 · 1 repository · arXiv:2307.10234
-
Coupling Large Language Models with Logic Programming for Robust and General Reasoning from Text 15 Jul 2023 · 1 repository · arXiv:2307.07696Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Large Language Models as Superpositions of Cultural Perspectives 15 Jul 2023 · 0 repositories · arXiv:2307.07870
-
Leveraging Large Language Models to Generate Answer Set Programs 15 Jul 2023 · 1 repository · arXiv:2307.07699Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Do not Mask Randomly: Effective Domain-adaptive Pre-training by Masking In-domain Keywords 14 Jul 2023 · 0 repositories · arXiv:2307.07160
-
Large Language Models Understand and Can be Enhanced by Emotional Stimuli 14 Jul 2023 · 0 repositories · arXiv:2307.11760
-
Fairness of ChatGPT and the Role Of Explainable-Guided Prompts 14 Jul 2023 · 1 repository · arXiv:2307.11761
-
Improving BERT with Hybrid Pooling Network and Drop Mask 14 Jul 2023 · 0 repositories · arXiv:2307.07258
-
MorphPiece : A Linguistic Tokenizer for Large Language Models 14 Jul 2023 · 0 repositories · arXiv:2307.07262
-
Sensi-BERT: Towards Sensitivity Driven Fine-Tuning for Parameter-Efficient BERT 14 Jul 2023 · 0 repositories · arXiv:2307.11764
-
Towards spoken dialect identification of Irish 14 Jul 2023 · 0 repositories · arXiv:2307.07436
-
TVPR: Text-to-Video Person Retrieval and a New Benchmark 14 Jul 2023 · 0 repositories · arXiv:2307.07184
-
A Study on Differentiable Logic and LLMs for EPIC-KITCHENS-100 Unsupervised Domain Adaptation Challenge for Action Recognition 2023 13 Jul 2023 · 0 repositories · arXiv:2307.06569
-
Agreement Tracking for Multi-Issue Negotiation Dialogues 13 Jul 2023 · 0 repositories · arXiv:2307.06524
-
Convolutional Neural Networks for Sentiment Analysis on Weibo Data: A Natural Language Processing Approach 13 Jul 2023 · 0 repositories · arXiv:2307.06540
-
Negated Complementary Commonsense using Large Language Models 13 Jul 2023 · 1 repository · arXiv:2307.06794
-
SecureFalcon: Are We There Yet in Automated Software Vulnerability Detection with LLMs? 13 Jul 2023 · 0 repositories · arXiv:2307.06616
-
Tackling Fake News in Bengali: Unraveling the Impact of Summarization vs. Augmentation on Pre-trained Language Models 13 Jul 2023 · 1 repository · arXiv:2307.06979
-
Retrieval Augmented Generation using Engineering Design Knowledge 13 Jul 2023 · 2 repositories · arXiv:2307.06985
-
Ashaar: Automatic Analysis and Generation of Arabic Poetry Using Deep Learning Approaches 12 Jul 2023 · 1 repository · arXiv:2307.06218
-
Detecting the Presence of COVID-19 Vaccination Hesitancy from South African Twitter Data Using Machine Learning 12 Jul 2023 · 0 repositories · arXiv:2307.15072
-
Distilling Large Language Models for Biomedical Knowledge Extraction: A Case Study on Adverse Drug Events 12 Jul 2023 · 0 repositories · arXiv:2307.06439
-
No Train No Gain: Revisiting Efficient Training Algorithms For Transformer-based Language Models 12 Jul 2023 · 1 repository · arXiv:2307.06440Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 3 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Prompt Generate Train (PGT): Few-shot Domain Adaption of Retrieval Augmented Generation Models for Open Book Question-Answering 12 Jul 2023 · 0 repositories · arXiv:2307.05915
-
Argumentative Segmentation Enhancement for Legal Summarization 11 Jul 2023 · 0 repositories · arXiv:2307.05081
-
DNAGPT: A Generalized Pre-trained Tool for Versatile DNA Sequence Analysis Tasks 11 Jul 2023 · 0 repositories · arXiv:2307.05628
-
Large Language Models 11 Jul 2023 · 0 repositories · arXiv:2307.05782
-
Named entity recognition using GPT for identifying comparable companies 11 Jul 2023 · 0 repositories · arXiv:2307.07420
-
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration 11 Jul 2023 · 2 repositories · arXiv:2307.05300
-
Vacaspati: A Diverse Corpus of Bangla Literature 11 Jul 2023 · 0 repositories · arXiv:2307.05083
-
AmadeusGPT: a natural language interface for interactive animal behavioral analysis 10 Jul 2023 · 1 repository · arXiv:2307.04858
-
ChatGPT for Digital Forensic Investigation: The Good, The Bad, and The Unknown 10 Jul 2023 · 1 repository · arXiv:2307.10195
-
SimpleMTOD: A Simple Language Model for Multimodal Task-Oriented Dialogue with Symbolic Scene Representation 10 Jul 2023 · 0 repositories · arXiv:2307.04907
-
Assessing the efficacy of large language models in generating accurate teacher responses 9 Jul 2023 · 0 repositories · arXiv:2307.04274
-
A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation 8 Jul 2023 · 0 repositories · arXiv:2307.03987
-
Is ChatGPT a Good Personality Recognizer? A Preliminary Study 8 Jul 2023 · 0 repositories · arXiv:2307.03952
-
DWReCO at CheckThat! 2023: Enhancing Subjectivity Detection through Style-based Data Sampling 7 Jul 2023 · 1 repository · arXiv:2307.03550
-
Goal-Conditioned Predictive Coding for Offline Reinforcement Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03406
-
How does AI chat change search behaviors? 7 Jul 2023 · 0 repositories · arXiv:2307.03826
-
RADAR: Robust AI-Text Detection via Adversarial Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03838
-
Text Simplification of Scientific Texts for Non-Expert Readers 7 Jul 2023 · 0 repositories · arXiv:2307.03569
-
TRAQ: Trustworthy Retrieval Augmented Question Answering via Conformal Prediction 7 Jul 2023 · 1 repository · arXiv:2307.04642
-
A Novel Site-Agnostic Multimodal Deep Learning Model to Identify Pro-Eating Disorder Content on Social Media 6 Jul 2023 · 0 repositories · arXiv:2307.06775
-
Can ChatGPT's Responses Boost Traditional Natural Language Processing? 6 Jul 2023 · 1 repository · arXiv:2307.04648
-
Improving Retrieval-Augmented Large Language Models via Data Importance Learning 6 Jul 2023 · 1 repository · arXiv:2307.03027Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Large Language Models Empowered Autonomous Edge AI for Connected Intelligence 6 Jul 2023 · 0 repositories · arXiv:2307.02779
-
Text Alignment Is An Efficient Unified Model for Massive NLP Tasks 6 Jul 2023 · 1 repository · arXiv:2307.02729
-
UIT-Saviors at MEDVQA-GI 2023: Improving Multimodal Learning with Image Enhancement for Gastrointestinal Visual Question Answering 6 Jul 2023 · 0 repositories · arXiv:2307.02783
-
CAME: Confidence-guided Adaptive Memory Efficient Optimization 5 Jul 2023 · 2 repositories · arXiv:2307.02047Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Emoji Prediction in Tweets using BERT 5 Jul 2023 · 1 repository · arXiv:2307.02054
-
Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations 5 Jul 2023 · 0 repositories · arXiv:2307.03678
-
Exploring Continual Learning for Code Generation Models 5 Jul 2023 · 0 repositories · arXiv:2307.02435
-
External Reasoning: Towards Multi-Large-Language-Models Interchangeable Assistance with Human Feedback 5 Jul 2023 · 1 repository · arXiv:2307.12057
-
Hoodwinked: Deception and Cooperation in a Text-Based Game for Language Models 5 Jul 2023 · 1 repository · arXiv:2308.01404
-
Leveraging Denoised Abstract Meaning Representation for Grammatical Error Correction 5 Jul 2023 · 0 repositories · arXiv:2307.02127
-
Multilingual Controllable Transformer-Based Lexical Simplification 5 Jul 2023 · 1 repository · arXiv:2307.02120
-
Named Entity Inclusion in Abstractive Text Summarization 5 Jul 2023 · 0 repositories · arXiv:2307.02570
-
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning 5 Jul 2023 · 0 repositories · arXiv:2307.02179
-
The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification 5 Jul 2023 · 0 repositories · arXiv:2307.02192
-
Embodied Task Planning with Large Language Models 4 Jul 2023 · 1 repository · arXiv:2307.01848Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
KDSTM: Neural Semi-supervised Topic Modeling with Knowledge Distillation 4 Jul 2023 · 0 repositories · arXiv:2307.01878Syntology 6 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
ALBERTI, a Multilingual Domain Specific Language Model for Poetry Analysis 3 Jul 2023 · 0 repositories · arXiv:2307.01387
-
Improving Language Plasticity via Pretraining with Active Forgetting 3 Jul 2023 · 1 repository · arXiv:2307.01163
-
Interpretability and Transparency-Driven Detection and Transformation of Textual Adversarial Examples (IT-DT) 3 Jul 2023 · 0 repositories · arXiv:2307.01225
-
Iterative Zero-Shot LLM Prompting for Knowledge Graph Construction 3 Jul 2023 · 0 repositories · arXiv:2307.01128
-
TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition 2 Jul 2023 · 0 repositories · arXiv:2307.00526
-
How far is Language Model from 100% Few-shot Named Entity Recognition in Medical Domain 1 Jul 2023 · 1 repository · arXiv:2307.00186
-
Large Language Models (GPT) for automating feedback on programming assignments 30 Jun 2023 · 0 repositories · arXiv:2307.00150
-
Meta-Reasoning: Semantics-Symbol Deconstruction for Large Language Models 30 Jun 2023 · 1 repository · arXiv:2306.17820
-
SPAE: Semantic Pyramid AutoEncoder for Multimodal Generation with Frozen LLMs 30 Jun 2023 · 0 repositories · arXiv:2306.17842
-
Stay on topic with Classifier-Free Guidance 30 Jun 2023 · 0 repositories · arXiv:2306.17806
-
Ticket-BERT: Labeling Incident Management Tickets with Language Models 30 Jun 2023 · 0 repositories · arXiv:2307.00108
-
A negation detection assessment of GPTs: analysis with the xNot360 dataset 29 Jun 2023 · 0 repositories · arXiv:2306.16638
-
Benchmarking Large Language Model Capabilities for Conditional Generation 29 Jun 2023 · 0 repositories · arXiv:2306.16793
-
BinaryViT: Pushing Binary Vision Transformers Towards Convolutional Models 29 Jun 2023 · 1 repository · arXiv:2306.16678
-
Classifying Crime Types using Judgment Documents from Social Media 29 Jun 2023 · 0 repositories · arXiv:2306.17020
-
Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors 29 Jun 2023 · 0 repositories · arXiv:2306.17156
-
Harnessing the Power of Hugging Face Transformers for Predicting Mental Health Disorders in Social Networks 29 Jun 2023 · 0 repositories · arXiv:2306.16891
-
An Efficient Sparse Inference Software Accelerator for Transformer-based Language Models on CPUs 28 Jun 2023 · 1 repository · arXiv:2306.16601
-
Pareto Optimal Learning for Estimating Large Language Model Errors 28 Jun 2023 · 0 repositories · arXiv:2306.16564
-
Beyond the Hype: Assessing the Performance, Trustworthiness, and Clinical Suitability of GPT3.5 28 Jun 2023 · 0 repositories · arXiv:2306.15887
-
Inferring the Goals of Communicating Agents from Actions and Instructions 28 Jun 2023 · 0 repositories · arXiv:2306.16207
-
Is ChatGPT a Biomedical Expert? -- Exploring the Zero-Shot Performance of Current GPT Models in Biomedical Tasks 28 Jun 2023 · 1 repository · arXiv:2306.16108
-
Multi-Site Clinical Federated Learning using Recursive and Attentive Models and NVFlare 28 Jun 2023 · 0 repositories · arXiv:2306.16367
-
Taqyim: Evaluating Arabic NLP Tasks Using ChatGPT Models 28 Jun 2023 · 1 repository · arXiv:2306.16322
-
Evaluating GPT-3.5 and GPT-4 on Grammatical Error Correction for Brazilian Portuguese 27 Jun 2023 · 0 repositories · arXiv:2306.15788
-
Gender Bias in BERT -- Measuring and Analysing Biases through Sentiment Rating in a Realistic Downstream Classification Task 27 Jun 2023 · 0 repositories · arXiv:2306.15298
-
Investigating Cross-Domain Behaviors of BERT in Review Understanding 27 Jun 2023 · 0 repositories · arXiv:2306.15123