Methods › General › Regularization › Attention Dropout › Papers, page 47
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 47 of 109: papers 4,601 to 4,700 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RECAP: Retrieval-Augmented Audio Captioning 18 Sep 2023 · 1 repository · arXiv:2309.09836Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Towards Ontology Construction with Language Models 18 Sep 2023 · 0 repositories · arXiv:2309.09898
-
Contrastive Decoding Improves Reasoning in Large Language Models 17 Sep 2023 · 0 repositories · arXiv:2309.09117
-
Detecting covariate drift in text data using document embeddings and dimensionality reduction 17 Sep 2023 · 1 repository · arXiv:2309.10000
-
Do Large GPT Models Discover Moral Dimensions in Language Representations? A Topological Study Of Sentence Embeddings 17 Sep 2023 · 0 repositories · arXiv:2309.09397
-
From Cooking Recipes to Robot Task Trees -- Improving Planning Correctness and Task Efficiency by Leveraging LLMs with a Knowledge Network 17 Sep 2023 · 0 repositories · arXiv:2309.09181
-
Empowering In-Browser Deep Learning Inference on Edge Devices with Just-in-Time Kernel Optimizations 16 Sep 2023 · 0 repositories · arXiv:2309.08978
-
Decoder-only Architecture for Speech Recognition with CTC Prompts and Text Data Augmentation 16 Sep 2023 · 0 repositories · arXiv:2309.08876
-
Has Sentiment Returned to the Pre-pandemic Level? A Sentiment Analysis Using U.S. College Subreddit Data from 2019 to 2022 16 Sep 2023 · 1 repository · arXiv:2309.08845
-
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 16 Sep 2023 · 1 repository · arXiv:2309.08963
-
A Modern Turkish Poet: Fine-Tuned GPT-2 15 Sep 2023 · 1 repository
-
Advancing the Evaluation of Traditional Chinese Language Models: Towards a Comprehensive Benchmark Suite 15 Sep 2023 · 1 repository · arXiv:2309.08448
-
AlbNER: A Corpus for Named Entity Recognition in Albanian 15 Sep 2023 · 0 repositories · arXiv:2309.08741
-
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models 15 Sep 2023 · 1 repository · arXiv:2309.08573
-
EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers 15 Sep 2023 · 2 repositories · arXiv:2309.08532Syntology 11 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending 15 Sep 2023 · 1 repository · arXiv:2309.08646
-
Detecting Relevant Information in High-Volume Chat Logs: Keyphrase Extraction for Grooming and Drug Dealing Forensic Analysis 15 Sep 2023 · 0 repositories · arXiv:2311.04905
-
GPT-Lab: Next Generation Of Optimal Chemistry Discovery By GPT Driven Robotic Lab 15 Sep 2023 · 0 repositories · arXiv:2309.16721
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
Large Language Models for Failure Mode Classification: An Investigation 15 Sep 2023 · 1 repository · arXiv:2309.08181
-
Structural Self-Supervised Objectives for Transformers 15 Sep 2023 · 1 repository · arXiv:2309.08272
-
Transformer Based Punctuation Restoration for Turkish 15 Sep 2023 · 1 repository
-
VulnSense: Efficient Vulnerability Detection in Ethereum Smart Contracts by Multimodal Learning with Graph Neural Network and Language Model 15 Sep 2023 · 0 repositories · arXiv:2309.08474
-
An Empirical Evaluation of Prompting Strategies for Large Language Models in Zero-Shot Clinical Natural Language Processing 14 Sep 2023 · 0 repositories · arXiv:2309.08008
-
Assessing the nature of large language models: A caution against anthropocentrism 14 Sep 2023 · 0 repositories · arXiv:2309.07683
-
Automatic Data Visualization Generation from Chinese Natural Language Questions 14 Sep 2023 · 0 repositories · arXiv:2309.07650
-
ChatGPT MT: Competitive for High- (but not Low-) Resource Languages 14 Sep 2023 · 2 repositories · arXiv:2309.07423Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
DBLPLink: An Entity Linker for the DBLP Scholarly Knowledge Graph 14 Sep 2023 · 1 repository · arXiv:2309.07545
-
DebCSE: Rethinking Unsupervised Contrastive Sentence Embedding Learning in the Debiasing Perspective 14 Sep 2023 · 0 repositories · arXiv:2309.07396
-
EnCodecMAE: Leveraging neural codecs for universal audio representation learning 14 Sep 2023 · 2 repositories · arXiv:2309.07391Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Text Classification of Cancer Clinical Trial Eligibility Criteria 14 Sep 2023 · 0 repositories · arXiv:2309.07812
-
Two Timin': Repairing Smart Contracts With A Two-Layered Approach 14 Sep 2023 · 0 repositories · arXiv:2309.07841
-
Benchmarking Procedural Language Understanding for Low-Resource Languages: A Case Study on Turkish 13 Sep 2023 · 1 repository · arXiv:2309.06698
-
Large Language Models Can Infer Psychological Dispositions of Social Media Users 13 Sep 2023 · 0 repositories · arXiv:2309.08631
-
Traveling Words: A Geometric Interpretation of Transformers 13 Sep 2023 · 1 repository · arXiv:2309.07315
-
Balanced and Explainable Social Media Analysis for Public Health with Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05951
-
Circuit Breaking: Removing Model Behaviors with Targeted Ablation 12 Sep 2023 · 1 repository · arXiv:2309.05973Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
RAP-Gen: Retrieval-Augmented Patch Generation with CodeT5 for Automatic Program Repair 12 Sep 2023 · 0 repositories · arXiv:2309.06057
-
Characterizing Latent Perspectives of Media Houses Towards Public Figures 12 Sep 2023 · 0 repositories · arXiv:2309.06112
-
PRESTI: Predicting Repayment Effort of Self-Admitted Technical Debt Using Textual Information 12 Sep 2023 · 0 repositories · arXiv:2309.06020
-
Comparing Llama-2 and GPT-3 LLMs for HPC kernels generation 12 Sep 2023 · 0 repositories · arXiv:2309.07103
-
Exploring Large Language Models for Ontology Alignment 12 Sep 2023 · 1 repository · arXiv:2309.07172
-
Overview of Memotion 3: Sentiment and Emotion Analysis of Codemixed Hinglish Memes 12 Sep 2023 · 0 repositories · arXiv:2309.06517
-
Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing 12 Sep 2023 · 0 repositories · arXiv:2309.05898
-
The Moral Machine Experiment on Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05958
-
Unveiling the potential of large language models in generating semantic and cross-language clones 12 Sep 2023 · 0 repositories · arXiv:2309.06424
-
Applying BioBERT to Extract Germline Gene-Disease Associations for Building a Knowledge Graph from the Biomedical Literature 11 Sep 2023 · 1 repository · arXiv:2309.13061
-
Black-Box Analysis: GPTs Across Time in Legal Textual Entailment Task 11 Sep 2023 · 0 repositories · arXiv:2309.05501
-
CrisisTransformers: Pre-trained language models and sentence encoders for crisis-related social media texts 11 Sep 2023 · 0 repositories · arXiv:2309.05494
-
Detecting Natural Language Biases with Prompt-based Learning 11 Sep 2023 · 0 repositories · arXiv:2309.05227
-
Memory Injections: Correcting Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models 11 Sep 2023 · 1 repository · arXiv:2309.05605Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
SparseSwin: Swin Transformer with Sparse Transformer Block 11 Sep 2023 · 1 repository · arXiv:2309.05224
-
Zero-shot Learning with Minimum Instruction to Extract Social Determinants and Family History from Clinical Notes using GPT Model 11 Sep 2023 · 0 repositories · arXiv:2309.05475
-
Implementing Learning Principles with a Personal AI Tutor: A Case Study 10 Sep 2023 · 0 repositories · arXiv:2309.13060
-
Learning Personalized User Preference from Cold Start in Multi-turn Conversations 10 Sep 2023 · 0 repositories · arXiv:2309.05127
-
Neural-Hidden-CRF: A Robust Weakly-Supervised Sequence Labeler 10 Sep 2023 · 1 repository · arXiv:2309.05086
-
RGAT: A Deeper Look into Syntactic Dependency Information for Coreference Resolution 10 Sep 2023 · 0 repositories · arXiv:2309.04977
-
Can NLP Models 'Identify', 'Distinguish', and 'Justify' Questions that Don't have a Definitive Answer? 8 Sep 2023 · 0 repositories · arXiv:2309.04635
-
Context-Aware Prompt Tuning for Vision-Language Model with Dual-Alignment 8 Sep 2023 · 0 repositories · arXiv:2309.04158
-
Encoding Multi-Domain Scientific Papers by Ensembling Multiple CLS Tokens 8 Sep 2023 · 1 repository · arXiv:2309.04333Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Fuzzy Fingerprinting Transformer Language-Models for Emotion Recognition in Conversations 8 Sep 2023 · 0 repositories · arXiv:2309.04292
-
Leveraging Pretrained Image-text Models for Improving Audio-Visual Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04628
-
UQ at #SMM4H 2023: ALEX for Public Health Analysis with Social Media 8 Sep 2023 · 1 repository · arXiv:2309.04213
-
Evaluating ChatGPT as a Recommender System: A Rigorous Approach 7 Sep 2023 · 1 repository · arXiv:2309.03613
-
Supervised Learning and Large Language Model Benchmarks on Mental Health Datasets: Cognitive Distortions and Suicidal Risks in Chinese Social Media 7 Sep 2023 · 2 repositories · arXiv:2309.03564
-
FLM-101B: An Open LLM and How to Train It with $100K Budget 7 Sep 2023 · 0 repositories · arXiv:2309.03852
-
Zero-Shot Audio Captioning via Audibility Guidance 7 Sep 2023 · 0 repositories · arXiv:2309.03884
-
Certifying LLM Safety against Adversarial Prompting 6 Sep 2023 · 1 repository · arXiv:2309.02705
-
HAE-RAE Bench: Evaluation of Korean Knowledge in Language Models 6 Sep 2023 · 1 repository · arXiv:2309.02706
-
Leave no Place Behind: Improved Geolocation in Humanitarian Documents 6 Sep 2023 · 0 repositories · arXiv:2309.02914
-
Offensive Hebrew Corpus and Detection using BERT 6 Sep 2023 · 1 repository · arXiv:2309.02724
-
Self-Supervised Masked Digital Elevation Models Encoding for Low-Resource Downstream Tasks 6 Sep 2023 · 0 repositories · arXiv:2309.03367
-
CodeApex: A Bilingual Programming Evaluation Benchmark for Large Language Models 5 Sep 2023 · 1 repository · arXiv:2309.01940
-
Do You Trust ChatGPT? -- Perceived Credibility of Human and AI-Generated Content 5 Sep 2023 · 0 repositories · arXiv:2309.02524
-
Incorporating Dictionaries into a Neural Network Architecture to Extract COVID-19 Medical Concepts From Social Media 5 Sep 2023 · 0 repositories · arXiv:2309.02188
-
Language Models for Novelty Detection in System Call Traces 5 Sep 2023 · 0 repositories · arXiv:2309.02206
-
Leveraging BERT Language Models for Multi-Lingual ESG Issue Identification 5 Sep 2023 · 0 repositories · arXiv:2309.02189
-
nanoT5: A PyTorch Framework for Pre-training and Fine-tuning T5-style Models with Limited Resources 5 Sep 2023 · 1 repository · arXiv:2309.02373Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
Sample Size in Natural Language Processing within Healthcare Research 5 Sep 2023 · 0 repositories · arXiv:2309.02237
-
Benchmarking Large Language Models in Retrieval-Augmented Generation 4 Sep 2023 · 1 repository · arXiv:2309.01431Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Do androids dream of fictional references? A bibliographic dialogue with ChatGPT3.5 4 Sep 2023 · 0 repositories · arXiv:2312.00789
-
Prompting or Fine-tuning? A Comparative Study of Large Language Models for Taxonomy Construction 4 Sep 2023 · 1 repository · arXiv:2309.01715
-
A Study on the Implementation of Generative AI Services Using an Enterprise Data-Based LLM Application Architecture 3 Sep 2023 · 0 repositories · arXiv:2309.01105
-
A Visual Interpretation-Based Self-Improved Classification System Using Virtual Adversarial Training 3 Sep 2023 · 0 repositories · arXiv:2309.01196
-
Saturn: An Optimized Data System for Large Model Deep Learning Workloads 3 Sep 2023 · 1 repository · arXiv:2309.01226
-
Knowledge Graph Embeddings for Multi-Lingual Structured Representations of Radiology Reports 2 Sep 2023 · 0 repositories · arXiv:2309.00917
-
Studying the impacts of pre-training using ChatGPT-generated text on downstream tasks 2 Sep 2023 · 0 repositories · arXiv:2309.05668
-
BatchPrompt: Accomplish more with less 1 Sep 2023 · 1 repository · arXiv:2309.00384
-
Large Language Models for Semantic Monitoring of Corporate Disclosures: A Case Study on Korea's Top 50 KOSPI Companies 1 Sep 2023 · 0 repositories · arXiv:2309.00208
-
Publicly Shareable Clinical Large Language Model Built on Synthetic Clinical Notes 1 Sep 2023 · 1 repository · arXiv:2309.00237Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SortedNet: A Scalable and Generalized Framework for Training Modular Deep Neural Networks 1 Sep 2023 · 0 repositories · arXiv:2309.00255
-
Taken out of context: On measuring situational awareness in LLMs 1 Sep 2023 · 1 repository · arXiv:2309.00667Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Why do universal adversarial attacks work on large language models?: Geometry might be the answer 1 Sep 2023 · 0 repositories · arXiv:2309.00254
-
BioCoder: A Benchmark for Bioinformatics Code Generation with Large Language Models 31 Aug 2023 · 1 repository · arXiv:2308.16458Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can humans help BERT gain "confidence"? 31 Aug 2023 · 0 repositories · arXiv:2309.06580
-
DictaBERT: A State-of-the-Art BERT Suite for Modern Hebrew 31 Aug 2023 · 0 repositories · arXiv:2308.16687
-
GPT has become financially literate: Insights from financial literacy tests of GPT and a preliminary test of how people use it as a source of advice 31 Aug 2023 · 0 repositories · arXiv:2309.00649
-
Ladder-of-Thought: Using Knowledge as Steps to Elevate Stance Detection 31 Aug 2023 · 0 repositories · arXiv:2308.16763
-
Linking microblogging sentiments to stock price movement: An application of GPT-4 31 Aug 2023 · 0 repositories · arXiv:2308.16771
-
SARATHI: Efficient LLM Inference by Piggybacking Decodes with Chunked Prefills 31 Aug 2023 · 0 repositories · arXiv:2308.16369