Methods › General › Regularization › Attention Dropout › Papers, page 45
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 45 of 109: papers 4,401 to 4,500 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Probing the Creativity of Large Language Models: Can models produce divergent semantic association? 17 Oct 2023 · 1 repository · arXiv:2310.11158Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
USDC: Unified Static and Dynamic Compression for Visual Transformer 17 Oct 2023 · 0 repositories · arXiv:2310.11117
-
Utilising a Large Language Model to Annotate Subject Metadata: A Case Study in an Australian National Research Data Catalogue 17 Oct 2023 · 0 repositories · arXiv:2310.11318
-
Battle of the Large Language Models: Dolly vs LLaMA vs Vicuna vs Guanaco vs Bard vs ChatGPT -- A Text-to-SQL Parsing Comparison 16 Oct 2023 · 0 repositories · arXiv:2310.10190
-
BioPlanner: Automatic Evaluation of LLMs on Protocol Planning in Biology 16 Oct 2023 · 1 repository · arXiv:2310.10632Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Data Contamination Through the Lens of Time 16 Oct 2023 · 1 repository · arXiv:2310.10628Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Fine-tuning ChatGPT for Automatic Scoring 16 Oct 2023 · 0 repositories · arXiv:2310.10072
-
Investigating Bias in Multilingual Language Models: Cross-Lingual Transfer of Debiasing Techniques 16 Oct 2023 · 1 repository · arXiv:2310.10310
-
Learning to Rank Context for Named Entity Recognition Using a Synthetic Dataset 16 Oct 2023 · 1 repository · arXiv:2310.10118
-
MoConVQ: Unified Physics-Based Motion Control via Scalable Discrete Representations 16 Oct 2023 · 0 repositories · arXiv:2310.10198
-
Prediction of Arabic Legal Rulings using Large Language Models 16 Oct 2023 · 0 repositories · arXiv:2310.10260
-
TRANSOM: An Efficient Fault-Tolerant System for Training LLMs 16 Oct 2023 · 1 repository · arXiv:2310.10046Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Configuration Validation with Large Language Models 15 Oct 2023 · 0 repositories · arXiv:2310.09690
-
Domain-Specific Language Model Post-Training for Indonesian Financial NLP 15 Oct 2023 · 1 repository · arXiv:2310.09736
-
Empirical study of pretrained multilingual language models for zero-shot cross-lingual knowledge transfer in generation 15 Oct 2023 · 0 repositories · arXiv:2310.09917
-
Image Augmentation with Controlled Diffusion for Weakly-Supervised Semantic Segmentation 15 Oct 2023 · 0 repositories · arXiv:2310.09760
-
Large Language Model-Aware In-Context Learning for Code Generation 15 Oct 2023 · 0 repositories · arXiv:2310.09748
-
Large Language Models for In-Context Student Modeling: Synthesizing Student's Behavior in Visual Programming 15 Oct 2023 · 1 repository · arXiv:2310.10690
-
DPZero: Private Fine-Tuning of Language Models without Backpropagation 14 Oct 2023 · 1 repository · arXiv:2310.09639Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 17 harvested samples) · 3 pointer-only (licence)
-
Efficient Model-Agnostic Multi-Group Equivariant Networks 14 Oct 2023 · 0 repositories · arXiv:2310.09675
-
Leveraging Generative AI: Improving Software Metadata Classification with Generated Code-Comment Pairs 14 Oct 2023 · 0 repositories · arXiv:2311.03365
-
Assessing and Enhancing the Robustness of Large Language Models with Task Structure Variations for Logical Reasoning 13 Oct 2023 · 1 repository · arXiv:2310.09430Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Enhancing BERT-Based Visual Question Answering through Keyword-Driven Sentence Selection 13 Oct 2023 · 0 repositories · arXiv:2310.09432
-
From Words and Exercises to Wellness: Farsi Chatbot for Self-Attachment Technique 13 Oct 2023 · 0 repositories · arXiv:2310.09362
-
Human-in-the-loop Machine Translation with Large Language Model 13 Oct 2023 · 1 repository · arXiv:2310.08908
-
Retrieval-Generation Alignment for End-to-End Task-Oriented Dialogue System 13 Oct 2023 · 1 repository · arXiv:2310.08877
-
Table-GPT: Table-tuned GPT for Diverse Table Tasks 13 Oct 2023 · 0 repositories · arXiv:2310.09263
-
QUIK: Towards End-to-End 4-Bit Inference on Generative Large Language Models 13 Oct 2023 · 1 repository · arXiv:2310.09259Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Analyzing Textual Data for Fatality Classification in Afghanistan's Armed Conflicts: A BERT Approach 12 Oct 2023 · 0 repositories · arXiv:2310.08653
-
Detection and prediction of clopidogrel treatment failures using longitudinal structured electronic health records 12 Oct 2023 · 0 repositories · arXiv:2310.08757
-
Evaluating The Effectiveness of Capsule Neural Network in Toxic Comment Classification using Pre-trained BERT Embeddings 12 Oct 2023 · 1 repository
-
Jailbreaking Black Box Large Language Models in Twenty Queries 12 Oct 2023 · 1 repository · arXiv:2310.08419Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 7 harvested samples)
-
Large language models can replicate cross-cultural differences in personality 12 Oct 2023 · 0 repositories · arXiv:2310.10679
-
LEMON: Lossless model expansion 12 Oct 2023 · 1 repository · arXiv:2310.07999Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LLM-augmented Preference Learning from Natural Language 12 Oct 2023 · 0 repositories · arXiv:2310.08523
-
Multiclass Classification of Policy Documents with Large Language Models 12 Oct 2023 · 0 repositories · arXiv:2310.08167
-
Promptor: A Conversational and Autonomous Prompt Generation Agent for Intelligent Text Entry Techniques 12 Oct 2023 · 0 repositories · arXiv:2310.08101
-
QASiNa: Religious Domain Question Answering using Sirah Nabawiyah 12 Oct 2023 · 1 repository · arXiv:2310.08102
-
The Uncertainty-based Retrieval Framework for Ancient Chinese CWS and POS 12 Oct 2023 · 1 repository · arXiv:2310.08496
-
Training Generative Question-Answering on Synthetic Data Obtained from an Instruct-tuned Model 12 Oct 2023 · 0 repositories · arXiv:2310.08072
-
Accelerating Vision Transformers Based on Heterogeneous Attention Patterns 11 Oct 2023 · 0 repositories · arXiv:2310.07664
-
Diversity of Thought Improves Reasoning Abilities of LLMs 11 Oct 2023 · 0 repositories · arXiv:2310.07088
-
Do Large Language Models have Shared Weaknesses in Medical Question Answering? 11 Oct 2023 · 0 repositories · arXiv:2310.07225
-
Fast-ELECTRA for Efficient Pre-training 11 Oct 2023 · 0 repositories · arXiv:2310.07347
-
Found in the Middle: Permutation Self-Consistency Improves Listwise Ranking in Large Language Models 11 Oct 2023 · 1 repository · arXiv:2310.07712Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
InstructRetro: Instruction Tuning post Retrieval-Augmented Pretraining 11 Oct 2023 · 1 repository · arXiv:2310.07713
-
Jaeger: A Concatenation-Based Multi-Transformer VQA Model 11 Oct 2023 · 0 repositories · arXiv:2310.07091
-
Large Language Models Are Zero-Shot Time Series Forecasters 11 Oct 2023 · 2 repositories · arXiv:2310.07820Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Toward Understanding BERT-Like Pre-Training for DNA Foundation Models 11 Oct 2023 · 0 repositories · arXiv:2310.07644
-
Sparse Universal Transformer 11 Oct 2023 · 2 repositories · arXiv:2310.07096Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Uncovering Hidden Connections: Iterative Search and Reasoning for Video-grounded Dialog 11 Oct 2023 · 2 repositories · arXiv:2310.07259Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Diffusion Models for Wireless Communications 11 Oct 2023 · 0 repositories · arXiv:2310.07312
-
A Comparative Study of Transformer-based Neural Text Representation Techniques on Bug Triaging 10 Oct 2023 · 0 repositories · arXiv:2310.06913
-
Answer Candidate Type Selection: Text-to-Text Language Model for Closed Book Question Answering Meets Knowledge Graphs 10 Oct 2023 · 0 repositories · arXiv:2310.07008
-
Automated clinical coding using off-the-shelf large language models 10 Oct 2023 · 0 repositories · arXiv:2310.06552
-
GeoLLM: Extracting Geospatial Knowledge from Large Language Models 10 Oct 2023 · 1 repository · arXiv:2310.06213Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
GPT-4 as an Agronomist Assistant? Answering Agriculture Exams Using Large Language Models 10 Oct 2023 · 0 repositories · arXiv:2310.06225
-
Humans and language models diverge when predicting repeating text 10 Oct 2023 · 1 repository · arXiv:2310.06408Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Large Language Models for Propaganda Detection 10 Oct 2023 · 2 repositories · arXiv:2310.06422
-
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression 10 Oct 2023 · 3 repositories · arXiv:2310.06839
-
Sparse Fine-tuning for Inference Acceleration of Large Language Models 10 Oct 2023 · 1 repository · arXiv:2310.06927Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Auditing Gender Analyzers on Text Data 9 Oct 2023 · 0 repositories · arXiv:2310.06061
-
Automating Customer Service using LangChain: Building custom open-source GPT Chatbot for organizations 9 Oct 2023 · 0 repositories · arXiv:2310.05421
-
Cabbage Sweeter than Cake? Analysing the Potential of Large Language Models for Learning Conceptual Spaces 9 Oct 2023 · 0 repositories · arXiv:2310.05481
-
Foundation Models Meet Visualizations: Challenges and Opportunities 9 Oct 2023 · 0 repositories · arXiv:2310.05771
-
Exploring the Maze of Multilingual Modeling 9 Oct 2023 · 0 repositories · arXiv:2310.05404
-
SC-Safety: A Multi-round Open-ended Question Adversarial Safety Benchmark for Large Language Models in Chinese 9 Oct 2023 · 0 repositories · arXiv:2310.05818
-
The Program Testing Ability of Large Language Models for Code 9 Oct 2023 · 0 repositories · arXiv:2310.05727
-
Transformer Fusion with Optimal Transport 9 Oct 2023 · 1 repository · arXiv:2310.05719Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Are Emily and Greg Still More Employable than Lakisha and Jamal? Investigating Algorithmic Hiring Bias in the Era of ChatGPT 8 Oct 2023 · 0 repositories · arXiv:2310.05135
-
Benchmarking Large Language Models with Augmented Instructions for Fine-grained Information Extraction 8 Oct 2023 · 0 repositories · arXiv:2310.05092
-
Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection 8 Oct 2023 · 0 repositories · arXiv:2310.05115
-
Distantly-Supervised Joint Extraction with Noise-Robust Learning 8 Oct 2023 · 1 repository · arXiv:2310.04994
-
Enhancing Pre-Trained Language Models with Sentence Position Embeddings for Rhetorical Roles Recognition in Legal Opinions 8 Oct 2023 · 0 repositories · arXiv:2310.05276
-
LLM4VV: Developing LLM-Driven Testsuite for Compiler Validation 8 Oct 2023 · 1 repository · arXiv:2310.04963
-
RAC-BERT: Character Radical Enhanced BERT for Ancient Chinese 8 Oct 2023 · 0 repositories
-
Zero-Shot Detection of Machine-Generated Codes 8 Oct 2023 · 1 repository · arXiv:2310.05103
-
Do self-supervised speech and language models extract similar representations as human brain? 7 Oct 2023 · 0 repositories · arXiv:2310.04645
-
Large Language Models Only Pass Primary School Exams in Indonesia: A Comprehensive Test on IndoMMLU 7 Oct 2023 · 1 repository · arXiv:2310.04928Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT 7 Oct 2023 · 2 repositories · arXiv:2310.04673Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Question-focused Summarization by Decomposing Articles into Facts and Opinions and Retrieving Entities 7 Oct 2023 · 0 repositories · arXiv:2310.04880
-
A Process for Topic Modelling Via Word Embeddings 6 Oct 2023 · 0 repositories · arXiv:2312.03705
-
Automatic Aspect Extraction from Scientific Texts 6 Oct 2023 · 1 repository · arXiv:2310.04074
-
Copy Suppression: Comprehensively Understanding an Attention Head 6 Oct 2023 · 1 repository · arXiv:2310.04625Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Keyword Augmented Retrieval: Novel framework for Information Retrieval integrated with speech interface 6 Oct 2023 · 0 repositories · arXiv:2310.04205
-
Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models 6 Oct 2023 · 2 repositories · arXiv:2310.04406Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Quantized Transformer Language Model Implementations on Edge Devices 6 Oct 2023 · 0 repositories · arXiv:2310.03971
-
Segmented Harmonic Loss: Handling Class-Imbalanced Multi-Label Clinical Data for Medical Coding with Large Language Models 6 Oct 2023 · 0 repositories · arXiv:2310.04595
-
Effective Slogan Generation with Noise Perturbation 6 Oct 2023 · 1 repository · arXiv:2310.04472
-
Agent Instructs Large Language Models to be General Zero-Shot Reasoners 5 Oct 2023 · 1 repository · arXiv:2310.03710Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Automating Human Tutor-Style Programming Feedback: Leveraging GPT-4 Tutor Model for Hint Generation and GPT-3.5 Student Model for Hint Validation 5 Oct 2023 · 2 repositories · arXiv:2310.03780Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning 5 Oct 2023 · 1 repository · arXiv:2310.03249Syntology official (archive's flag): 28 ran · 28 ran (of which 0 constructed an object rather than computing a result; 28 with no instrument failure: 0 honoured, 0 violated, 28 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 28 harvested samples) · 28 pointer-only (licence)
-
DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines 5 Oct 2023 · 3 repositories · arXiv:2310.03714Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To! 5 Oct 2023 · 1 repository · arXiv:2310.03693Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks 5 Oct 2023 · 1 repository · arXiv:2310.03684
-
A Survey of GPT-3 Family Large Language Models Including ChatGPT and GPT-4 4 Oct 2023 · 0 repositories · arXiv:2310.12321
-
COVID-19 South African Vaccine Hesitancy Models Show Boost in Performance Upon Fine-Tuning on M-pox Tweets 4 Oct 2023 · 0 repositories · arXiv:2310.04453
-
Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning 4 Oct 2023 · 1 repository · arXiv:2310.03094
-
Memoria: Resolving Fateful Forgetting Problem through Human-Inspired Memory Architecture 4 Oct 2023 · 1 repository · arXiv:2310.03052Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 10 harvested samples)
-
NOLA: Compressing LoRA using Linear Combination of Random Basis 4 Oct 2023 · 1 repository · arXiv:2310.02556Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)