Methods › General › Regularization › Attention Dropout › Papers, page 43
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 43 of 109: papers 4,201 to 4,300 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Speech-based Slot Filling using Large Language Models 13 Nov 2023 · 0 repositories · arXiv:2311.07418
-
STEER: Unified Style Transfer with Expert Reinforcement 13 Nov 2023 · 1 repository · arXiv:2311.07167
-
Teach me with a Whisper: Enhancing Large Language Models for Analyzing Spoken Transcripts using Speech Embeddings 13 Nov 2023 · 0 repositories · arXiv:2311.07014
-
Controllable Topic-Focused Abstractive Summarization 12 Nov 2023 · 0 repositories · arXiv:2311.06724
-
From Complex to Simple: Unraveling the Cognitive Tree for Reasoning with Small Language Models 12 Nov 2023 · 0 repositories · arXiv:2311.06754
-
GIELLM: Japanese General Information Extraction Large Language Model Utilizing Mutual Reinforcement Effect 12 Nov 2023 · 0 repositories · arXiv:2311.06838
-
Retrieval and Generative Approaches for a Pregnancy Chatbot in Nepali with Stemmed and Non-Stemmed Data : A Comparative Study 12 Nov 2023 · 0 repositories · arXiv:2311.06898
-
Large Language Models are In-context Teachers for Knowledge Reasoning 12 Nov 2023 · 0 repositories · arXiv:2311.06985
-
Data Contamination Quiz: A Tool to Detect and Estimate Contamination in Large Language Models 10 Nov 2023 · 2 repositories · arXiv:2311.06233Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Establishing Performance Baselines in Fine-Tuning, Retrieval-Augmented Generation and Soft-Prompting for Non-Specialist LLM Users 10 Nov 2023 · 0 repositories · arXiv:2311.05903
-
Exploring Fine-tuning ChatGPT for News Recommendation 10 Nov 2023 · 0 repositories · arXiv:2311.05850
-
Smart Agent-Based Modeling: On the Use of Large Language Models in Computer Simulations 10 Nov 2023 · 4 repositories · arXiv:2311.06330Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Deep Natural Language Feature Learning for Interpretable Prediction 9 Nov 2023 · 1 repository · arXiv:2311.05754
-
GeoFormer: Predicting Human Mobility using Generative Pre-trained Transformer (GPT) 9 Nov 2023 · 0 repositories · arXiv:2311.05092
-
Large Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document Summarisation 9 Nov 2023 · 0 repositories · arXiv:2311.05169
-
Leveraging Artificial Intelligence Technology for Mapping Research to Sustainable Development Goals: A Case Study 9 Nov 2023 · 0 repositories · arXiv:2311.16162
-
LogShield: A Transformer-based APT Detection System Leveraging Self-Attention 9 Nov 2023 · 0 repositories · arXiv:2311.05733
-
Agent Lumos: Unified and Modular Training for Open-Source Language Agents 9 Nov 2023 · 2 repositories · arXiv:2311.05657Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Vision Encoder-Decoder Models for AI Coaching 9 Nov 2023 · 2 repositories · arXiv:2311.16161
-
DACBERT: Leveraging Dependency Agreement for Cost-Efficient Bert Pretraining 8 Nov 2023 · 0 repositories · arXiv:2311.04799
-
Deep Learning Brasil at ABSAPT 2022: Portuguese Transformer Ensemble Approaches 8 Nov 2023 · 1 repository · arXiv:2311.05051
-
DeepLearningBrasil@LT-EDI-2023: Exploring Deep Learning Techniques for Detecting Depression in Social Media Text 8 Nov 2023 · 1 repository · arXiv:2311.05047
-
Determination of toxic comments and unintended model bias minimization using Deep learning approach 8 Nov 2023 · 1 repository · arXiv:2311.04789
-
Massive Editing for Large Language Models via Meta Learning 8 Nov 2023 · 1 repository · arXiv:2311.04661Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Pre-training LLMs using human-like development data corpus 8 Nov 2023 · 0 repositories · arXiv:2311.04666
-
Rethinking Benchmark and Contamination for Language Models with Rephrased Samples 8 Nov 2023 · 1 repository · arXiv:2311.04850Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Enhancing LLM Intelligence with ARM-RAG: Auxiliary Rationale Memory for Retrieval Augmented Generation 7 Nov 2023 · 0 repositories · arXiv:2311.04177
-
Evaluating Large Language Models in Ophthalmology 7 Nov 2023 · 0 repositories · arXiv:2311.04933
-
Identifying and Mitigating Vulnerabilities in LLM-Integrated Applications 7 Nov 2023 · 0 repositories · arXiv:2311.16153
-
Towards Interpretable Sequence Continuation: Analyzing Shared Circuits in Large Language Models 7 Nov 2023 · 1 repository · arXiv:2311.04131Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Modelling Sentiment Analysis: LLMs and data augmentation techniques 7 Nov 2023 · 1 repository · arXiv:2311.04139
-
Neuro-GPT: Towards A Foundation Model for EEG 7 Nov 2023 · 1 repository · arXiv:2311.03764Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Personality Style Recognition via Machine Learning: Identifying Anaclitic and Introjective Personality Styles from Patients' Speech 7 Nov 2023 · 0 repositories · arXiv:2311.04088
-
Accumulating Word Representations in Multi-level Context Integration for ERC Task 6 Nov 2023 · 1 repository
-
Adapting Pre-trained Generative Models for Extractive Question Answering 6 Nov 2023 · 0 repositories · arXiv:2311.02961
-
DeepInception: Hypnotize Large Language Model to Be Jailbreaker 6 Nov 2023 · 1 repository · arXiv:2311.03191Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
In-Context Learning for Knowledge Base Question Answering for Unmanned Systems based on Large Language Models 6 Nov 2023 · 0 repositories · arXiv:2311.02956
-
Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch 6 Nov 2023 · 3 repositories · arXiv:2311.03099Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Unraveling Downstream Gender Bias from Large Language Models: A Study on AI Educational Writing Assistance 6 Nov 2023 · 1 repository · arXiv:2311.03311
-
AI-TA: Towards an Intelligent Question-Answer Teaching Assistant using Open-Source LLMs 5 Nov 2023 · 1 repository · arXiv:2311.02775
-
Evaluating the Potential of Leading Large Language Models in Reasoning Biology Questions 5 Nov 2023 · 0 repositories · arXiv:2311.07582
-
Extraction of Atypical Aspects from Customer Reviews: Datasets and Experiments with Language Models 5 Nov 2023 · 2 repositories · arXiv:2311.02702
-
UID as a Guiding Metric for Automated Authorship Obfuscation 5 Nov 2023 · 0 repositories · arXiv:2312.03709
-
You Only Forward Once: Prediction and Rationalization in A Single Forward Pass 4 Nov 2023 · 0 repositories · arXiv:2311.02344
-
Automating Governing Knowledge Commons and Contextual Integrity (GKC-CI) Privacy Policy Annotations with Large Language Models 3 Nov 2023 · 1 repository · arXiv:2311.02192
-
COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning 3 Nov 2023 · 0 repositories · arXiv:2311.02248
-
Data-Free Distillation of Language Model by Text-to-Text Transfer 3 Nov 2023 · 0 repositories · arXiv:2311.01689
-
Efficient Black-Box Adversarial Attacks on Neural Text Detectors 3 Nov 2023 · 1 repository · arXiv:2311.01873
-
Exploring the Numerical Reasoning Capabilities of Language Models: A Comprehensive Analysis on Tabular Data 3 Nov 2023 · 0 repositories · arXiv:2311.02216
-
FaMeSumm: Investigating and Improving Faithfulness of Medical Summarization 3 Nov 2023 · 1 repository · arXiv:2311.02271Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Indo LEGO-ABSA: A Multitask Generative Aspect Based Sentiment Analysis for Indonesian Language 3 Nov 2023 · 1 repository · arXiv:2311.01757
-
Simplifying Transformer Blocks 3 Nov 2023 · 1 repository · arXiv:2311.01906
-
Better Together: Enhancing Generative Knowledge Graph Completion with Language Models and Neighborhood Information 2 Nov 2023 · 1 repository · arXiv:2311.01326
-
Learning Defect Prediction from Unrealistic Data 2 Nov 2023 · 0 repositories · arXiv:2311.00931
-
Long Story Short: a Summarize-then-Search Method for Long Video Question Answering 2 Nov 2023 · 1 repository · arXiv:2311.01233
-
MAAIG: Motion Analysis And Instruction Generation 2 Nov 2023 · 0 repositories · arXiv:2311.00980
-
Measuring Five Accountable Talk Moves to Improve Instruction at Scale 2 Nov 2023 · 0 repositories · arXiv:2311.10749
-
Server-side Rescoring of Spoken Entity-centric Knowledge Queries for Virtual Assistants 2 Nov 2023 · 0 repositories · arXiv:2311.01398
-
An Improved Transformer-based Model for Detecting Phishing, Spam, and Ham: A Large Language Model Approach 1 Nov 2023 · 0 repositories · arXiv:2311.04913
-
Are Large Language Models Reliable Judges? A Study on the Factuality Evaluation Capabilities of LLMs 1 Nov 2023 · 0 repositories · arXiv:2311.00681
-
Attention Alignment and Flexible Positional Embeddings Improve Transformer Length Extrapolation 1 Nov 2023 · 0 repositories · arXiv:2311.00684
-
Continuous Training and Fine-tuning for Domain-Specific Language Models in Medical Question Answering 1 Nov 2023 · 0 repositories · arXiv:2311.00204
-
Data Augmentation for Code Translation with Comparable Corpora and Multiple References 1 Nov 2023 · 1 repository · arXiv:2311.00317Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Entity Alignment Method of Science and Technology Patent based on Graph Convolution Network and Information Fusion 1 Nov 2023 · 0 repositories · arXiv:2311.00300
-
Is GPT Powerful Enough to Analyze the Emotions of Memes? 1 Nov 2023 · 0 repositories · arXiv:2311.00223
-
Syntactic Inductive Bias in Transformer Language Models: Especially Helpful for Low-Resource Languages? 1 Nov 2023 · 1 repository · arXiv:2311.00268
-
Unsupervised Lexical Simplification with Context Augmentation 1 Nov 2023 · 1 repository · arXiv:2311.00310
-
BERTwich: Extending BERT's Capabilities to Model Dialectal and Noisy Text 31 Oct 2023 · 0 repositories · arXiv:2311.00116
-
Breaking the Token Barrier: Chunking and Convolution for Efficient Long Text Classification with BERT 31 Oct 2023 · 0 repositories · arXiv:2310.20558
-
Do large language models solve verbal analogies like children do? 31 Oct 2023 · 0 repositories · arXiv:2310.20384
-
Does GPT-4 pass the Turing test? 31 Oct 2023 · 0 repositories · arXiv:2310.20216
-
EELBERT: Tiny Models through Dynamic Embeddings 31 Oct 2023 · 0 repositories · arXiv:2310.20144
-
Efficient Classification of Student Help Requests in Programming Courses Using Large Language Models 31 Oct 2023 · 0 repositories · arXiv:2310.20105
-
FA Team at the NTCIR-17 UFO Task 31 Oct 2023 · 0 repositories · arXiv:2310.20322
-
GAR-meets-RAG Paradigm for Zero-Shot Information Retrieval 31 Oct 2023 · 0 repositories · arXiv:2310.20158
-
Increasing The Performance of Cognitively Inspired Data-Efficient Language Models via Implicit Structure Building 31 Oct 2023 · 1 repository · arXiv:2310.20589
-
Interactive Multi-fidelity Learning for Cost-effective Adaptation of Language Model with Sparse Human Supervision 31 Oct 2023 · 0 repositories · arXiv:2310.20153
-
PsyCoT: Psychological Questionnaire as Powerful Chain-of-Thought for Personality Detection 31 Oct 2023 · 1 repository · arXiv:2310.20256
-
Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests 31 Oct 2023 · 0 repositories · arXiv:2310.20320
-
BTRec: BERT-Based Trajectory Recommendation for Personalized Tours 30 Oct 2023 · 1 repository · arXiv:2310.19886
-
Generating Medical Prescriptions with Conditional Transformer 30 Oct 2023 · 1 repository · arXiv:2310.19727Syntology official (archive's flag): 10 ran · 10 ran (of which 4 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Herd: Using multiple, smaller LLMs to match the performances of proprietary, large LLMs via an intelligent composer 30 Oct 2023 · 0 repositories · arXiv:2310.19902
-
Interpretable-by-Design Text Understanding with Iteratively Generated Concept Bottleneck 30 Oct 2023 · 1 repository · arXiv:2310.19660
-
Jina Embeddings 2: 8192-Token General-Purpose Text Embeddings for Long Documents 30 Oct 2023 · 2 repositories · arXiv:2310.19923Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
LitCab: Lightweight Language Model Calibration over Short- and Long-form Responses 30 Oct 2023 · 1 repository · arXiv:2310.19208
-
Partial Tensorized Transformers for Natural Language Processing 30 Oct 2023 · 0 repositories · arXiv:2310.20077
-
Remember what you did so you know what to do next 30 Oct 2023 · 0 repositories · arXiv:2311.01468
-
Split-NER: Named Entity Recognition via Two Question-Answering-based Classifications 30 Oct 2023 · 1 repository · arXiv:2310.19942Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization 30 Oct 2023 · 1 repository · arXiv:2310.20033Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
EtiCor: Corpus for Analyzing LLMs for Etiquettes 29 Oct 2023 · 1 repository · arXiv:2310.18974
-
From Chatbots to PhishBots? -- Preventing Phishing scams created using ChatGPT, Google Bard and Claude 29 Oct 2023 · 0 repositories · arXiv:2310.19181
-
Prompt-Engineering and Transformer-based Question Generation and Evaluation 29 Oct 2023 · 0 repositories · arXiv:2310.18867
-
Retrofitting Light-weight Language Models for Emotions using Supervised Contrastive Learning 29 Oct 2023 · 0 repositories · arXiv:2310.18930
-
Efficient kernel surrogates for neural network-based regression 28 Oct 2023 · 0 repositories · arXiv:2310.18612
-
The Synergy of Speculative Decoding and Batching in Serving Large Language Models 28 Oct 2023 · 0 repositories · arXiv:2310.18813
-
TLM: Token-Level Masking for Transformers 28 Oct 2023 · 1 repository · arXiv:2310.18738Syntology official (archive's flag): 11 ran · 23 ran (of which 0 constructed an object rather than computing a result; 22 with no instrument failure: 0 honoured, 0 violated, 22 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 28 harvested samples) · 28 pointer-only (licence)
-
Large language models for aspect-based sentiment analysis 27 Oct 2023 · 1 repository · arXiv:2310.18025Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Lost in Translation, Found in Spans: Identifying Claims in Multilingual Social Media 27 Oct 2023 · 1 repository · arXiv:2310.18205
-
OffMix-3L: A Novel Code-Mixed Dataset in Bangla-English-Hindi for Offensive Language Identification 27 Oct 2023 · 1 repository · arXiv:2310.18387
-
SentMix-3L: A Bangla-English-Hindi Code-Mixed Dataset for Sentiment Analysis 27 Oct 2023 · 1 repository · arXiv:2310.18023