Methods › General › Regularization › Attention Dropout › Papers, page 31
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 31 of 109: papers 3,001 to 3,100 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Utilizing Large Language Models to Generate Synthetic Data to Increase the Performance of BERT-Based Neural Networks 8 May 2024 · 0 repositories · arXiv:2405.06695
-
An LLM-Tool Compiler for Fused Parallel Function Calling 7 May 2024 · 0 repositories · arXiv:2405.17438
-
Enriched BERT Embeddings for Scholarly Publication Classification 7 May 2024 · 1 repository · arXiv:2405.04136
-
ERATTA: Extreme RAG for Table To Answers with Large Language Models 7 May 2024 · 0 repositories · arXiv:2405.03963
-
Evaluating Text Summaries Generated by Large Language Models Using OpenAI's GPT 7 May 2024 · 0 repositories · arXiv:2405.04053
-
GPT-Enabled Cybersecurity Training: A Tailored Approach for Effective Awareness 7 May 2024 · 0 repositories · arXiv:2405.04138
-
How does GPT-2 Predict Acronyms? Extracting and Understanding a Circuit via Mechanistic Interpretability 7 May 2024 · 1 repository · arXiv:2405.04156Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Long Context Alignment with Short Instructions and Synthesized Positions 7 May 2024 · 0 repositories · arXiv:2405.03939
-
Remote Diffusion 7 May 2024 · 0 repositories · arXiv:2405.04717
-
Revisiting Character-level Adversarial Attacks for Language Models 7 May 2024 · 1 repository · arXiv:2405.04346Syntology official (archive's flag): 23 ran · 23 ran (of which 3 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 10 where Syntology's instrument failed) · 8 unverified (of 31 harvested samples)
-
Robust Implementation of Retrieval-Augmented Generation on Edge-based Computing-in-Memory Architectures 7 May 2024 · 0 repositories · arXiv:2405.04700
-
SUTRA: Scalable Multilingual Language Model Architecture 7 May 2024 · 0 repositories · arXiv:2405.06694
-
The Silicon Ceiling: Auditing GPT's Race and Gender Biases in Hiring 7 May 2024 · 0 repositories · arXiv:2405.04412
-
Utilizing GPT to Enhance Text Summarization: A Strategy to Minimize Hallucinations 7 May 2024 · 0 repositories · arXiv:2405.04039
-
Anchored Answers: Unravelling Positional Bias in GPT-2's Multiple-Choice Questions 6 May 2024 · 1 repository · arXiv:2405.03205
-
Characterizing the Dilemma of Performance and Index Size in Billion-Scale Vector Search and Breaking It with Second-Tier Memory 6 May 2024 · 0 repositories · arXiv:2405.03267
-
Compressing Long Context for Enhancing RAG with AMR-based Concept Distillation 6 May 2024 · 0 repositories · arXiv:2405.03085
-
Detecting Android Malware: From Neural Embeddings to Hands-On Validation with BERTroid 6 May 2024 · 0 repositories · arXiv:2405.03620
-
Detecting Anti-Semitic Hate Speech using Transformer-based Large Language Models 6 May 2024 · 0 repositories · arXiv:2405.03794
-
ERAGent: Enhancing Retrieval-Augmented Language Models with Improved Accuracy, Efficiency, and Personalization 6 May 2024 · 1 repository · arXiv:2405.06683
-
Hire Me or Not? Examining Language Model's Behavior with Occupation Attributes 6 May 2024 · 1 repository · arXiv:2405.06687Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Large Language Models Reveal Information Operation Goals, Tactics, and Narrative Frames 6 May 2024 · 1 repository · arXiv:2405.03688
-
Can Large Language Models Make the Grade? An Empirical Study Evaluating LLMs Ability to Mark Short Answer Questions in K-12 Education 5 May 2024 · 0 repositories · arXiv:2405.02985
-
Labeling supervised fine-tuning data with the scaling law 5 May 2024 · 2 repositories · arXiv:2405.02817
-
Leveraging Lecture Content for Improved Feedback: Explorations with GPT-4 and Retrieval Augmented Generation 5 May 2024 · 0 repositories · arXiv:2405.06681
-
Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models 5 May 2024 · 0 repositories · arXiv:2405.02917
-
Stochastic RAG: End-to-End Retrieval-Augmented Generation through Expected Utility Maximization 5 May 2024 · 0 repositories · arXiv:2405.02816
-
Unraveling the Dominance of Large Language Models Over Transformer Models for Bangla Natural Language Inference: A Comprehensive Study 5 May 2024 · 1 repository · arXiv:2405.02937
-
A Combination of BERT and Transformer for Vietnamese Spelling Correction 4 May 2024 · 0 repositories · arXiv:2405.02573
-
Assessing Adversarial Robustness of Large Language Models: An Empirical Study 4 May 2024 · 0 repositories · arXiv:2405.02764
-
Analyzing Narrative Processing in Large Language Models (LLMs): Using GPT4 to test BERT 3 May 2024 · 0 repositories · arXiv:2405.02024
-
Comparative Analysis of Retrieval Systems in the Real World 3 May 2024 · 0 repositories · arXiv:2405.02048
-
DALLMi: Domain Adaption for LLM-based Multi-label Classifier 3 May 2024 · 1 repository · arXiv:2405.01883
-
Evaluating Large Language Models for Structured Science Summarization in the Open Research Knowledge Graph 3 May 2024 · 0 repositories · arXiv:2405.02105
-
Exploiting ChatGPT for Diagnosing Autism-Associated Language Disorders and Identifying Distinct Features 3 May 2024 · 1 repository · arXiv:2405.01799
-
Exploring Combinatorial Problem Solving with Large Language Models: A Case Study on the Travelling Salesman Problem Using GPT-3.5 Turbo 3 May 2024 · 0 repositories · arXiv:2405.01997
-
Attribution in Scientific Literature: New Benchmark and Methods 3 May 2024 · 0 repositories · arXiv:2405.02228
-
Structural Pruning of Pre-trained Language Models via Neural Architecture Search 3 May 2024 · 1 repository · arXiv:2405.02267
-
A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law 2 May 2024 · 1 repository · arXiv:2405.01769
-
Bayesian Optimization with LLM-Based Acquisition Functions for Natural Language Preference Elicitation 2 May 2024 · 0 repositories · arXiv:2405.00981
-
Early Transformers: A study on Efficient Training of Transformer Models through Early-Bird Lottery Tickets 2 May 2024 · 0 repositories · arXiv:2405.02353
-
Investigating Wit, Creativity, and Detectability of Large Language Models in Domain-Specific Writing Style Adaptation of Reddit's Showerthoughts 2 May 2024 · 1 repository · arXiv:2405.01660
-
The Effectiveness of LLMs as Annotators: A Comparative Overview and Empirical Analysis of Direct Representation 2 May 2024 · 0 repositories · arXiv:2405.01299
-
UQA: Corpus for Urdu Question Answering 2 May 2024 · 3 repositories · arXiv:2405.01458
-
A Named Entity Recognition and Topic Modeling-based Solution for Locating and Better Assessment of Natural Disasters in Social Media 1 May 2024 · 0 repositories · arXiv:2405.00903
-
CourseAssist: Pedagogically Appropriate AI Tutor for Computer Science Education 1 May 2024 · 0 repositories · arXiv:2407.10246
-
How Can I Improve? Using GPT to Highlight the Desired and Undesired Parts of Open-ended Responses 1 May 2024 · 0 repositories · arXiv:2405.00291
-
Integrating A.I. in Higher Education: Protocol for a Pilot Study with 'SAMCares: An Adaptive Learning Hub' 1 May 2024 · 1 repository · arXiv:2405.00330
-
Opinion Mining Using Pre-Trained Large Language Models: Identifying the Type, Polarity, Intensity, Expression, and Source of Private States 1 May 2024 · 1 repository
-
Better & Faster Large Language Models via Multi-token Prediction 30 Apr 2024 · 1 repository · arXiv:2404.19737
-
Can Large Language Models put 2 and 2 together? Probing for Entailed Arithmetical Relationships 30 Apr 2024 · 0 repositories · arXiv:2404.19432
-
Do Large Language Models Understand Conversational Implicature -- A case study with a chinese sitcom 30 Apr 2024 · 1 repository · arXiv:2404.19509Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Graph Neural Network Approach to Semantic Type Detection in Tables 30 Apr 2024 · 1 repository · arXiv:2405.00123
-
Graphical Reasoning: LLM-based Semi-Open Relation Extraction 30 Apr 2024 · 1 repository · arXiv:2405.00216
-
PANGeA: Procedural Artificial Narrative using Generative AI for Turn-Based Video Games 30 Apr 2024 · 0 repositories · arXiv:2404.19721
-
Towards a Search Engine for Machines: Unified Ranking for Multiple Retrieval-Augmented Large Language Models 30 Apr 2024 · 1 repository · arXiv:2405.00175
-
TuBA: Cross-Lingual Transferability of Backdoor Attacks in LLMs with Instruction Tuning 30 Apr 2024 · 0 repositories · arXiv:2404.19597
-
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models 29 Apr 2024 · 1 repository · arXiv:2404.18353
-
Evaluating and Mitigating Linguistic Discrimination in Large Language Models 29 Apr 2024 · 0 repositories · arXiv:2404.18534
-
FeDeRA:Efficient Fine-tuning of Language Models in Federated Learning Leveraging Weight Decomposition 29 Apr 2024 · 0 repositories · arXiv:2404.18848
-
GPT-4 passes most of the 297 written Polish Board Certification Examinations 29 Apr 2024 · 0 repositories · arXiv:2405.01589
-
PECC: Problem Extraction and Coding Challenges 29 Apr 2024 · 1 repository · arXiv:2404.18766
-
Time Machine GPT 29 Apr 2024 · 0 repositories · arXiv:2404.18543
-
Can Perplexity Predict Fine-Tuning Performance? An Investigation of Tokenization Effects on Sequential Language Models for Nepali 28 Apr 2024 · 0 repositories · arXiv:2404.18071
-
L3Cube-MahaNews: News-based Short Text and Long Document Classification Datasets in Marathi 28 Apr 2024 · 1 repository · arXiv:2404.18216
-
Tabular Embedding Model (TEM): Finetuning Embedding Models For Tabular RAG Applications 28 Apr 2024 · 0 repositories · arXiv:2405.01585
-
Detection of Conspiracy Theories Beyond Keyword Bias in German-Language Telegram Using Large Language Models 27 Apr 2024 · 0 repositories · arXiv:2404.17985
-
Enhancing Pre-Trained Generative Language Models with Question Attended Span Extraction on Machine Reading Comprehension 27 Apr 2024 · 0 repositories · arXiv:2404.17991
-
Evaluation of Few-Shot Learning for Classification Tasks in the Polish Language 27 Apr 2024 · 0 repositories · arXiv:2404.17832
-
GPT for Games: A Scoping Review (2020-2023) 27 Apr 2024 · 0 repositories · arXiv:2404.17794
-
MRScore: Evaluating Radiology Report Generation with LLM-based Reward System 27 Apr 2024 · 0 repositories · arXiv:2404.17778
-
Tool Calling: Enhancing Medication Consultation via Retrieval-Augmented Large Language Models 27 Apr 2024 · 0 repositories · arXiv:2404.17897
-
Automated Data Visualization from Natural Language via Large Language Models: An Exploratory Study 26 Apr 2024 · 1 repository · arXiv:2404.17136
-
"ChatGPT Is Here to Help, Not to Replace Anybody" -- An Evaluation of Students' Opinions On Integrating ChatGPT In CS Courses 26 Apr 2024 · 0 repositories · arXiv:2404.17443
-
Enhancing Legal Compliance and Regulation Analysis with Large Language Models 26 Apr 2024 · 0 repositories · arXiv:2404.17522
-
Human-Imperceptible Retrieval Poisoning Attacks in LLM-Powered Applications 26 Apr 2024 · 0 repositories · arXiv:2404.17196
-
Prompting Towards Alleviating Code-Switched Data Scarcity in Under-Resourced Languages with GPT as a Pivot 26 Apr 2024 · 0 repositories · arXiv:2404.17216
-
Quantifying Memorization and Detecting Training Data of Pre-trained Language Models using Japanese Newspaper 26 Apr 2024 · 0 repositories · arXiv:2404.17143
-
Retrieval-Augmented Generation with Knowledge Graphs for Customer Service Question Answering 26 Apr 2024 · 0 repositories · arXiv:2404.17723
-
A Short Survey of Human Mobility Prediction in Epidemic Modeling from Transformers to LLMs 25 Apr 2024 · 0 repositories · arXiv:2404.16921
-
Análise de ambiguidade linguística em modelos de linguagem de grande escala (LLMs) 25 Apr 2024 · 0 repositories · arXiv:2404.16653
-
Evaluating Consistency and Reasoning Capabilities of Large Language Models 25 Apr 2024 · 0 repositories · arXiv:2404.16478
-
Incorporating Lexical and Syntactic Knowledge for Unsupervised Cross-Lingual Transfer 25 Apr 2024 · 1 repository · arXiv:2404.16627
-
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages 25 Apr 2024 · 1 repository · arXiv:2404.16816
-
Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans and GPT-4 25 Apr 2024 · 0 repositories · arXiv:2404.16692
-
WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models 25 Apr 2024 · 1 repository · arXiv:2404.16308Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
A Comprehensive Survey on Evaluating Large Language Model Applications in the Medical Industry 24 Apr 2024 · 0 repositories · arXiv:2404.15777
-
Automated Creation of Source Code Variants of a Cryptographic Hash Function Implementation Using Generative Pre-Trained Transformer Models 24 Apr 2024 · 0 repositories · arXiv:2404.15681
-
BERT vs GPT for financial engineering 24 Apr 2024 · 0 repositories · arXiv:2405.12990
-
Detecting Conceptual Abstraction in LLMs 24 Apr 2024 · 0 repositories · arXiv:2404.15848
-
From Local to Global: A Graph RAG Approach to Query-Focused Summarization 24 Apr 2024 · 3 repositories · arXiv:2404.16130Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Prompt Leakage effect and defense strategies for multi-turn LLM interactions 24 Apr 2024 · 0 repositories · arXiv:2404.16251
-
Learning Long-form Video Prior via Generative Pre-Training 24 Apr 2024 · 1 repository · arXiv:2404.15909
-
Studying Large Language Model Behaviors Under Context-Memory Conflicts With Real Documents 24 Apr 2024 · 1 repository · arXiv:2404.16032Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Telco-RAG: Navigating the Challenges of Retrieval-Augmented Language Models for Telecommunications 24 Apr 2024 · 1 repository · arXiv:2404.15939
-
The Promise and Challenges of Using LLMs to Accelerate the Screening Process of Systematic Reviews 24 Apr 2024 · 0 repositories · arXiv:2404.15667
-
AI and Machine Learning for Next Generation Science Assessments 23 Apr 2024 · 0 repositories · arXiv:2405.06660
-
Automated Multi-Language to English Machine Translation Using Generative Pre-Trained Transformers 23 Apr 2024 · 0 repositories · arXiv:2404.14680
-
IryoNLP at MEDIQA-CORR 2024: Tackling the Medical Error Detection & Correction Task On the Shoulders of Medical Agents 23 Apr 2024 · 0 repositories · arXiv:2404.15488
-
Evaluating the Efficacy of Large Language Models in Identifying Phishing Attempts 23 Apr 2024 · 0 repositories · arXiv:2404.15485