Methods › General › Regularization › Attention Dropout › Papers, page 34
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 34 of 109: papers 3,301 to 3,400 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Improving Retrieval for RAG based Question Answering Models on Financial Documents 23 Mar 2024 · 0 repositories · arXiv:2404.07221
-
LlamBERT: Large-scale low-cost data annotation in NLP 23 Mar 2024 · 1 repository · arXiv:2403.15938
-
Towards a RAG-based Summarization Agent for the Electron-Ion Collider 23 Mar 2024 · 1 repository · arXiv:2403.15729
-
Using Large Language Models for OntoClean-based Ontology Refinement 23 Mar 2024 · 0 repositories · arXiv:2403.15864
-
SOEN-101: Code Generation by Emulating Software Process Models Using Large Language Model Agents 23 Mar 2024 · 0 repositories · arXiv:2403.15852
-
Adapprox: Adaptive Approximation in Adam Optimization via Randomized Low-Rank Matrices 22 Mar 2024 · 0 repositories · arXiv:2403.14958
-
Blended RAG: Improving RAG (Retriever-Augmented Generation) Accuracy with Semantic Search and Hybrid Query-Based Retrievers 22 Mar 2024 · 1 repository · arXiv:2404.07220Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Can large language models explore in-context? 22 Mar 2024 · 0 repositories · arXiv:2403.15371
-
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation 22 Mar 2024 · 1 repository · arXiv:2403.14965
-
MasonTigers at SemEval-2024 Task 1: An Ensemble Approach for Semantic Textual Relatedness 22 Mar 2024 · 0 repositories · arXiv:2403.14990
-
Measuring Gender and Racial Biases in Large Language Models 22 Mar 2024 · 0 repositories · arXiv:2403.15281
-
On Zero-Shot Counterspeech Generation by LLMs 22 Mar 2024 · 1 repository · arXiv:2403.14938
-
Optimal path for Biomedical Text Summarization Using Pointer GPT 22 Mar 2024 · 0 repositories · arXiv:2404.08654
-
Selecting Query-bag as Pseudo Relevance Feedback for Information-seeking Conversations 22 Mar 2024 · 0 repositories · arXiv:2404.04272
-
SensoryT5: Infusing Sensorimotor Norms into T5 for Enhanced Fine-grained Emotion Classification 22 Mar 2024 · 0 repositories · arXiv:2403.15574
-
Text Clustering with Large Language Model Embeddings 22 Mar 2024 · 0 repositories · arXiv:2403.15112
-
Emergent World Models and Latent Variable Estimation in Chess-Playing Language Models 21 Mar 2024 · 1 repository · arXiv:2403.15498Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
FIT-RAG: Black-Box RAG with Factual Information and Token Reduction 21 Mar 2024 · 0 repositories · arXiv:2403.14374
-
LLM-based Extraction of Contradictions from Patents 21 Mar 2024 · 0 repositories · arXiv:2403.14258
-
PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model 21 Mar 2024 · 1 repository · arXiv:2403.14598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding 21 Mar 2024 · 1 repository · arXiv:2403.14743
-
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving 20 Mar 2024 · 0 repositories · arXiv:2403.13331
-
AUD-TGN: Advancing Action Unit Detection with Temporal Convolution and GPT-2 in Wild Audiovisual Contexts 20 Mar 2024 · 0 repositories · arXiv:2403.13678
-
Ax-to-Grind Urdu: Benchmark Dataset for Urdu Fake News Detection 20 Mar 2024 · 1 repository · arXiv:2403.14037
-
Efficient argument classification with compact language models and ChatGPT-4 refinements 20 Mar 2024 · 0 repositories · arXiv:2403.15473
-
Incentivizing News Consumption on Social Media Platforms Using Large Language Models and Realistic Bot Accounts 20 Mar 2024 · 1 repository · arXiv:2403.13362
-
Motion Generation from Fine-grained Textual Descriptions 20 Mar 2024 · 1 repository · arXiv:2403.13518
-
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs 20 Mar 2024 · 0 repositories · arXiv:2403.13801
-
PARAMANU-AYN: Pretrain from scratch or Continual Pretraining of LLMs for Legal Domain Adaptation? 20 Mar 2024 · 0 repositories · arXiv:2403.13681
-
Automated Data Curation for Robust Language Model Fine-Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.12776
-
Automatic Summarization of Doctor-Patient Encounter Dialogues Using Large Language Model through Prompt Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.13089
-
Can AI Outperform Human Experts in Creating Social Media Creatives? 19 Mar 2024 · 0 repositories · arXiv:2404.00018
-
Fine-Tuning Pre-trained Language Models to Detect In-Game Trash Talks 19 Mar 2024 · 0 repositories · arXiv:2403.15458
-
Instructing Large Language Models to Identify and Ignore Irrelevant Conditions 19 Mar 2024 · 1 repository · arXiv:2403.12744
-
Pipelined Biomedical Event Extraction Rivaling Joint Learning 19 Mar 2024 · 0 repositories · arXiv:2403.12386
-
TT-BLIP: Enhancing Fake News Detection Using BLIP and Tri-Transformer 19 Mar 2024 · 0 repositories · arXiv:2403.12481
-
A Disease Labeler for Chinese Chest X-Ray Report Generation 18 Mar 2024 · 0 repositories · arXiv:2404.16852
-
CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification 18 Mar 2024 · 1 repository · arXiv:2403.11904
-
Construction of Hyper-Relational Knowledge Graphs Using Pre-Trained Large Language Models 18 Mar 2024 · 0 repositories · arXiv:2403.11786
-
EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.12171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Embedded Named Entity Recognition using Probing Classifiers 18 Mar 2024 · 2 repositories · arXiv:2403.11747Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Embracing the Generative AI Revolution: Advancing Tertiary Education in Cybersecurity with GPT 18 Mar 2024 · 0 repositories · arXiv:2403.11402
-
Ensuring Safe and High-Quality Outputs: A Guideline Library Approach for Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11838
-
Evaluating Named Entity Recognition: A comparative analysis of mono- and multilingual transformer models on a novel Brazilian corporate earnings call transcripts dataset 18 Mar 2024 · 2 repositories · arXiv:2403.12212
-
GPT-4 as Evaluator: Evaluating Large Language Models on Pest Management in Agriculture 18 Mar 2024 · 0 repositories · arXiv:2403.11858
-
HateCOT: An Explanation-Enhanced Dataset for Generalizable Offensive Speech Detection via Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11456
-
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments 18 Mar 2024 · 1 repository · arXiv:2403.11807Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 2 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Meta-Prompting for Automating Zero-shot Visual Recognition with LLMs 18 Mar 2024 · 1 repository · arXiv:2403.11755Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Metaphor Understanding Challenge Dataset for LLMs 18 Mar 2024 · 0 repositories · arXiv:2403.11810
-
Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure 18 Mar 2024 · 1 repository · arXiv:2403.11425
-
Leveraging Large Language Models to Detect npm Malicious Packages 18 Mar 2024 · 0 repositories · arXiv:2403.12196
-
Data is all you need: Finetuning LLMs for Chip Design via an Automated design-data augmentation framework 17 Mar 2024 · 1 repository · arXiv:2403.11202Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Forging the Forger: An Attempt to Improve Authorship Verification via Data Augmentation 17 Mar 2024 · 0 repositories · arXiv:2403.11265
-
JORA: JAX Tensor-Parallel LoRA Library for Retrieval Augmented Fine-Tuning 17 Mar 2024 · 1 repository · arXiv:2403.11366
-
Can Large Language Models abstract Medical Coded Language? 16 Mar 2024 · 0 repositories · arXiv:2403.10822
-
Empirical Studies of Parameter Efficient Methods for Large Language Models of Code and Knowledge Transfer to R 16 Mar 2024 · 1 repository · arXiv:2405.01553
-
From Melting Pots to Misrepresentations: Exploring Harms in Generative AI 16 Mar 2024 · 0 repositories · arXiv:2403.10776
-
Large language model-powered chatbots for internationalizing student support in higher education 16 Mar 2024 · 0 repositories · arXiv:2403.14702
-
Application of GPT Language Models for Innovation in Activities in University Teaching 15 Mar 2024 · 0 repositories · arXiv:2403.14694
-
DRAGIN: Dynamic Retrieval Augmented Generation based on the Information Needs of Large Language Models 15 Mar 2024 · 1 repository · arXiv:2403.10081Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Enhancing LLM Factual Accuracy with RAG to Counter Hallucinations: A Case Study on Domain-Specific Queries in Private Knowledge-Bases 15 Mar 2024 · 1 repository · arXiv:2403.10446Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
ExeGPT: Constraint-Aware Resource Scheduling for LLM Inference 15 Mar 2024 · 0 repositories · arXiv:2404.07947
-
Knowledge Condensation and Reasoning for Knowledge-based VQA 15 Mar 2024 · 0 repositories · arXiv:2403.10037
-
RAFT: Adapting Language Model to Domain Specific RAG 15 Mar 2024 · 1 repository · arXiv:2403.10131
-
Repoformer: Selective Retrieval for Repository-Level Code Completion 15 Mar 2024 · 0 repositories · arXiv:2403.10059
-
ViTCN: Vision Transformer Contrastive Network For Reasoning 15 Mar 2024 · 0 repositories · arXiv:2403.09962
-
AI on AI: Exploring the Utility of GPT as an Expert Annotator of AI Publications 14 Mar 2024 · 0 repositories · arXiv:2403.09097
-
Basque and Spanish Counter Narrative Generation: Data Creation and Evaluation 14 Mar 2024 · 0 repositories · arXiv:2403.09159
-
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences 14 Mar 2024 · 2 repositories · arXiv:2403.09032Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating LLMs for Gender Disparities in Notable Persons 14 Mar 2024 · 0 repositories · arXiv:2403.09148
-
Fisher Mask Nodes for Language Model Merging 14 Mar 2024 · 1 repository · arXiv:2403.09891
-
Incorporating Graph Attention Mechanism into Geometric Problem Solving Based on Deep Reinforcement Learning 14 Mar 2024 · 1 repository · arXiv:2403.14690
-
Information Extraction: An application to the domain of hyper-local financial data on developing countries 14 Mar 2024 · 0 repositories · arXiv:2403.09077
-
Komodo: A Linguistic Expedition into Indonesia's Regional Languages 14 Mar 2024 · 0 repositories · arXiv:2403.09362
-
Leap: molecular synthesisability scoring with intermediates 14 Mar 2024 · 0 repositories · arXiv:2403.13005
-
Optimistic Verifiable Training by Controlling Hardware Nondeterminism 14 Mar 2024 · 1 repository · arXiv:2403.09603Syntology official (archive's flag): 9 ran · 9 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
RAGGED: Towards Informed Design of Retrieval Augmented Generation Systems 14 Mar 2024 · 1 repository · arXiv:2403.09040Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
Rectifying Demonstration Shortcut in In-Context Learning 14 Mar 2024 · 1 repository · arXiv:2403.09488Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Retrieval augmented text-to-SQL generation for epidemiological question answering using electronic health records 14 Mar 2024 · 1 repository · arXiv:2403.09226
-
Sabiá-2: A New Generation of Portuguese Large Language Models 14 Mar 2024 · 0 repositories · arXiv:2403.09887
-
Autoregressive Score Generation for Multi-trait Essay Scoring 13 Mar 2024 · 1 repository · arXiv:2403.08332
-
Distilling Named Entity Recognition Models for Endangered Species from Large Language Models 13 Mar 2024 · 0 repositories · arXiv:2403.15430
-
Do Language Models Care About Text Quality? Evaluating Web-Crawled Corpora Across 11 Languages 13 Mar 2024 · 0 repositories · arXiv:2403.08693
-
Embedded Translations for Low-resource Automated Glossing 13 Mar 2024 · 0 repositories · arXiv:2403.08189
-
Generative Pretrained Structured Transformers: Unsupervised Syntactic Language Models at Scale 13 Mar 2024 · 2 repositories · arXiv:2403.08293Syntology official (archive's flag): 2 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Research on the Application of Deep Learning-based BERT Model in Sentiment Analysis 13 Mar 2024 · 0 repositories · arXiv:2403.08217
-
Rich Semantic Knowledge Enhanced Large Language Models for Few-shot Chinese Spell Checking 13 Mar 2024 · 0 repositories · arXiv:2403.08492
-
Chronos: Learning the Language of Time Series 12 Mar 2024 · 6 repositories · arXiv:2403.07815Syntology official (archive's flag): 8 ran · 23 ran (of which 0 constructed an object rather than computing a result; 22 with no instrument failure: 3 honoured, 1 violated, 18 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 28 harvested samples) · 5 pointer-only (licence)
-
Contextual Clarity: Generating Sentences with Transformer Models using Context-Reverso Data 12 Mar 2024 · 1 repository · arXiv:2403.08103
-
Enhancing Readmission Prediction with Deep Learning: Extracting Biomedical Concepts from Clinical Texts 12 Mar 2024 · 0 repositories · arXiv:2403.09722
-
GPT-generated Text Detection: Benchmark Dataset and Tensor-based Detection Method 12 Mar 2024 · 1 repository · arXiv:2403.07321Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Investigating the performance of Retrieval-Augmented Generation and fine-tuning for the development of AI-driven knowledge-based systems 12 Mar 2024 · 1 repository · arXiv:2403.09727
-
LookupFFN: Making Transformers Compute-lite for CPU inference 12 Mar 2024 · 1 repository · arXiv:2403.07221Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
MoralBERT: A Fine-Tuned Language Model for Capturing Moral Values in Social Discussions 12 Mar 2024 · 1 repository · arXiv:2403.07678
-
Rethinking ASTE: A Minimalist Tagging Scheme Alongside Contrastive Learning 12 Mar 2024 · 0 repositories · arXiv:2403.07342
-
Rethinking Generative Large Language Model Evaluation for Semantic Comprehension 12 Mar 2024 · 0 repositories · arXiv:2403.07872
-
SIFiD: Reassess Summary Factual Inconsistency Detection with LLM 12 Mar 2024 · 0 repositories · arXiv:2403.07557
-
The future of document indexing: GPT and Donut revolutionize table of content processing 12 Mar 2024 · 0 repositories · arXiv:2403.07553
-
A multi-cohort study on prediction of acute brain dysfunction states using selective state space models 11 Mar 2024 · 0 repositories · arXiv:2403.07201
-
Development of a Reliable and Accessible Caregiving Language Model (CaLM) 11 Mar 2024 · 0 repositories · arXiv:2403.06857