Methods › General › Regularization › Attention Dropout › Papers, page 39
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 39 of 109: papers 3,801 to 3,900 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
FinLLMs: A Framework for Financial Reasoning Dataset Generation with Large Language Models 19 Jan 2024 · 0 repositories · arXiv:2401.10744
-
LangBridge: Multilingual Reasoning Without Multilingual Supervision 19 Jan 2024 · 1 repository · arXiv:2401.10695Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Mining experimental data from Materials Science literature with Large Language Models: an evaluation study 19 Jan 2024 · 1 repository · arXiv:2401.11052
-
Reinforcement learning for question answering in programming domain using public community scoring as a human feedback 19 Jan 2024 · 0 repositories · arXiv:2401.10882
-
ChatQA: Surpassing GPT-4 on Conversational QA and RAG 18 Jan 2024 · 0 repositories · arXiv:2401.10225
-
Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs 18 Jan 2024 · 1 repository · arXiv:2401.10065Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Gender Bias in Machine Translation and The Era of Large Language Models 18 Jan 2024 · 0 repositories · arXiv:2401.10016
-
Image Translation as Diffusion Visual Programmers 18 Jan 2024 · 0 repositories · arXiv:2401.09742
-
Leveraging Biases in Large Language Models: "bias-kNN'' for Effective Few-Shot Learning 18 Jan 2024 · 0 repositories · arXiv:2401.09783
-
When Neural Code Completion Models Size up the Situation: Attaining Cheaper and Faster Completion through Dynamic Model Inference 18 Jan 2024 · 1 repository · arXiv:2401.09964
-
BERTologyNavigator: Advanced Question Answering with BERT-based Semantics 17 Jan 2024 · 0 repositories · arXiv:2401.09553
-
Deciphering Textual Authenticity: A Generalized Strategy through the Lens of Large Language Semantics for Detecting Human vs. Machine-Generated Text 17 Jan 2024 · 1 repository · arXiv:2401.09407
-
Efficient slot labelling 17 Jan 2024 · 0 repositories · arXiv:2401.09343
-
Improving Classification Performance With Human Feedback: Label a few, we label the rest 17 Jan 2024 · 0 repositories · arXiv:2401.09555
-
Learning from Implicit User Feedback, Emotions and Demographic Information in Task-Oriented and Document-Grounded Dialogues 17 Jan 2024 · 1 repository · arXiv:2401.09248
-
MADA: Meta-Adaptive Optimizers through hyper-gradient Descent 17 Jan 2024 · 0 repositories · arXiv:2401.08893
-
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model 17 Jan 2024 · 15 repositories · arXiv:2401.09417Syntology official: no sample here; runs from other or unrecorded repositories · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
A Reproducibility Study of Goldilocks: Just-Right Tuning of BERT for TAR 16 Jan 2024 · 1 repository · arXiv:2401.08104
-
Application of LLM Agents in Recruitment: A Novel Framework for Resume Screening 16 Jan 2024 · 0 repositories · arXiv:2401.08315
-
Enhancing Robustness of LLM-Synthetic Text Detectors for Academic Writing: A Comprehensive Analysis 16 Jan 2024 · 0 repositories · arXiv:2401.08046
-
Exploiting Inter-Layer Expert Affinity for Accelerating Mixture-of-Experts Model Inference 16 Jan 2024 · 1 repository · arXiv:2401.08383
-
RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture 16 Jan 2024 · 0 repositories · arXiv:2401.08406
-
RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning 16 Jan 2024 · 1 repository · arXiv:2401.08326Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples)
-
Tuning Language Models by Proxy 16 Jan 2024 · 2 repositories · arXiv:2401.08565Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
A character-based steganography using masked language modeling 15 Jan 2024 · 1 repository
-
A Novel Approach for Automatic Program Repair using Round-Trip Translation with Large Language Models 15 Jan 2024 · 1 repository · arXiv:2401.07994
-
Graph database while computationally efficient filters out quickly the ESG integrated equities in investment management 15 Jan 2024 · 0 repositories · arXiv:2401.07483
-
Towards Efficient Methods in Medical Question Answering using Knowledge Graph Embeddings 15 Jan 2024 · 1 repository · arXiv:2401.07977Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Leveraging the power of transformers for guilt detection in text 15 Jan 2024 · 0 repositories · arXiv:2401.07414
-
SemEval-2017 Task 4: Sentiment Analysis in Twitter using BERT 15 Jan 2024 · 1 repository · arXiv:2401.07944
-
The Chronicles of RAG: The Retriever, the Chunk and the Generator 15 Jan 2024 · 0 repositories · arXiv:2401.07883
-
Understanding YTHDF2-mediated mRNA Degradation By m6A-BERT-Deg 15 Jan 2024 · 1 repository · arXiv:2401.08004
-
Harnessing Large Language Models Over Transformer Models for Detecting Bengali Depressive Social Media Text: A Comprehensive Study 14 Jan 2024 · 1 repository · arXiv:2401.07310
-
Learning to be Homo Economicus: Can an LLM Learn Preferences from Choice 14 Jan 2024 · 0 repositories · arXiv:2401.07345
-
MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation 14 Jan 2024 · 0 repositories · arXiv:2401.07314
-
Promptformer: Prompted Conformer Transducer for ASR 14 Jan 2024 · 0 repositories · arXiv:2401.07360
-
Streamlining the Selection Phase of Systematic Literature Reviews (SLRs) Using AI-Enabled GPT-4 Assistant API 14 Jan 2024 · 0 repositories · arXiv:2402.18582
-
A Novel Multi-Stage Prompting Approach for Language Agnostic MCQ Generation using GPT 13 Jan 2024 · 1 repository · arXiv:2401.07098
-
Assessing Large Language Models in Mechanical Engineering Education: A Study on Mechanics-Focused Conceptual Understanding 13 Jan 2024 · 0 repositories · arXiv:2401.12983
-
Bridging the Preference Gap between Retrievers and LLMs 13 Jan 2024 · 0 repositories · arXiv:2401.06954
-
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation 13 Jan 2024 · 0 repositories · arXiv:2401.08694
-
An investigation of structures responsible for gender bias in BERT and DistilBERT 12 Jan 2024 · 0 repositories · arXiv:2401.06495
-
Comparing GPT-4 and Open-Source Language Models in Misinformation Mitigation 12 Jan 2024 · 0 repositories · arXiv:2401.06920
-
Human-AI Collaborative Essay Scoring: A Dual-Process Framework with LLMs 12 Jan 2024 · 1 repository · arXiv:2401.06431
-
How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs 12 Jan 2024 · 2 repositories · arXiv:2401.06373Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Improved Learned Sparse Retrieval with Corpus-Specific Vocabularies 12 Jan 2024 · 1 repository · arXiv:2401.06703
-
Intention Analysis Makes LLMs A Good Jailbreak Defender 12 Jan 2024 · 1 repository · arXiv:2401.06561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation 12 Jan 2024 · 1 repository · arXiv:2401.06583
-
Mission: Impossible Language Models 12 Jan 2024 · 1 repository · arXiv:2401.06416Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
PersianMind: A Cross-Lingual Persian-English Large Language Model 12 Jan 2024 · 0 repositories · arXiv:2401.06466
-
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes 12 Jan 2024 · 1 repository · arXiv:2401.06930
-
Analyzing Regional Impacts of Climate Change using Natural Language Processing Techniques 11 Jan 2024 · 0 repositories · arXiv:2401.06817
-
Investigating Data Contamination for Pre-training Language Models 11 Jan 2024 · 0 repositories · arXiv:2401.06059
-
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs 11 Jan 2024 · 0 repositories · arXiv:2401.05940
-
Prompt-based mental health screening from social media text 11 Jan 2024 · 0 repositories · arXiv:2401.05912
-
The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models 11 Jan 2024 · 1 repository · arXiv:2401.05618Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning 10 Jan 2024 · 1 repository · arXiv:2401.05268Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
I am a Strange Dataset: Metalinguistic Tests for Language Models 10 Jan 2024 · 1 repository · arXiv:2401.05300
-
InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks 10 Jan 2024 · 1 repository · arXiv:2401.05507Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Can Active Label Correction Improve LLM-based Modular AI Systems? 10 Jan 2024 · 0 repositories · arXiv:2401.05467
-
Monte Carlo Tree Search for Recipe Generation using GPT-2 10 Jan 2024 · 0 repositories · arXiv:2401.05199
-
Reinforcement Learning for Optimizing RAG for Domain Chatbots 10 Jan 2024 · 0 repositories · arXiv:2401.06800
-
An Assessment on Comprehending Mental Health through Large Language Models 9 Jan 2024 · 0 repositories · arXiv:2401.04592
-
DepressionEmo: A novel dataset for multilabel classification of depression emotions 9 Jan 2024 · 1 repository · arXiv:2401.04655
-
Fighting Fire with Fire: Adversarial Prompting to Generate a Misinformation Detection Dataset 9 Jan 2024 · 0 repositories · arXiv:2401.04481
-
Language Detection for Transliterated Content 9 Jan 2024 · 0 repositories · arXiv:2401.04619
-
Phishing Website Detection through Multi-Model Analysis of HTML Content 9 Jan 2024 · 0 repositories · arXiv:2401.04820
-
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark 8 Jan 2024 · 1 repository · arXiv:2401.03991
-
Anatomy of Neural Language Models 8 Jan 2024 · 1 repository · arXiv:2401.03797
-
Distortions in Judged Spatial Relations in Large Language Models 8 Jan 2024 · 0 repositories · arXiv:2401.04218
-
Advancing bioinformatics with large language models: components, applications and perspectives 8 Jan 2024 · 0 repositories · arXiv:2401.04155
-
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems 8 Jan 2024 · 1 repository · arXiv:2401.05443
-
Mixtral of Experts 8 Jan 2024 · 6 repositories · arXiv:2401.04088Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
RoBERTurk: Adjusting RoBERTa for Turkish 7 Jan 2024 · 0 repositories · arXiv:2401.03515
-
Exploring Defeasibility in Causal Reasoning 6 Jan 2024 · 0 repositories · arXiv:2401.03183
-
PIXAR: Auto-Regressive Language Modeling in Pixel Space 6 Jan 2024 · 0 repositories · arXiv:2401.03321
-
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors 6 Jan 2024 · 0 repositories · arXiv:2401.03238
-
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding 5 Jan 2024 · 1 repository · arXiv:2401.03003Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Can Large Language Models Understand Molecules? 5 Jan 2024 · 2 repositories · arXiv:2402.00024
-
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 5 Jan 2024 · 1 repository · arXiv:2401.02954
-
Natural Language Programming in Medicine: Administering Evidence Based Clinical Workflows with Autonomous Agents Powered by Generative Large Language Models 5 Jan 2024 · 0 repositories · arXiv:2401.02851
-
German Text Embedding Clustering Benchmark 5 Jan 2024 · 1 repository · arXiv:2401.02709
-
Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks 5 Jan 2024 · 2 repositories · arXiv:2401.02731
-
Are LLMs Robust for Spoken Dialogues? 4 Jan 2024 · 0 repositories · arXiv:2401.02297
-
Beyond Extraction: Contextualising Tabular Data for Efficient Summarisation by Language Models 4 Jan 2024 · 0 repositories · arXiv:2401.02333
-
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages 4 Jan 2024 · 1 repository · arXiv:2401.02254
-
Re-evaluating the Memory-balanced Pipeline Parallelism: BPipe 4 Jan 2024 · 0 repositories · arXiv:2401.02088
-
Text2MDT: Extracting Medical Decision Trees from Medical Texts 4 Jan 2024 · 1 repository · arXiv:2401.02034
-
Studying and Recommending Information Highlighting in Stack Overflow Answers 3 Jan 2024 · 1 repository · arXiv:2401.01472
-
Enhancing Multilingual Information Retrieval in Mixed Human Resources Environments: A RAG Model Implementation for Multicultural Enterprise 3 Jan 2024 · 0 repositories · arXiv:2401.01511
-
The Internet of Things in the Era of Generative AI: Vision and Challenges 3 Jan 2024 · 0 repositories · arXiv:2401.01923
-
Iterative Mask Filling: An Effective Text Augmentation Method Using Masked Language Modeling 3 Jan 2024 · 0 repositories · arXiv:2401.01830
-
MLPs Compass: What is learned when MLPs are combined with PLMs? 3 Jan 2024 · 0 repositories · arXiv:2401.01667
-
Natural Language Processing and Multimodal Stock Price Prediction 3 Jan 2024 · 0 repositories · arXiv:2401.01487
-
Revisiting Zero-Shot Abstractive Summarization in the Era of Large Language Models from the Perspective of Position Bias 3 Jan 2024 · 1 repository · arXiv:2401.01989
-
TPC-ViT: Token Propagation Controller for Efficient Vision Transformer 3 Jan 2024 · 0 repositories · arXiv:2401.01470
-
Towards Robust Semantic Segmentation against Patch-based Attack via Attention Refinement 3 Jan 2024 · 0 repositories · arXiv:2401.01750
-
Vietnamese Poem Generation & The Prospect Of Cross-Language Poem-To-Poem Translation 2 Jan 2024 · 1 repository · arXiv:2401.01078
-
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models 1 Jan 2024 · 1 repository · arXiv:2401.00757Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Computational Framework for Behavioral Assessment of LLM Therapists 1 Jan 2024 · 1 repository · arXiv:2401.00820