Methods › General › Regularization › Attention Dropout › Papers, page 56
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 56 of 109: papers 5,501 to 5,600 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Neural Keyphrase Generation: Analysis and Evaluation 27 Apr 2023 · 0 repositories · arXiv:2304.13883
-
Origin Tracing and Detecting of LLMs 27 Apr 2023 · 0 repositories · arXiv:2304.14072
-
pyBibX -- A Python Library for Bibliometric and Scientometric Analysis Powered with Artificial Intelligence Tools 27 Apr 2023 · 1 repository · arXiv:2304.14516
-
SweCTRL-Mini: a data-transparent Transformer-based large language model for controllable text generation in Swedish 27 Apr 2023 · 1 repository · arXiv:2304.13994
-
Prompting GPT-3.5 for Text-to-SQL with De-semanticization and Skeleton Retrieval 26 Apr 2023 · 0 repositories · arXiv:2304.13301
-
Evaluation of GPT-3.5 and GPT-4 for supporting real-world information needs in healthcare delivery 26 Apr 2023 · 0 repositories · arXiv:2304.13714
-
Exploring the Curious Case of Code Prompts 26 Apr 2023 · 1 repository · arXiv:2304.13250Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Extracting Structured Seed-Mediated Gold Nanorod Growth Procedures from Literature with GPT-3 26 Apr 2023 · 0 repositories · arXiv:2304.13846
-
Fine Tuning with Abnormal Examples 26 Apr 2023 · 0 repositories · arXiv:2304.13783
-
HausaNLP at SemEval-2023 Task 12: Leveraging African Low Resource TweetData for Sentiment Analysis 26 Apr 2023 · 1 repository · arXiv:2304.13634
-
Technical Report: Impact of Position Bias on Language Models in Token Classification 26 Apr 2023 · 2 repositories · arXiv:2304.13567
-
Towards Multi-Modal DBMSs for Seamless Querying of Texts and Tables 26 Apr 2023 · 0 repositories · arXiv:2304.13559
-
Introducing MBIB -- the first Media Bias Identification Benchmark Task and Dataset Collection 25 Apr 2023 · 1 repository · arXiv:2304.13148
-
Measuring Massive Multitask Chinese Understanding 25 Apr 2023 · 2 repositories · arXiv:2304.12986Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
NLP-LTU at SemEval-2023 Task 10: The Impact of Data Augmentation and Semi-Supervised Learning Techniques on Text Classification Performance on an Imbalanced Dataset 25 Apr 2023 · 0 repositories · arXiv:2304.12847
-
Semantic Compression With Large Language Models 25 Apr 2023 · 0 repositories · arXiv:2304.12512
-
The Potential of Visual ChatGPT For Remote Sensing 25 Apr 2023 · 0 repositories · arXiv:2304.13009
-
What does BERT learn about prosody? 25 Apr 2023 · 0 repositories · arXiv:2304.12706
-
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM 24 Apr 2023 · 1 repository · arXiv:2304.11872
-
PARAGRAPH2GRAPH: A GNN-based framework for layout paragraph analysis 24 Apr 2023 · 1 repository · arXiv:2304.11810
-
Pre-trained Embeddings for Entity Resolution: An Experimental Analysis [Experiment, Analysis & Benchmark] 24 Apr 2023 · 1 repository · arXiv:2304.12329
-
SocialDial: A Benchmark for Socially-Aware Dialogue Systems 24 Apr 2023 · 1 repository · arXiv:2304.12026
-
Text-to-Audio Generation using Instruction-Tuned LLM and Latent Diffusion Model 24 Apr 2023 · 1 repository · arXiv:2304.13731
-
Processing Natural Language on Embedded Devices: How Well Do Transformer Models Perform? 23 Apr 2023 · 2 repositories · arXiv:2304.11520
-
Boosting Theory-of-Mind Performance in Large Language Models via Prompting 22 Apr 2023 · 1 repository · arXiv:2304.11490
-
L3Cube-IndicSBERT: A simple approach for learning cross-lingual sentence representations using multilingual BERT 22 Apr 2023 · 0 repositories · arXiv:2304.11434
-
A Group-Specific Approach to NLP for Hate Speech Detection 21 Apr 2023 · 1 repository · arXiv:2304.11223
-
BERT Based Clinical Knowledge Extraction for Biomedical Knowledge Graph Construction and Analysis 21 Apr 2023 · 0 repositories · arXiv:2304.10996
-
Building Multimodal AI Chatbots 21 Apr 2023 · 1 repository · arXiv:2305.03512
-
Evaluating Transformer Language Models on Arithmetic Operations Using Number Decomposition 21 Apr 2023 · 1 repository · arXiv:2304.10977
-
Inducing anxiety in large language models can induce bias 21 Apr 2023 · 0 repositories · arXiv:2304.11111
-
Multi-Modal Deep Learning for Credit Rating Prediction Using Text and Numerical Data Streams 21 Apr 2023 · 1 repository · arXiv:2304.10740
-
Text2Time: Transformer-based Article Time Period Prediction 21 Apr 2023 · 0 repositories · arXiv:2304.10859
-
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination 21 Apr 2023 · 0 repositories · arXiv:2304.14347
-
Who's the Best Detective? LLMs vs. MLs in Detecting Incoherent Fourth Grade Math Answers 21 Apr 2023 · 0 repositories · arXiv:2304.11257
-
Domain-specific Continued Pretraining of Language Models for Capturing Long Context in Mental Health 20 Apr 2023 · 0 repositories · arXiv:2304.10447
-
Is Cross-modal Information Retrieval Possible without Training? 20 Apr 2023 · 0 repositories · arXiv:2304.11095
-
Meta Semantics: Towards better natural language understanding and reasoning 20 Apr 2023 · 0 repositories · arXiv:2304.10663
-
Movie Box Office Prediction With Self-Supervised and Visually Grounded Pretraining 20 Apr 2023 · 0 repositories · arXiv:2304.10311
-
Safety Assessment of Chinese Large Language Models 20 Apr 2023 · 2 repositories · arXiv:2304.10436
-
SINC: Spatial Composition of 3D Human Motions for Simultaneous Action Generation 20 Apr 2023 · 0 repositories · arXiv:2304.10417
-
Word Sense Induction with Knowledge Distillation from BERT 20 Apr 2023 · 0 repositories · arXiv:2304.10642
-
Catch Me If You Can: Identifying Fraudulent Physician Reviews with Large Language Models Using Generative Pre-Trained Transformers 19 Apr 2023 · 0 repositories · arXiv:2304.09948
-
GeneGPT: Augmenting Large Language Models with Domain Tools for Improved Access to Biomedical Information 19 Apr 2023 · 1 repository · arXiv:2304.09667Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
How to Do Things with Deep Learning Code 19 Apr 2023 · 0 repositories · arXiv:2304.09406
-
Scaling Transformer to 1M tokens and beyond with RMT 19 Apr 2023 · 3 repositories · arXiv:2304.11062Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Supporting Human-AI Collaboration in Auditing LLMs with LLMs 19 Apr 2023 · 0 repositories · arXiv:2304.09991
-
SurgicalGPT: End-to-End Language-Vision GPT for Visual Question Answering in Surgery 19 Apr 2023 · 1 repository · arXiv:2304.09974Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TieFake: Title-Text Similarity and Emotion-Aware Fake News Detection 19 Apr 2023 · 2 repositories · arXiv:2304.09421
-
BIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information Retrieval 18 Apr 2023 · 0 repositories · arXiv:2304.09333
-
CancerGPT: Few-shot Drug Pair Synergy Prediction using Large Pre-trained Language Models 18 Apr 2023 · 0 repositories · arXiv:2304.10946
-
LLM-based Interaction for Content Generation: A Case Study on the Perception of Employees in an IT department 18 Apr 2023 · 0 repositories · arXiv:2304.09064
-
Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling 18 Apr 2023 · 1 repository · arXiv:2304.09145Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
An Empirical Study of Multitask Learning to Improve Open Domain Dialogue Systems 17 Apr 2023 · 1 repository · arXiv:2304.08115
-
Context-Dependent Embedding Utterance Representations for Emotion Recognition in Conversations 17 Apr 2023 · 1 repository · arXiv:2304.08216
-
From Zero to Hero: Examining the Power of Symbolic Tasks in Instruction Tuning 17 Apr 2023 · 1 repository · arXiv:2304.07995Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
InstructUIE: Multi-task Instruction Tuning for Unified Information Extraction 17 Apr 2023 · 1 repository · arXiv:2304.08085
-
LongForm: Effective Instruction Tuning with Reverse Instructions 17 Apr 2023 · 2 repositories · arXiv:2304.08460
-
Multimodal Short Video Rumor Detection System Based on Contrastive Learning 17 Apr 2023 · 0 repositories · arXiv:2304.08401
-
New Product Development (NPD) through Social Media-based Analysis by Comparing Word2Vec and BERT Word Embeddings 17 Apr 2023 · 0 repositories · arXiv:2304.08369
-
Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding 17 Apr 2023 · 0 repositories · arXiv:2304.10548
-
The MiniPile Challenge for Data-Efficient Language Models 17 Apr 2023 · 1 repository · arXiv:2304.08442
-
A Virtual Simulation-Pilot Agent for Training of Air Traffic Controllers 16 Apr 2023 · 0 repositories · arXiv:2304.07842
-
ArguGPT: evaluating, understanding and identifying argumentative essays generated by GPT models 16 Apr 2023 · 2 repositories · arXiv:2304.07666
-
Enhancing Automated Program Repair through Fine-tuning and Prompt Engineering 16 Apr 2023 · 0 repositories · arXiv:2304.07840
-
Sabiá: Portuguese Large Language Models 16 Apr 2023 · 0 repositories · arXiv:2304.07880
-
SikuGPT: A Generative Pre-trained Model for Intelligent Information Processing of Ancient Texts from the Perspective of Digital Humanities 16 Apr 2023 · 1 repository · arXiv:2304.07778
-
Towards Better Instruction Following Language Models for Chinese: Investigating the Impact of Training Data and Evaluation 16 Apr 2023 · 2 repositories · arXiv:2304.07854
-
Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models 15 Apr 2023 · 0 repositories · arXiv:2304.07619
-
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs 14 Apr 2023 · 2 repositories · arXiv:2304.08244Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
ChatGPT: Applications, Opportunities, and Threats 14 Apr 2023 · 0 repositories · arXiv:2304.09103
-
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data 14 Apr 2023 · 1 repository · arXiv:2304.08247Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
SimpLex: a lexical text simplification architecture 14 Apr 2023 · 1 repository · arXiv:2304.07002
-
Stochastic Code Generation 14 Apr 2023 · 0 repositories · arXiv:2304.08243
-
Automated Mapping of CVE Vulnerability Records to MITRE CWE Weaknesses 13 Apr 2023 · 0 repositories · arXiv:2304.11130
-
ChatGPT cites the most-cited articles and journals, relying solely on Google Scholar's citation counts. As a result, AI may amplify the Matthew Effect in environmental science 13 Apr 2023 · 0 repositories · arXiv:2304.06794
-
Evaluation of Social Biases in Recent Large Pre-Trained Models 13 Apr 2023 · 0 repositories · arXiv:2304.06861
-
PGTask: Introducing the Task of Profile Generation from Dialogues 13 Apr 2023 · 1 repository · arXiv:2304.06634
-
Shall We Pretrain Autoregressive Language Models with Retrieval? A Comprehensive Study 13 Apr 2023 · 1 repository · arXiv:2304.06762
-
What does CLIP know about a red circle? Visual prompt engineering for VLMs 13 Apr 2023 · 0 repositories · arXiv:2304.06712
-
Detection of Fake Generated Scientific Abstracts 12 Apr 2023 · 1 repository · arXiv:2304.06148
-
Evaluation of ChatGPT Model for Vulnerability Detection 12 Apr 2023 · 0 repositories · arXiv:2304.07232
-
Localizing Model Behavior with Path Patching 12 Apr 2023 · 1 repository · arXiv:2304.05969Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Approximating Online Human Evaluation of Social Chatbots with Prompting 11 Apr 2023 · 0 repositories · arXiv:2304.05253
-
Bayesian Optimization of Catalysis With In-Context Learning 11 Apr 2023 · 2 repositories · arXiv:2304.05341
-
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis 11 Apr 2023 · 0 repositories · arXiv:2304.05534
-
Exploring the Use of Foundation Models for Named Entity Recognition and Lemmatization Tasks in Slavic Languages 11 Apr 2023 · 0 repositories · arXiv:2304.05336
-
Multi-step Jailbreaking Privacy Attacks on ChatGPT 11 Apr 2023 · 1 repository · arXiv:2304.05197Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Towards preserving word order importance through Forced Invalidation 11 Apr 2023 · 1 repository · arXiv:2304.05221
-
Training Large Language Models Efficiently with Sparsity and Dataflow 11 Apr 2023 · 0 repositories · arXiv:2304.05511
-
Automated Reading Passage Generation with OpenAI's Large Language Model 10 Apr 2023 · 0 repositories · arXiv:2304.04616
-
Incorporating Structured Sentences with Time-enhanced BERT for Fully-inductive Temporal Relation Prediction 10 Apr 2023 · 0 repositories · arXiv:2304.04717
-
Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study 10 Apr 2023 · 1 repository · arXiv:2304.04339Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
On the Possibilities of AI-Generated Text Detection 10 Apr 2023 · 0 repositories · arXiv:2304.04736
-
Are Large Language Models Ready for Healthcare? A Comparative Study on Clinical Language Understanding 9 Apr 2023 · 1 repository · arXiv:2304.05368
-
Learning to Tokenize for Generative Retrieval 9 Apr 2023 · 1 repository · arXiv:2304.04171
-
Factify 2: A Multimodal Fake News and Satire News Dataset 8 Apr 2023 · 1 repository · arXiv:2304.03897
-
FlexMoE: Scaling Large-scale Sparse Pre-trained Model Training via Dynamic Device Placement 8 Apr 2023 · 0 repositories · arXiv:2304.03946
-
GPT4Rec: A Generative Framework for Personalized Recommendation and User Interests Interpretation 8 Apr 2023 · 0 repositories · arXiv:2304.03879
-
Interpretable Multi Labeled Bengali Toxic Comments Classification using Deep Learning 8 Apr 2023 · 1 repository · arXiv:2304.04087