Methods › General › Regularization › Attention Dropout › Papers, page 57
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 57 of 109: papers 5,601 to 5,700 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Multi-class Categorization of Reasons behind Mental Disturbance in Long Texts 8 Apr 2023 · 0 repositories · arXiv:2304.04118
-
tmn at SemEval-2023 Task 9: Multilingual Tweet Intimacy Detection using XLM-T, Google Translate, and Ensemble Learning 8 Apr 2023 · 1 repository · arXiv:2304.04054
-
Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4 7 Apr 2023 · 1 repository · arXiv:2304.03439
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
Deep Learning for Opinion Mining and Topic Classification of Course Reviews 6 Apr 2023 · 0 repositories · arXiv:2304.03394
-
GPT detectors are biased against non-native English writers 6 Apr 2023 · 2 repositories · arXiv:2304.02819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Making AI Less "Thirsty": Uncovering and Addressing the Secret Water Footprint of AI Models 6 Apr 2023 · 1 repository · arXiv:2304.03271Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Micron-BERT: BERT-based Facial Micro-Expression Recognition 6 Apr 2023 · 1 repository · arXiv:2304.03195Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Multi-label classification of open-ended questions with BERT 6 Apr 2023 · 0 repositories · arXiv:2304.02945
-
Towards Interpretable Mental Health Analysis with Large Language Models 6 Apr 2023 · 2 repositories · arXiv:2304.03347
-
Zero-Shot Next-Item Recommendation using Large Pretrained Language Models 6 Apr 2023 · 1 repository · arXiv:2304.03153Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Conceptual structure coheres in human cognition but not in large language models 5 Apr 2023 · 0 repositories · arXiv:2304.02754
-
Bengali Fake Review Detection using Semi-supervised Generative Adversarial Networks 5 Apr 2023 · 0 repositories · arXiv:2304.02739
-
ChartReader: A Unified Framework for Chart Derendering and Comprehension without Heuristic Rules 5 Apr 2023 · 1 repository · arXiv:2304.02173Syntology official (archive's flag): 16 ran · 16 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 10 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Context-Aware Classification of Legal Document Pages 5 Apr 2023 · 0 repositories · arXiv:2304.02787
-
Document-Level Machine Translation with Large Language Models 5 Apr 2023 · 1 repository · arXiv:2304.02210
-
Large Language Models as Master Key: Unlocking the Secrets of Materials Science with GPT 5 Apr 2023 · 0 repositories · arXiv:2304.02213
-
Blockwise Compression of Transformer-based Models without Retraining 4 Apr 2023 · 0 repositories · arXiv:2304.01483
-
Generating Natural Language from Logic Expressions with Structural Representation 4 Apr 2023 · 1 repository
-
Geotechnical Parrot Tales (GPT): Harnessing Large Language Models in geotechnical engineering 4 Apr 2023 · 0 repositories · arXiv:2304.02138
-
GPT-4 to GPT-3.5: 'Hold My Scalpel' -- A Look at the Competency of OpenAI's GPT on the Plastic Surgery In-Service Training Exam 4 Apr 2023 · 0 repositories · arXiv:2304.01503
-
Improved Visual Fine-tuning with Natural Language Supervision 4 Apr 2023 · 1 repository · arXiv:2304.01489Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation 4 Apr 2023 · 0 repositories · arXiv:2304.01746
-
LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models 4 Apr 2023 · 2 repositories · arXiv:2304.01933Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
REFINER: Reasoning Feedback on Intermediate Representations 4 Apr 2023 · 1 repository · arXiv:2304.01904Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
San-BERT: Extractive Summarization for Sanskrit Documents using BERT and it's variants 4 Apr 2023 · 0 repositories · arXiv:2304.01894
-
Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models 4 Apr 2023 · 0 repositories · arXiv:2304.01852
-
Detection of Homophobia & Transphobia in Dravidian Languages: Exploring Deep Learning Methods 3 Apr 2023 · 0 repositories · arXiv:2304.01241
-
GreekBART: The First Pretrained Greek Sequence-to-Sequence Model 3 Apr 2023 · 2 repositories · arXiv:2304.00869
-
Hate Speech Targets Detection in Parler using BERT 3 Apr 2023 · 1 repository · arXiv:2304.01179
-
MiniRBT: A Two-stage Distilled Small Chinese Pre-trained Model 3 Apr 2023 · 1 repository · arXiv:2304.00717
-
PEACH: Pre-Training Sequence-to-Sequence Multilingual Models for Translation with Semi-Supervised Pseudo-Parallel Document Generation 3 Apr 2023 · 1 repository · arXiv:2304.01282
-
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT 3 Apr 2023 · 2 repositories · arXiv:2304.01246
-
The StatCan Dialogue Dataset: Retrieving Data Tables through Conversations with Genuine Intents 3 Apr 2023 · 1 repository · arXiv:2304.01412
-
Does Human Collaboration Enhance the Accuracy of Identifying LLM-Generated Deepfake Texts? 3 Apr 2023 · 2 repositories · arXiv:2304.01002Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Better Language Models of Code through Self-Improvement 2 Apr 2023 · 1 repository · arXiv:2304.01228
-
Classifying COVID-19 Related Tweets for Fake News Detection and Sentiment Analysis with BERT-based Models 2 Apr 2023 · 0 repositories · arXiv:2304.00636
-
LLMMaps -- A Visual Metaphor for Stratified Evaluation of Large Language Models 2 Apr 2023 · 1 repository · arXiv:2304.00457
-
The Other Side of Compression: Measuring Bias in Pruned Transformers 1 Apr 2023 · 1 repository
-
BERTino: an Italian DistilBERT model 31 Mar 2023 · 1 repository · arXiv:2303.18121
-
Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations 31 Mar 2023 · 1 repository · arXiv:2303.18027
-
GPT-4 can pass the Korean National Licensing Examination for Korean Medicine Doctors 31 Mar 2023 · 0 repositories · arXiv:2303.17807
-
Extracting Thyroid Nodules Characteristics from Ultrasound Reports Using Transformer-based Natural Language Processing Methods 31 Mar 2023 · 0 repositories · arXiv:2304.00115
-
JobHam-place with smart recommend job options and candidate filtering options 31 Mar 2023 · 0 repositories · arXiv:2303.17930
-
Quick Dense Retrievers Consume KALE: Post Training Kullback Leibler Alignment of Embeddings for Asymmetrical dual encoders 31 Mar 2023 · 0 repositories · arXiv:2304.01016
-
Aligning a medium-size GPT model in English to a small closed domain in Spanish 30 Mar 2023 · 0 repositories · arXiv:2303.17649
-
Evaluation of GPT and BERT-based models on identifying protein-protein interactions in biomedical text 30 Mar 2023 · 0 repositories · arXiv:2303.17728
-
Fine-Tuning BERT with Character-Level Noise for Zero-Shot Transfer to Dialects and Closely-Related Languages 30 Mar 2023 · 0 repositories · arXiv:2303.17683
-
Humans in Humans Out: On GPT Converging Toward Common Sense in both Success and Failure 30 Mar 2023 · 0 repositories · arXiv:2303.17276
-
oBERTa: Improving Sparse Transfer Learning via improved initialization, distillation, and pruning regimes 30 Mar 2023 · 0 repositories · arXiv:2303.17612
-
Synthesis of Mathematical programs from Natural Language Specifications 30 Mar 2023 · 0 repositories · arXiv:2304.03287
-
Advances in apparent conceptual physics reasoning in GPT-4 29 Mar 2023 · 0 repositories · arXiv:2303.17012
-
AnnoLLM: Making Large Language Models to Be Better Crowdsourced Annotators 29 Mar 2023 · 2 repositories · arXiv:2303.16854Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
AutoAD: Movie Description in Context 29 Mar 2023 · 1 repository · arXiv:2303.16899Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 8 unverified (of 13 harvested samples)
-
BERT4ETH: A Pre-trained Transformer for Ethereum Fraud Detection 29 Mar 2023 · 1 repository · arXiv:2303.18138
-
Evaluating GPT-3.5 and GPT-4 Models on Brazilian University Admission Exams 29 Mar 2023 · 1 repository · arXiv:2303.17003Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
How do decoding algorithms distribute information in dialogue responses? 29 Mar 2023 · 0 repositories · arXiv:2303.17006
-
Larger Probes Tell a Different Story: Extending Psycholinguistic Datasets Via In-Context Learning 29 Mar 2023 · 1 repository · arXiv:2303.16445
-
Summarizing Indian Languages using Multilingual Transformers based Models 29 Mar 2023 · 0 repositories · arXiv:2303.16657
-
Ten Quick Tips for Harnessing the Power of ChatGPT/GPT-4 in Computational Biology 29 Mar 2023 · 1 repository · arXiv:2303.16429
-
ViewRefer: Grasp the Multi-view Knowledge for 3D Visual Grounding with GPT and Prototype Guidance 29 Mar 2023 · 7 repositories · arXiv:2303.16894Syntology official (archive's flag): 4 ran · 8 ran (of which 2 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Improving Large Language Models for Clinical Named Entity Recognition via Prompt Engineering 29 Mar 2023 · 1 repository · arXiv:2303.16416
-
Explicit Planning Helps Language Models in Logical Reasoning 28 Mar 2023 · 2 repositories · arXiv:2303.15714Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
NeuralMind-UNICAMP at 2022 TREC NeuCLIR: Large Boring Rerankers for Cross-lingual Retrieval 28 Mar 2023 · 1 repository · arXiv:2303.16145
-
On Codex Prompt Engineering for OCL Generation: An Empirical Study 28 Mar 2023 · 0 repositories · arXiv:2303.16244
-
One Adapter for All Programming Languages? Adapter Tuning for Code Search and Summarization 28 Mar 2023 · 1 repository · arXiv:2303.15822
-
Zero-Shot Generalizable End-to-End Task-Oriented Dialog System using Context Summarization and Domain Schema 28 Mar 2023 · 1 repository · arXiv:2303.16252
-
KPEval: Towards Fine-Grained Semantic-Based Keyphrase Evaluation 27 Mar 2023 · 1 repository · arXiv:2303.15422
-
TextMI: Textualize Multimodal Information for Integrating Non-verbal Cues in Pre-trained Language Models 27 Mar 2023 · 0 repositories · arXiv:2303.15430
-
Exploring Multimodal Sentiment Analysis via CBAM Attention and Double-layer BiLSTM Architecture 26 Mar 2023 · 0 repositories · arXiv:2303.14708
-
Analyzing the Performance of GPT-3.5 and GPT-4 in Grammatical Error Correction 25 Mar 2023 · 0 repositories · arXiv:2303.14342
-
Automatic Generation of Multiple-Choice Questions 25 Mar 2023 · 0 repositories · arXiv:2303.14576
-
Can Large Language Models assist in Hazard Analysis? 25 Mar 2023 · 0 repositories · arXiv:2303.15473
-
GPT is becoming a Turing machine: Here are some ways to program it 25 Mar 2023 · 0 repositories · arXiv:2303.14310
-
Indonesian Text-to-Image Synthesis with Sentence-BERT and FastGAN 25 Mar 2023 · 1 repository · arXiv:2303.14517
-
Spatio-Temporal driven Attention Graph Neural Network with Block Adjacency matrix (STAG-NN-BA) 25 Mar 2023 · 0 repositories · arXiv:2303.14322
-
Depression detection in social media posts using affective and social norm features 24 Mar 2023 · 0 repositories · arXiv:2303.14279
-
"Get ready for a party": Exploring smarter smart spaces with help from large language models 24 Mar 2023 · 1 repository · arXiv:2303.14143
-
Personalizing Task-oriented Dialog Systems via Zero-shot Generalizable Reward Function 24 Mar 2023 · 0 repositories · arXiv:2303.13797
-
SEAL: Semantic Frame Execution And Localization for Perceiving Afforded Robot Actions 24 Mar 2023 · 0 repositories · arXiv:2303.14067
-
SIGMORPHON 2023 Shared Task of Interlinear Glossing: Baseline Model 24 Mar 2023 · 1 repository · arXiv:2303.14234
-
Toward Open-domain Slot Filling via Self-supervised Co-training 24 Mar 2023 · 0 repositories · arXiv:2303.13801
-
Where to Go Next for Recommender Systems? ID- vs. Modality-based Recommender Models Revisited 24 Mar 2023 · 1 repository · arXiv:2303.13835
-
A Novel Patent Similarity Measurement Methodology: Semantic Distance and Technological Distance 23 Mar 2023 · 1 repository · arXiv:2303.16767
-
Beyond Universal Transformer: block reusing with adaptor in Transformer for automatic speech recognition 23 Mar 2023 · 0 repositories · arXiv:2303.13072
-
DBLP-QuAD: A Question Answering Dataset over the DBLP Scholarly Knowledge Graph 23 Mar 2023 · 1 repository · arXiv:2303.13351
-
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT 23 Mar 2023 · 0 repositories · arXiv:2303.13013
-
GETT-QA: Graph Embedding based T2T Transformer for Knowledge Graph Question Answering 23 Mar 2023 · 1 repository · arXiv:2303.13284
-
Retrieval-Augmented Classification with Decoupled Representation 23 Mar 2023 · 1 repository · arXiv:2303.13065
-
Analyzing the Generalizability of Deep Contextualized Language Representations For Text Classification 22 Mar 2023 · 0 repositories · arXiv:2303.12936
-
Generate labeled training data using Prompt Programming and GPT-3. An example of Big Five Personality Classification 22 Mar 2023 · 0 repositories · arXiv:2303.12279
-
Open-source Frame Semantic Parsing 22 Mar 2023 · 1 repository · arXiv:2303.12788
-
Semantic Communication with Memory 22 Mar 2023 · 0 repositories · arXiv:2303.12335
-
TRON: Transformer Neural Network Acceleration with Non-Coherent Silicon Photonics 22 Mar 2023 · 0 repositories · arXiv:2303.12914
-
A Complete Survey on Generative AI (AIGC): Is ChatGPT from GPT-4 to GPT-5 All You Need? 21 Mar 2023 · 0 repositories · arXiv:2303.11717
-
ChatGPT and a New Academic Reality: Artificial Intelligence-Written Research Papers and the Ethics of the Large Language Models in Scholarly Publishing 21 Mar 2023 · 0 repositories · arXiv:2303.13367
-
cTBLS: Augmenting Large Language Models with Conversational Tables 21 Mar 2023 · 1 repository · arXiv:2303.12024
-
Fine-tuning ClimateBert transformer with ClimaText for the disclosure analysis of climate-related financial risks 21 Mar 2023 · 0 repositories · arXiv:2303.13373
-
Is BERT Blind? Exploring the Effect of Vision-and-Language Pretraining on Visual Language Understanding 21 Mar 2023 · 1 repository · arXiv:2303.12513
-
Learning A Sparse Transformer Network for Effective Image Deraining 21 Mar 2023 · 1 repository · arXiv:2303.11950Syntology official (archive's flag): 8 ran · 8 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 8 samples that ran constructed an object rather than computing a result (of 11 harvested samples) · 11 pointer-only (licence)