Methods › General › Regularization › Dropout › Papers, page 34
Dropout
Papers archive 2025-07-28
archive papers tagged: 27,472 · with a code link: 12,129 · where Syntology ran a sample: 3,620 (3,044 with a run with no instrument failure, 576 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,620 of 27,472 tagged: 3,044 with a run with no instrument failure, 576 where every run was a failure of Syntology's instrument)
Page 34 of 275: papers 3,301 to 3,400 of 27,472, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning 5 Dec 2024 · 1 repository · arXiv:2412.04661
-
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments? 5 Dec 2024 · 0 repositories · arXiv:2412.03856
-
TransAdapter: Vision Transformer for Feature-Centric Unsupervised Domain Adaptation 5 Dec 2024 · 1 repository · arXiv:2412.04073
-
Uniform Discretized Integrated Gradients: An effective attribution based method for explaining large language models 5 Dec 2024 · 0 repositories · arXiv:2412.03886
-
A Water Efficiency Dataset for African Data Centers 4 Dec 2024 · 0 repositories · arXiv:2412.03716
-
Advanced Risk Prediction and Stability Assessment of Banks Using Time Series Transformer Models 4 Dec 2024 · 0 repositories · arXiv:2412.03606
-
Advancing Conversational Psychotherapy: Integrating Privacy, Dual-Memory, and Domain Expertise with Large Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.02987
-
AntLM: Bridging Causal and Masked Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.03275
-
Controlling the Mutation in Large Language Models for the Efficient Evolution of Algorithms 4 Dec 2024 · 0 repositories · arXiv:2412.03250
-
Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts? 4 Dec 2024 · 0 repositories · arXiv:2412.03235
-
EMPATH: MediaPipe-Aided Ensemble Learning with Attention-Based Transformers for Accurate Recognition of Bangla Word-Level Sign Language 4 Dec 2024 · 1 repository
-
FANAL -- Financial Activity News Alerting Language Modeling Framework 4 Dec 2024 · 0 repositories · arXiv:2412.03527
-
Interpreting Transformers for Jet Tagging 4 Dec 2024 · 1 repository · arXiv:2412.03673Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
LEP-QNN: Loan Eligibility Prediction Using Quantum Neural Networks 4 Dec 2024 · 0 repositories · arXiv:2412.03158
-
MaterialPicker: Multi-Modal Material Generation with Diffusion Transformers 4 Dec 2024 · 0 repositories · arXiv:2412.03225
-
Multi-Branch Mutual-Distillation Transformer for EEG-Based Seizure Subtype Classification 4 Dec 2024 · 0 repositories · arXiv:2412.15224
-
Multimodal Sentiment Analysis Based on BERT and ResNet 4 Dec 2024 · 0 repositories · arXiv:2412.03625
-
Navigation World Models 4 Dec 2024 · 1 repository · arXiv:2412.03572Syntology 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention 4 Dec 2024 · 0 repositories · arXiv:2412.03520
-
Theoretical limitations of multi-layer Transformer 4 Dec 2024 · 1 repository · arXiv:2412.02975
-
Achieving Semantic Consistency: Contextualized Word Representations for Political Text Analysis 3 Dec 2024 · 0 repositories · arXiv:2412.04505
-
CAISSON: Concept-Augmented Inference Suite of Self-Organizing Neural Networks 3 Dec 2024 · 0 repositories · arXiv:2412.02835
-
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity 3 Dec 2024 · 0 repositories · arXiv:2412.02252
-
CPTQuant -- A Novel Mixed Precision Post-Training Quantization Techniques for Large Language Models 3 Dec 2024 · 0 repositories · arXiv:2412.03599
-
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators 3 Dec 2024 · 1 repository · arXiv:2412.02467
-
FCL-ViT: Task-Aware Attention Tuning for Continual Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02509
-
Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model 3 Dec 2024 · 0 repositories · arXiv:2412.02802
-
GQWformer: A Quantum-based Transformer for Graph Representation Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02285
-
Gracefully Filtering Backdoor Samples for Generative Large Language Models without Retraining 3 Dec 2024 · 1 repository · arXiv:2412.02454
-
Impact of Data Snooping on Deep Learning Models for Locating Vulnerabilities in Lifted Code 3 Dec 2024 · 0 repositories · arXiv:2412.02048
-
Learn More by Using Less: Distributed Learning with Energy-Constrained Devices 3 Dec 2024 · 0 repositories · arXiv:2412.02289
-
Leveraging Large Language Models for Comparative Literature Summarization with Reflective Incremental Mechanisms 3 Dec 2024 · 0 repositories · arXiv:2412.02149
-
MAGMA: Manifold Regularization for MAEs 3 Dec 2024 · 1 repository · arXiv:2412.02871
-
OCR Hinders RAG: Evaluating the Cascading Impact of OCR on Retrieval-Augmented Generation 3 Dec 2024 · 1 repository · arXiv:2412.02592
-
Optimization of Transformer heart disease prediction model based on particle swarm optimization algorithm 3 Dec 2024 · 0 repositories · arXiv:2412.02801
-
Patent-CR: A Dataset for Patent Claim Revision 3 Dec 2024 · 0 repositories · arXiv:2412.02549
-
RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models 3 Dec 2024 · 1 repository · arXiv:2412.02830
-
Recovering implicit physics model under real-world constraints 3 Dec 2024 · 0 repositories · arXiv:2412.02215
-
Revisiting the Initial Steps in Adaptive Gradient Descent Optimization 3 Dec 2024 · 0 repositories · arXiv:2412.02153
-
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction 3 Dec 2024 · 0 repositories · arXiv:2412.02698
-
Semantic Tokens in Retrieval Augmented Generation 3 Dec 2024 · 0 repositories · arXiv:2412.02563
-
The Asymptotic Behavior of Attention in Transformers 3 Dec 2024 · 0 repositories · arXiv:2412.02682
-
Transformer-Based Auxiliary Loss for Face Recognition Across Age Variations 3 Dec 2024 · 0 repositories · arXiv:2412.02198
-
Automated Extraction of Acronym-Expansion Pairs from Scientific Papers 2 Dec 2024 · 0 repositories · arXiv:2412.01093
-
Convolutional Transformer Neural Collaborative Filtering 2 Dec 2024 · 0 repositories · arXiv:2412.01376
-
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation 2 Dec 2024 · 0 repositories · arXiv:2412.01429
-
FGATT: A Robust Framework for Wireless Data Imputation Using Fuzzy Graph Attention Networks and Transformer Encoders 2 Dec 2024 · 0 repositories · arXiv:2412.01979
-
GETAE: Graph information Enhanced deep neural NeTwork ensemble ArchitecturE for fake news detection 2 Dec 2024 · 1 repository · arXiv:2412.01825
-
Global Average Feature Augmentation for Robust Semantic Segmentation with Transformers 2 Dec 2024 · 0 repositories · arXiv:2412.01941
-
High-Throughput Detection of Risk Factors to Sudden Cardiac Arrest in Youth Athletes: A Smartwatch-Based Screening Platform 2 Dec 2024 · 0 repositories · arXiv:2412.12118
-
Identifying Reliable Predictions in Detection Transformers 2 Dec 2024 · 0 repositories · arXiv:2412.01782
-
MBA-RAG: a Bandit Approach for Adaptive Retrieval-Augmented Generation through Question Complexity 2 Dec 2024 · 1 repository · arXiv:2412.01572Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Multimodal Fusion Learning with Dual Attention for Medical Imaging 2 Dec 2024 · 1 repository · arXiv:2412.01248
-
Mutli-View 3D Reconstruction using Knowledge Distillation 2 Dec 2024 · 1 repository · arXiv:2412.02039
-
NYT-Connections: A Deceptively Simple Text Classification Task that Stumps System-1 Thinkers 2 Dec 2024 · 0 repositories · arXiv:2412.01621
-
PKRD-CoT: A Unified Chain-of-thought Prompting for Multi-Modal Large Language Models in Autonomous Driving 2 Dec 2024 · 0 repositories · arXiv:2412.02025
-
R-Bot: An LLM-based Query Rewrite System 2 Dec 2024 · 0 repositories · arXiv:2412.01661
-
ReHub: Linear Complexity Graph Transformers with Adaptive Hub-Spoke Reassignment 2 Dec 2024 · 0 repositories · arXiv:2412.01519
-
Reliable and scalable variable importance estimation via warm-start and early stopping 2 Dec 2024 · 1 repository · arXiv:2412.01120
-
SiTSE: Sinhala Text Simplification Dataset and Evaluation 2 Dec 2024 · 1 repository · arXiv:2412.01293
-
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models 2 Dec 2024 · 0 repositories · arXiv:2412.01353
-
Swin Transformer with Enhanced Dropout and Layer-wise Unfreezing for Facial Expression Recognition in Mental Health Detection 2 Dec 2024 · 1 repository
-
The Promise and Peril of Generative AI: Evidence from GPT-4 as Sell-Side Analysts 2 Dec 2024 · 0 repositories · arXiv:2412.01069
-
Tokenizing 3D Molecule Structure with Quantized Spherical Coordinates 2 Dec 2024 · 0 repositories · arXiv:2412.01564
-
A Comprehensive Guide to Explainable AI: From Classical Models to LLMs 1 Dec 2024 · 1 repository · arXiv:2412.00800
-
AniMer: Animal Pose and Shape Estimation Using Family Aware Transformer 1 Dec 2024 · 0 repositories · arXiv:2412.00837
-
Categorical Keypoint Positional Embedding for Robust Animal Re-Identification 1 Dec 2024 · 0 repositories · arXiv:2412.00818
-
Decision Transformer vs. Decision Mamba: Analysing the Complexity of Sequential Decision Making in Atari Games 1 Dec 2024 · 1 repository · arXiv:2412.00725
-
DSSRNN: Decomposition-Enhanced State-Space Recurrent Neural Network for Time-Series Analysis 1 Dec 2024 · 1 repository · arXiv:2412.00994
-
EDTformer: An Efficient Decoder Transformer for Visual Place Recognition 1 Dec 2024 · 1 repository · arXiv:2412.00784Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
EventGPT: Event Stream Understanding with Multimodal Large Language Models 1 Dec 2024 · 0 repositories · arXiv:2412.00832
-
Lightweight Contenders: Navigating Semi-Supervised Text Mining through Peer Collaboration and Self Transcendence 1 Dec 2024 · 1 repository · arXiv:2412.00883
-
MIMIC: Multimodal Islamophobic Meme Identification and Classification 1 Dec 2024 · 1 repository · arXiv:2412.00681
-
Pairwise Discernment of AffectNet Expressions with ArcFace 1 Dec 2024 · 0 repositories · arXiv:2412.01860
-
Precise Facial Landmark Detection by Dynamic Semantic Aggregation Transformer 1 Dec 2024 · 1 repository · arXiv:2412.00740
-
TGTOD: A Global Temporal Graph Transformer for Outlier Detection at Scale 1 Dec 2024 · 1 repository · arXiv:2412.00984
-
CDEMapper: Enhancing NIH Common Data Element Normalization using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00491
-
Cognitive Biases in Large Language Models: A Survey and Mitigation Experiments 30 Nov 2024 · 0 repositories · arXiv:2412.00323
-
Does Self-Attention Need Separate Weights in Transformers? 30 Nov 2024 · 0 repositories · arXiv:2412.00359
-
Dynamic Token Selection for Aerial-Ground Person Re-Identification 30 Nov 2024 · 0 repositories · arXiv:2412.00433
-
Empowering the Deaf and Hard of Hearing Community: Enhancing Video Captions Using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00342
-
Enhancing Skin Cancer Diagnosis (SCD) Using Late Discrete Wavelet Transform (DWT) and New Swarm-Based Optimizers 30 Nov 2024 · 1 repository · arXiv:2412.00472
-
Fairness at Every Intersection: Uncovering and Mitigating Intersectional Biases in Multimodal Clinical Predictions 30 Nov 2024 · 0 repositories · arXiv:2412.00606
-
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters 30 Nov 2024 · 0 repositories · arXiv:2412.00530
-
Homeostasis and Sparsity in Transformer 30 Nov 2024 · 0 repositories · arXiv:2412.00503
-
Multi-scale Feature Enhancement in Multi-task Learning for Medical Image Analysis 30 Nov 2024 · 0 repositories · arXiv:2412.00351
-
Advanced System Integration: Analyzing OpenAPI Chunking for Retrieval-Augmented Generation 29 Nov 2024 · 0 repositories · arXiv:2411.19804
-
Dynamic ETF Portfolio Optimization Using enhanced Transformer-Based Models for Covariance and Semi-Covariance Prediction(Work in Progress) 29 Nov 2024 · 0 repositories · arXiv:2411.19649
-
Excretion Detection in Pigsties Using Convolutional and Transformerbased Deep Neural Networks 29 Nov 2024 · 0 repositories · arXiv:2412.00256
-
Generating a Low-code Complete Workflow via Task Decomposition and RAG 29 Nov 2024 · 0 repositories · arXiv:2412.00239
-
Graph Neural Networks for Heart Failure Prediction on an EHR-Based Patient Similarity Graph 29 Nov 2024 · 1 repository · arXiv:2411.19742
-
HVAC-DPT: A Decision Pretrained Transformer for HVAC Control 29 Nov 2024 · 0 repositories · arXiv:2411.19746
-
Know Your RAG: Dataset Taxonomy and Generation Strategies for Evaluating RAG Systems 29 Nov 2024 · 0 repositories · arXiv:2411.19710
-
Knowledge Management for Automobile Failure Analysis Using Graph RAG 29 Nov 2024 · 0 repositories · arXiv:2411.19539
-
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification 29 Nov 2024 · 1 repository · arXiv:2411.19638
-
Multi-task CNN Behavioral Embedding Model For Transaction Fraud Detection 29 Nov 2024 · 0 repositories · arXiv:2411.19457
-
On Domain-Specific Post-Training for Multimodal Large Language Models 29 Nov 2024 · 0 repositories · arXiv:2411.19930
-
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation 29 Nov 2024 · 0 repositories · arXiv:2411.19528
-
RL-MILP Solver: A Reinforcement Learning Approach for Solving Mixed-Integer Linear Programs with Graph Neural Networks 29 Nov 2024 · 0 repositories · arXiv:2411.19517
-
SAT-HMR: Real-Time Multi-Person 3D Mesh Estimation via Scale-Adaptive Tokens 29 Nov 2024 · 0 repositories · arXiv:2411.19824