Methods › General › Regularization › Dropout › Papers, page 28
Dropout
Papers archive 2025-07-28
archive papers tagged: 27,472 · with a code link: 12,129 · where Syntology ran a sample: 3,620 (3,044 with a run with no instrument failure, 576 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,620 of 27,472 tagged: 3,044 with a run with no instrument failure, 576 where every run was a failure of Syntology's instrument)
Page 28 of 275: papers 2,701 to 2,800 of 27,472, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SpecTf: Transformers Enable Data-Driven Imaging Spectroscopy Cloud Detection 9 Jan 2025 · 1 repository · arXiv:2501.04916
-
The dynamics of meaning through time: Assessment of Large Language Models 9 Jan 2025 · 0 repositories · arXiv:2501.05552
-
The more polypersonal the better -- a short look on space geometry of fine-tuned layers 9 Jan 2025 · 0 repositories · arXiv:2501.05503
-
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation 9 Jan 2025 · 1 repository · arXiv:2501.05014
-
Advancing Retrieval-Augmented Generation for Persian: Development of Language Models, Comprehensive Benchmarks, and Best Practices for Optimization 8 Jan 2025 · 0 repositories · arXiv:2501.04858
-
Circuit Complexity Bounds for Visual Autoregressive Model 8 Jan 2025 · 0 repositories · arXiv:2501.04299
-
Comparison of Neural Models for X-ray Image Classification in COVID-19 Detection 8 Jan 2025 · 0 repositories · arXiv:2501.04196
-
Integrating LLMs with ITS: Recent Advances, Potentials, Challenges, and Future Directions 8 Jan 2025 · 0 repositories · arXiv:2501.04437
-
Knowledge Retrieval Based on Generative AI 8 Jan 2025 · 0 repositories · arXiv:2501.04635
-
MB-TaylorFormer V2: Improved Multi-branch Linear Transformer Expanded by Taylor Formula for Image Restoration 8 Jan 2025 · 2 repositories · arXiv:2501.04486
-
Multi-task retriever fine-tuning for domain-specific and efficient RAG 8 Jan 2025 · 0 repositories · arXiv:2501.04652
-
On weight and variance uncertainty in neural networks for regression tasks 8 Jan 2025 · 1 repository · arXiv:2501.04272
-
Quantum-inspired Embeddings Projection and Similarity Metrics for Representation Learning 8 Jan 2025 · 1 repository · arXiv:2501.04591
-
Re-ranking the Context for Multimodal Retrieval Augmented Generation 8 Jan 2025 · 0 repositories · arXiv:2501.04695
-
Scaling Large Language Model Training on Frontier with Low-Bandwidth Partitioning 8 Jan 2025 · 0 repositories · arXiv:2501.04266
-
AuxDepthNet: Real-Time Monocular 3D Object Detection with Depth-Sensitive Features 7 Jan 2025 · 0 repositories · arXiv:2501.03700
-
CFFormer: Cross CNN-Transformer Channel Attention and Spatial Feature Fusion for Improved Segmentation of Low Quality Medical Images 7 Jan 2025 · 0 repositories · arXiv:2501.03629
-
DGSSA: Domain generalization with structural and stylistic augmentation for retinal vessel segmentation 7 Jan 2025 · 0 repositories · arXiv:2501.03466
-
Efficient and Accurate Tuberculosis Diagnosis: Attention Residual U-Net and Vision Transformer Based Detection Framework 7 Jan 2025 · 0 repositories · arXiv:2501.03538
-
Finding A Voice: Evaluating African American Dialect Generation for Chatbot Technology 7 Jan 2025 · 1 repository · arXiv:2501.03441
-
How to Select Pre-Trained Code Models for Reuse? A Learning Perspective 7 Jan 2025 · 1 repository · arXiv:2501.03783
-
HP-BERT: A framework for longitudinal study of Hinduphobia on social media via LLMs 7 Jan 2025 · 1 repository · arXiv:2501.05482
-
IntegrityAI at GenAI Detection Task 2: Detecting Machine-Generated Academic Essays in English and Arabic Using ELECTRA and Stylometry 7 Jan 2025 · 0 repositories · arXiv:2501.05476
-
Language and Planning in Robotic Navigation: A Multilingual Evaluation of State-of-the-Art Models 7 Jan 2025 · 0 repositories · arXiv:2501.05478
-
LM-Net: A Light-weight and Multi-scale Network for Medical Image Segmentation 7 Jan 2025 · 1 repository · arXiv:2501.03838
-
MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems 7 Jan 2025 · 1 repository · arXiv:2501.03468
-
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding 7 Jan 2025 · 0 repositories · arXiv:2501.05479
-
RAG-Check: Evaluating Multimodal Retrieval Augmented Generation Performance 7 Jan 2025 · 0 repositories · arXiv:2501.03995
-
Reading with Intent -- Neutralizing Intent 7 Jan 2025 · 0 repositories · arXiv:2501.03475
-
SNR-EQ-JSCC: Joint Source-Channel Coding with SNR-Based Embedding and Query 7 Jan 2025 · 0 repositories · arXiv:2501.04732
-
Text to Band Gap: Pre-trained Language Models as Encoders for Semiconductor Band Gap Prediction 7 Jan 2025 · 1 repository · arXiv:2501.03456
-
Three-dimensional attention Transformer for state evaluation in real-time strategy games 7 Jan 2025 · 0 repositories · arXiv:2501.03832
-
A Novel Vision Transformer for Camera-LiDAR Fusion based Traffic Object Segmentation 6 Jan 2025 · 0 repositories · arXiv:2501.02858
-
Adaptive Pruning of Pretrained Transformer via Differential Inclusions 6 Jan 2025 · 0 repositories · arXiv:2501.03289
-
CHAT: Beyond Contrastive Graph Transformer for Link Prediction in Heterogeneous Networks 6 Jan 2025 · 0 repositories · arXiv:2501.02760
-
Developing an Artificial Intelligence Tool for Personalized Breast Cancer Treatment Plans based on the NCCN Guidelines 6 Jan 2025 · 0 repositories · arXiv:2502.15698
-
Intelligent logistics management robot path planning algorithm integrating transformer and GCN network 6 Jan 2025 · 0 repositories · arXiv:2501.02749
-
FlippedRAG: Black-Box Opinion Manipulation Adversarial Attacks to Retrieval-Augmented Generation Models 6 Jan 2025 · 0 repositories · arXiv:2501.02968
-
GLoG-CSUnet: Enhancing Vision Transformers with Adaptable Radiomic Features for Medical Image Segmentation 6 Jan 2025 · 1 repository · arXiv:2501.02788
-
Graph-based Retrieval Augmented Generation for Dynamic Few-shot Text Classification 6 Jan 2025 · 0 repositories · arXiv:2501.02844
-
Integrating Language-Image Prior into EEG Decoding for Cross-Task Zero-Calibration RSVP-BCI 6 Jan 2025 · 0 repositories · arXiv:2501.02841
-
Mixture-of-Experts Graph Transformers for Interpretable Particle Collision Detection 6 Jan 2025 · 1 repository · arXiv:2501.03432
-
Multi-Modal One-Shot Federated Ensemble Learning for Medical Data with Vision Large Language Model 6 Jan 2025 · 0 repositories · arXiv:2501.03292
-
Political Events using RAG with LLMs 6 Jan 2025 · 0 repositories · arXiv:2502.15701
-
QuIM-RAG: Advancing Retrieval-Augmented Generation with Inverted Question Matching for Enhanced QA Performance 6 Jan 2025 · 0 repositories · arXiv:2501.02702
-
SALT: Sales Autocompletion Linked Business Tables Dataset 6 Jan 2025 · 1 repository · arXiv:2501.03413
-
Scalable Forward-Forward Algorithm 6 Jan 2025 · 0 repositories · arXiv:2501.03176
-
Sensorformer: Cross-patch attention with global-patch compression is effective for high-dimensional multivariate time series forecasting 6 Jan 2025 · 0 repositories · arXiv:2501.03284
-
Sequence Complementor: Complementing Transformers For Time Series Forecasting with Learnable Sequences 6 Jan 2025 · 0 repositories · arXiv:2501.02735
-
Tree-based RAG-Agent Recommendation System: A Case Study in Medical Test Data 6 Jan 2025 · 0 repositories · arXiv:2501.02727
-
VicSim: Enhancing Victim Simulation with Emotional and Linguistic Fidelity 6 Jan 2025 · 0 repositories · arXiv:2501.03139
-
Decoding fMRI Data into Captions using Prefix Language Modeling 5 Jan 2025 · 1 repository · arXiv:2501.02570
-
DeTrack: In-model Latent Denoising Learning for Visual Object Tracking 5 Jan 2025 · 0 repositories · arXiv:2501.02467
-
Empowering Bengali Education with AI: Solving Bengali Math Word Problems through Transformer Models 5 Jan 2025 · 0 repositories · arXiv:2501.02599
-
Evaluating Large Language Models Against Human Annotators in Latent Content Analysis: Sentiment, Political Leaning, Emotional Intensity, and Sarcasm 5 Jan 2025 · 0 repositories · arXiv:2501.02532
-
GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking 5 Jan 2025 · 0 repositories · arXiv:2501.02690
-
HonkaiChat: Companions from Anime that feel alive! 5 Jan 2025 · 0 repositories · arXiv:2501.03277
-
LWFNet: Coherent Doppler Wind Lidar-Based Network for Wind Field Retrieval 5 Jan 2025 · 0 repositories · arXiv:2501.02613
-
PTEENet: Post-Trained Early-Exit Neural Networks Augmentation for Inference Cost Optimization 5 Jan 2025 · 0 repositories · arXiv:2501.02508
-
Towards New Benchmark for AI Alignment & Sentiment Analysis in Socially Important Issues: A Comparative Study of Human and LLMs in the Context of AGI 5 Jan 2025 · 0 repositories · arXiv:2501.02531
-
Context Aware Lemmatization and Morphological Tagging Method in Turkish 4 Jan 2025 · 0 repositories · arXiv:2501.02361
-
Examining the Robustness of Homogeneity Bias to Hyperparameter Adjustments in GPT-4 4 Jan 2025 · 0 repositories · arXiv:2501.02211
-
Exploring the Capabilities and Limitations of Large Language Models for Radiation Oncology Decision Support 4 Jan 2025 · 0 repositories · arXiv:2501.02346
-
Graph-Aware Isomorphic Attention for Adaptive Dynamics in Transformers 4 Jan 2025 · 1 repository · arXiv:2501.02393
-
Knowledge Graph Retrieval-Augmented Generation for LLM-based Recommendation 4 Jan 2025 · 0 repositories · arXiv:2501.02226
-
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena 4 Jan 2025 · 0 repositories · arXiv:2501.03266
-
The Application of Large Language Models in Recommendation Systems 4 Jan 2025 · 0 repositories · arXiv:2501.02178
-
A Separable Self-attention Inspired by the State Space Model for Computer Vision 3 Jan 2025 · 1 repository · arXiv:2501.02040
-
A Survey on Large Language Models with some Insights on their Capabilities and Limitations 3 Jan 2025 · 0 repositories · arXiv:2501.04040
-
AgentRefine: Enhancing Agent Generalization through Refinement Tuning 3 Jan 2025 · 0 repositories · arXiv:2501.01702
-
BARTPredict: Empowering IoT Security with LLM-Driven Cyber Threat Prediction 3 Jan 2025 · 0 repositories · arXiv:2501.01664
-
BERT4MIMO: A Foundation Model using BERT Architecture for Massive MIMO Channel State Information Prediction 3 Jan 2025 · 1 repository · arXiv:2501.01802
-
Classifier-Guided Captioning Across Modalities 3 Jan 2025 · 0 repositories · arXiv:2501.03183
-
End-to-End Long Document Summarization using Gradient Caching 3 Jan 2025 · 0 repositories · arXiv:2501.01805
-
GoBERT: Gene Ontology Graph Informed BERT for Universal Gene Function Prediction 3 Jan 2025 · 0 repositories · arXiv:2501.01930
-
Improving Transducer-Based Spoken Language Understanding with Self-Conditioned CTC and Knowledge Transfer 3 Jan 2025 · 0 repositories · arXiv:2501.01936
-
LLMs & Legal Aid: Understanding Legal Needs Exhibited Through User Queries 3 Jan 2025 · 0 repositories · arXiv:2501.01711
-
MIRAGE: Exploring How Large Language Models Perform in Complex Social Interactive Environments 3 Jan 2025 · 1 repository · arXiv:2501.01652
-
PersonaAI: Leveraging Retrieval-Augmented Generation and Personalized Context for AI-Driven Digital Avatars 3 Jan 2025 · 0 repositories · arXiv:2503.15489
-
Quantitative Gait Analysis from Single RGB Videos Using a Dual-Input Transformer-Based Network 3 Jan 2025 · 1 repository · arXiv:2501.01689
-
Towards Hard and Soft Shadow Removal via Dual-Branch Separation Network and Vision Transformer 3 Jan 2025 · 0 repositories · arXiv:2501.01864
-
Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions 3 Jan 2025 · 1 repository · arXiv:2501.01872Syntology official: no sample here; runs from other or unrecorded repositories · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
VidFormer: A novel end-to-end framework fused by 3DCNN and Transformer for Video-based Remote Physiological Measurement 3 Jan 2025 · 0 repositories · arXiv:2501.01691
-
3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer 2 Jan 2025 · 0 repositories · arXiv:2501.01163
-
An Efficient Attention Mechanism for Sequential Recommendation Tasks: HydraRec 2 Jan 2025 · 0 repositories · arXiv:2501.01242
-
BeliN: A Novel Corpus for Bengali Religious News Headline Generation using Contextual Feature Fusion 2 Jan 2025 · 1 repository · arXiv:2501.01069
-
Disambiguation of Chinese Polyphones in an End-to-End Framework with Semantic Features Extracted by Pre-trained BERT 2 Jan 2025 · 0 repositories · arXiv:2501.01102
-
EHCTNet: Enhanced Hybrid of CNN and Transformer Network for Remote Sensing Image Change Detection 2 Jan 2025 · 0 repositories · arXiv:2501.01238
-
Graph Generative Pre-trained Transformer 2 Jan 2025 · 0 repositories · arXiv:2501.01073
-
Large Language Models for Mental Health Diagnostic Assessments: Exploring The Potential of Large Language Models for Assisting with Mental Health Diagnostic Assessments -- The Depression and Anxiety Case 2 Jan 2025 · 0 repositories · arXiv:2501.01305
-
Learning Spectral Methods by Transformers 2 Jan 2025 · 0 repositories · arXiv:2501.01312
-
Long-range Brain Graph Transformer 2 Jan 2025 · 1 repository · arXiv:2501.01100Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Missing Data as Augmentation in the Earth Observation Domain: A Multi-View Learning Approach 2 Jan 2025 · 1 repository · arXiv:2501.01132
-
MSWA: Refining Local Attention with Multi-ScaleWindow Attention 2 Jan 2025 · 0 repositories · arXiv:2501.01039
-
Multi-Head Explainer: A General Framework to Improve Explainability in CNNs and Transformers 2 Jan 2025 · 0 repositories · arXiv:2501.01311
-
Multi-Modal Video Feature Extraction for Popularity Prediction 2 Jan 2025 · 0 repositories · arXiv:2501.01422
-
nnY-Net: Swin-NeXt with Cross-Attention for 3D Medical Images Segmentation 2 Jan 2025 · 0 repositories · arXiv:2501.01406
-
Operator Learning for Reconstructing Flow Fields from Sparse Measurements: an Energy Transformer Approach 2 Jan 2025 · 0 repositories · arXiv:2501.08339
-
Predicting the Performance of Black-box LLMs through Self-Queries 2 Jan 2025 · 1 repository · arXiv:2501.01558Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models 2 Jan 2025 · 2 repositories · arXiv:2501.01423