Methods › General › Stochastic Optimization › Adam › Papers, page 26
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 26 of 244: papers 2,501 to 2,600 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Transformer-Based Wireless Capsule Endoscopy Bleeding Tissue Detection and Classification 26 Dec 2024 · 1 repository · arXiv:2412.19218
-
Adopting Trustworthy AI for Sleep Disorder Prediction: Deep Time Series Analysis with Temporal Attention Mechanism and Counterfactual Explanations 25 Dec 2024 · 0 repositories · arXiv:2412.18971
-
DCIS: Efficient Length Extrapolation of LLMs via Divide-and-Conquer Scaling Factor Search 25 Dec 2024 · 1 repository · arXiv:2412.18811
-
Distortion-Aware Adversarial Attacks on Bounding Boxes of Object Detectors 25 Dec 2024 · 1 repository · arXiv:2412.18815
-
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation 25 Dec 2024 · 0 repositories · arXiv:2412.18907
-
Evaluating the Adversarial Robustness of Detection Transformers 25 Dec 2024 · 0 repositories · arXiv:2412.18718
-
Injecting Bias into Text Classification Models using Backdoor Attacks 25 Dec 2024 · 0 repositories · arXiv:2412.18975
-
Ister: Inverted Seasonal-Trend Decomposition Transformer for Explainable Multivariate Time Series Forecasting 25 Dec 2024 · 0 repositories · arXiv:2412.18798
-
MTCAE-DFER: Multi-Task Cascaded Autoencoder for Dynamic Facial Expression Recognition 25 Dec 2024 · 1 repository · arXiv:2412.18988
-
Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks 25 Dec 2024 · 0 repositories · arXiv:2412.18729
-
Position-aware Graph Transformer for Recommendation 25 Dec 2024 · 0 repositories · arXiv:2412.18731
-
Resource-Efficient Transformer Architecture: Optimizing Memory and Execution Time for Real-Time Applications 25 Dec 2024 · 0 repositories · arXiv:2501.00042
-
SAFLITE: Fuzzing Autonomous Systems via Large Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18727
-
Torque-Aware Momentum 25 Dec 2024 · 0 repositories · arXiv:2412.18790
-
UNIC-Adapter: Unified Image-instruction Adapter with Multi-modal Transformer for Image Generation 25 Dec 2024 · 0 repositories · arXiv:2412.18928
-
Using Large Language Models for Automated Grading of Student Writing about Science 25 Dec 2024 · 0 repositories · arXiv:2412.18719
-
Whose Morality Do They Speak? Unraveling Cultural Bias in Multilingual Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18863
-
AutoSculpt: A Pattern-based Model Auto-pruning Framework Using Reinforcement Learning and Graph Learning 24 Dec 2024 · 0 repositories · arXiv:2412.18091
-
Comprehensive Assessment of BERT-Based Methods for Predicting Antimicrobial Peptides 24 Dec 2024 · 1 repository
-
Decentralized Intelligence in GameFi: Embodied AI Agents and the Convergence of DeFi and Virtual Ecosystems 24 Dec 2024 · 1 repository · arXiv:2412.18601
-
DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation 24 Dec 2024 · 1 repository · arXiv:2412.18597Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Do Language Models Understand the Cognitive Tasks Given to Them? Investigations with the N-Back Paradigm 24 Dec 2024 · 0 repositories · arXiv:2412.18120
-
EvoPat: A Multi-LLM-based Patents Summarization and Analysis Agent 24 Dec 2024 · 0 repositories · arXiv:2412.18100
-
GeAR: Graph-enhanced Agent for Retrieval-augmented Generation 24 Dec 2024 · 0 repositories · arXiv:2412.18431
-
Improving Factuality with Explicit Working Memory 24 Dec 2024 · 0 repositories · arXiv:2412.18069
-
Leveraging Convolutional Neural Network-Transformer Synergy for Predictive Modeling in Risk-Based Applications 24 Dec 2024 · 0 repositories · arXiv:2412.18222
-
Molly: Making Large Language Model Agents Solve Python Problem More Logically 24 Dec 2024 · 0 repositories · arXiv:2412.18093
-
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English 24 Dec 2024 · 1 repository · arXiv:2412.18415
-
Pirates of the RAG: Adaptively Attacking LLMs to Leak Knowledge Bases 24 Dec 2024 · 0 repositories · arXiv:2412.18295
-
Predator Prey Scavenger Model using Holling's Functional Response of Type III and Physics-Informed Deep Neural Networks 24 Dec 2024 · 0 repositories · arXiv:2412.18344
-
Research on the Proximity Relationships of Psychosomatic Disease Knowledge Graph Modules Extracted by Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18419
-
Segment-Based Attention Masking for GPTs 24 Dec 2024 · 1 repository · arXiv:2412.18487
-
TAB: Transformer Attention Bottlenecks enable User Intervention and Debugging in Vision-Language Models 24 Dec 2024 · 1 repository · arXiv:2412.18675
-
TimelyLLM: Segmented LLM Serving System for Time-sensitive Robotic Applications 24 Dec 2024 · 0 repositories · arXiv:2412.18695
-
Unlocking the Potential of Multiple BERT Models for Bangla Question Answering in NCTB Textbooks 24 Dec 2024 · 0 repositories · arXiv:2412.18440
-
A Survey of Query Optimization in Large Language Models 23 Dec 2024 · 0 repositories · arXiv:2412.17558
-
CiteBART: Learning to Generate Citations for Local Citation Recommendation 23 Dec 2024 · 1 repository · arXiv:2412.17534Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
Comparative Analysis of Document-Level Embedding Methods for Similarity Scoring on Shakespeare Sonnets and Taylor Swift Lyrics 23 Dec 2024 · 0 repositories · arXiv:2412.17552
-
DiffFormer: a Differential Spatial-Spectral Transformer for Hyperspectral Image Classification 23 Dec 2024 · 1 repository · arXiv:2412.17350
-
Edge-AI for Agriculture: Lightweight Vision Models for Disease Detection in Resource-Limited Settings 23 Dec 2024 · 0 repositories · arXiv:2412.18635
-
Efficient fine-tuning methodology of text embedding models for information retrieval: contrastive learning penalty (clp) 23 Dec 2024 · 1 repository · arXiv:2412.17364
-
Fast Gradient Computation for RoPE Attention in Almost Linear Time 23 Dec 2024 · 0 repositories · arXiv:2412.17316
-
LayerDropBack: A Universally Applicable Approach for Accelerating Training of Deep Networks 23 Dec 2024 · 1 repository · arXiv:2412.18027
-
Multimodal Preference Data Synthetic Alignment with Reward Model 23 Dec 2024 · 1 repository · arXiv:2412.17417
-
STeInFormer: Spatial-Temporal Interaction Transformer Architecture for Remote Sensing Change Detection 23 Dec 2024 · 1 repository · arXiv:2412.17247
-
Theoretical Constraints on the Expressive Power of RoPE-based Tensor Attention Transformers 23 Dec 2024 · 0 repositories · arXiv:2412.18040
-
Token Statistics Transformer: Linear-Time Attention via Variational Rate Reduction 23 Dec 2024 · 1 repository · arXiv:2412.17810
-
URoadNet: Dual Sparse Attentive U-Net for Multiscale Road Network Extraction 23 Dec 2024 · 0 repositories · arXiv:2412.17573
-
A Reality Check on Context Utilisation for Retrieval-Augmented Generation 22 Dec 2024 · 1 repository · arXiv:2412.17031
-
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps 22 Dec 2024 · 0 repositories · arXiv:2412.17113
-
An OpenMind for 3D medical vision self-supervised learning 22 Dec 2024 · 1 repository · arXiv:2412.17041
-
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models 22 Dec 2024 · 0 repositories · arXiv:2501.03246
-
DR-Encoder: Encode Low-rank Gradients with Random Prior for Large Language Models Differentially Privately 22 Dec 2024 · 0 repositories · arXiv:2412.17053
-
Grams: Gradient Descent with Adaptive Momentum Scaling 22 Dec 2024 · 1 repository · arXiv:2412.17107Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Multifaceted User Modeling in Recommendation: A Federated Foundation Models Approach 22 Dec 2024 · 1 repository · arXiv:2412.16969
-
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora 22 Dec 2024 · 0 repositories · arXiv:2412.16976
-
PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental Health 22 Dec 2024 · 1 repository · arXiv:2412.16882
-
Reconsidering SMT Over NMT for Closely Related Languages: A Case Study of Persian-Hindi Pair 22 Dec 2024 · 0 repositories · arXiv:2412.16877
-
Robustness of Large Language Models Against Adversarial Attacks 22 Dec 2024 · 0 repositories · arXiv:2412.17011
-
SubstationAI: Multimodal Large Model-Based Approaches for Analyzing Substation Equipment Faults 22 Dec 2024 · 0 repositories · arXiv:2412.17077
-
Survey on Abstractive Text Summarization: Dataset, Models, and Metrics 22 Dec 2024 · 2 repositories · arXiv:2412.17165
-
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction 22 Dec 2024 · 0 repositories · arXiv:2412.16919
-
AlzheimerRAG: Multimodal Retrieval Augmented Generation for PubMed articles 21 Dec 2024 · 0 repositories · arXiv:2412.16701
-
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans? 21 Dec 2024 · 0 repositories · arXiv:2412.16772
-
Distilling Large Language Models for Efficient Clinical Information Extraction 21 Dec 2024 · 0 repositories · arXiv:2501.00031
-
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16486
-
Flash3D: Super-scaling Point Transformers through Joint Hardware-Geometry Locality 21 Dec 2024 · 1 repository · arXiv:2412.16481
-
Formal Language Knowledge Corpus for Retrieval Augmented Generation 21 Dec 2024 · 0 repositories · arXiv:2412.16689
-
From Histopathology Images to Cell Clouds: Learning Slide Representations with Hierarchical Cell Transformer 21 Dec 2024 · 0 repositories · arXiv:2412.16715
-
Identifying Cyberbullying Roles in Social Media 21 Dec 2024 · 0 repositories · arXiv:2412.16417
-
Improving FIM Code Completions via Context & Curriculum Based Learning 21 Dec 2024 · 0 repositories · arXiv:2412.16589
-
Lillama: Large Language Models Compression via Low-Rank Feature Distillation 21 Dec 2024 · 0 repositories · arXiv:2412.16719
-
Object Detection Approaches to Identifying Hand Images with High Forensic Values 21 Dec 2024 · 0 repositories · arXiv:2412.16431
-
Paraformer: Parameterization of Sub-grid Scale Processes Using Transformers 21 Dec 2024 · 0 repositories · arXiv:2412.16763
-
Quantum-Like Contextuality in Large Language Models 21 Dec 2024 · 1 repository · arXiv:2412.16806
-
Research on Violent Text Detection System Based on BERT-fasttext Model 21 Dec 2024 · 0 repositories · arXiv:2412.16455
-
STKDRec: Spatial-Temporal Knowledge Distillation for Takeaway Recommendation 21 Dec 2024 · 1 repository · arXiv:2412.16502
-
TimeRAG: BOOSTING LLM Time Series Forecasting via Retrieval-Augmented Generation 21 Dec 2024 · 0 repositories · arXiv:2412.16643
-
Towards More Robust Retrieval-Augmented Generation: Evaluating RAG Under Adversarial Poisoning Attacks 21 Dec 2024 · 1 repository · arXiv:2412.16708
-
VSFormer: Value and Shape-Aware Transformer with Prior-Enhanced Self-Attention for Multivariate Time Series Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16515
-
Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline 20 Dec 2024 · 0 repositories · arXiv:2412.15660
-
Adversarial Robustness through Dynamic Ensemble Learning 20 Dec 2024 · 0 repositories · arXiv:2412.16254
-
Benchmarking LLMs and SLMs for patient reported outcomes 20 Dec 2024 · 0 repositories · arXiv:2412.16291
-
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation 20 Dec 2024 · 0 repositories · arXiv:2412.16135
-
Decoding Linguistic Nuances in Mental Health Text Classification Using Expressive Narrative Stories 20 Dec 2024 · 0 repositories · arXiv:2412.16302
-
Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring 20 Dec 2024 · 0 repositories · arXiv:2412.16108
-
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks 20 Dec 2024 · 1 repository · arXiv:2412.15605Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Explainable AI for Multivariate Time Series Pattern Exploration: Latent Space Visual Analytics with Temporal Fusion Transformer and Variational Autoencoders in Power Grid Event Diagnosis 20 Dec 2024 · 0 repositories · arXiv:2412.16098
-
Foxtsage vs. Adam: Revolution or Evolution in Optimization? 20 Dec 2024 · 0 repositories · arXiv:2412.17855
-
Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context 20 Dec 2024 · 0 repositories · arXiv:2412.16359
-
Humanlike Cognitive Patterns as Emergent Phenomena in Large Language Models 20 Dec 2024 · 0 repositories · arXiv:2412.15501
-
HybGRAG: Hybrid Retrieval-Augmented Generation on Textual and Relational Knowledge Bases 20 Dec 2024 · 0 repositories · arXiv:2412.16311
-
Linguistic Features Extracted by GPT-4 Improve Alzheimer's Disease Detection based on Spontaneous Speech 20 Dec 2024 · 1 repository · arXiv:2412.15772
-
Multi-dimensional Visual Prompt Enhanced Image Restoration via Mamba-Transformer Aggregation 20 Dec 2024 · 1 repository · arXiv:2412.15845
-
PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics 20 Dec 2024 · 0 repositories · arXiv:2412.16120
-
SeagrassFinder: Deep Learning for Eelgrass Detection and Coverage Estimation in the Wild 20 Dec 2024 · 0 repositories · arXiv:2412.16147
-
Towards Interpretable Radiology Report Generation via Concept Bottlenecks using a Multi-Agentic RAG 20 Dec 2024 · 1 repository · arXiv:2412.16086
-
XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation 20 Dec 2024 · 1 repository · arXiv:2412.15529Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
A Full Transformer-based Framework for Automatic Pain Estimation using Videos 19 Dec 2024 · 0 repositories · arXiv:2412.15095
-
A Survey of RWKV 19 Dec 2024 · 1 repository · arXiv:2412.14847