Methods › General › Stochastic Optimization › Adam › Papers, page 2
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 2 of 244: papers 101 to 200 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Query Routing for Retrieval-Augmented Language Models 29 May 2025 · 0 repositories · arXiv:2505.23052
-
Reducing Latency in LLM-Based Natural Language Commands Processing for Robot Navigation 29 May 2025 · 0 repositories · arXiv:2506.00075
-
Table-R1: Inference-Time Scaling for Table Reasoning 29 May 2025 · 1 repository · arXiv:2505.23621
-
The Rich and the Simple: On the Implicit Bias of Adam and SGD 29 May 2025 · 0 repositories · arXiv:2505.24022
-
The Warmup Dilemma: How Learning Rate Strategies Impact Speech-to-Text Model Convergence 29 May 2025 · 1 repository · arXiv:2505.23420
-
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos 29 May 2025 · 1 repository · arXiv:2505.23693
-
Agent-UniRAG: A Trainable Open-Source LLM Agent Framework for Unified Retrieval-Augmented Generation Systems 28 May 2025 · 0 repositories · arXiv:2505.22571
-
Attention-Enhanced Prompt Decision Transformers for UAV-Assisted Communications with AoI 28 May 2025 · 0 repositories · arXiv:2505.22170
-
Breaking the Cloak! Unveiling Chinese Cloaked Toxicity with Homophone Graph and Toxic Lexicon 28 May 2025 · 0 repositories · arXiv:2505.22184
-
Climate Finance Bench 28 May 2025 · 1 repository · arXiv:2505.22752
-
Cross-modal RAG: Sub-dimensional Retrieval-Augmented Text-to-Image Generation 28 May 2025 · 1 repository · arXiv:2505.21956
-
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control 28 May 2025 · 0 repositories · arXiv:2505.22642
-
HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer 28 May 2025 · 2 repositories · arXiv:2505.22705Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Improving QA Efficiency with DistilBERT: Fine-Tuning and Inference on mobile Intel CPUs 28 May 2025 · 0 repositories · arXiv:2505.22937
-
Multi-MLLM Knowledge Distillation for Out-of-Context News Detection 28 May 2025 · 0 repositories · arXiv:2505.22517
-
MultiFormer: A Multi-Person Pose Estimation System Based on CSI and Attention Mechanism 28 May 2025 · 0 repositories · arXiv:2505.22555
-
PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning 28 May 2025 · 0 repositories · arXiv:2505.22085
-
RAGPPI: RAG Benchmark for Protein-Protein Interactions in Drug Discovery 28 May 2025 · 1 repository · arXiv:2505.23823
-
Say What You Mean: Natural Language Access Control with Large Language Models for Internet of Things 28 May 2025 · 0 repositories · arXiv:2505.23835
-
SkewRoute: Training-Free LLM Routing for Knowledge Graph Retrieval-Augmented Generation via Score Skewness of Retrieved Context 28 May 2025 · 0 repositories · arXiv:2505.23841
-
Update Your Transformer to the Latest Release: Re-Basin of Task Vectors 28 May 2025 · 1 repository · arXiv:2505.22697Syntology official (archive's flag): 1 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 3 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning 28 May 2025 · 1 repository · arXiv:2505.22019Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A domain adaptation neural network for digital twin-supported fault diagnosis 27 May 2025 · 1 repository · arXiv:2505.21046
-
AgriFM: A Multi-source Temporal Remote Sensing Foundation Model for Crop Mapping 27 May 2025 · 1 repository · arXiv:2505.21357
-
Beyond 1D: Vision Transformers and Multichannel Signal Images for PPG-to-ECG Reconstruction 27 May 2025 · 0 repositories · arXiv:2505.21767
-
Continuous-Time Attention: PDE-Guided Mechanisms for Long-Sequence Transformers 27 May 2025 · 0 repositories · arXiv:2505.20666
-
Diagnosing and Resolving Cloud Platform Instability with Multi-modal RAG LLMs 27 May 2025 · 0 repositories · arXiv:2505.21419
-
From prosthetic memory to prosthetic denial: Auditing whether large language models are prone to mass atrocity denialism 27 May 2025 · 0 repositories · arXiv:2505.21753
-
HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling 27 May 2025 · 0 repositories · arXiv:2505.20836
-
HTMNet: A Hybrid Network with Transformer-Mamba Bottleneck Multimodal Fusion for Transparent and Reflective Objects Depth Completion 27 May 2025 · 0 repositories · arXiv:2505.20904
-
Long Context Scaling: Divide and Conquer via Multi-Agent Question-driven Collaboration 27 May 2025 · 0 repositories · arXiv:2505.20625
-
Minute-Long Videos with Dual Parallelisms 27 May 2025 · 1 repository · arXiv:2505.21070
-
MoPFormer: Motion-Primitive Transformer for Wearable-Sensor Activity Recognition 27 May 2025 · 0 repositories · arXiv:2505.20744
-
Pause Tokens Strictly Increase the Expressivity of Constant-Depth Transformers 27 May 2025 · 0 repositories · arXiv:2505.21024
-
PolarGrad: A Class of Matrix-Gradient Optimizers from a Unifying Preconditioning Perspective 27 May 2025 · 0 repositories · arXiv:2505.21799
-
Privacy-Preserving Chest X-ray Report Generation via Multimodal Federated Learning with ViT and GPT-2 27 May 2025 · 0 repositories · arXiv:2505.21715
-
SOSBENCH: Benchmarking Safety Alignment on Scientific Knowledge 27 May 2025 · 0 repositories · arXiv:2505.21605
-
Absolute Coordinates Make Motion Generation Easy 26 May 2025 · 0 repositories · arXiv:2505.19377
-
Aggregated Structural Representation with Large Language Models for Human-Centric Layout Generation 26 May 2025 · 0 repositories · arXiv:2505.19554
-
AMQA: An Adversarial Dataset for Benchmarking Bias of LLMs in Medicine and Healthcare 26 May 2025 · 1 repository · arXiv:2505.19562
-
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents 26 May 2025 · 0 repositories · arXiv:2505.19494
-
Automated evaluation of children's speech fluency for low-resource languages 26 May 2025 · 0 repositories · arXiv:2505.19671
-
Benchmarking Multimodal Knowledge Conflict for Large Multimodal Models 26 May 2025 · 1 repository · arXiv:2505.19509
-
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages 26 May 2025 · 0 repositories · arXiv:2505.19851
-
Calibrating Pre-trained Language Classifiers on LLM-generated Noisy Labels via Iterative Refinement 26 May 2025 · 1 repository · arXiv:2505.19675
-
CardioPatternFormer: Pattern-Guided Attention for Interpretable ECG Classification with Transformer Architecture 26 May 2025 · 0 repositories · arXiv:2505.20481
-
Continuous-Time Analysis of Heavy Ball Momentum in Min-Max Games 26 May 2025 · 0 repositories · arXiv:2505.19537
-
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language 26 May 2025 · 0 repositories · arXiv:2505.19971
-
Dependency Parsing is More Parameter-Efficient with Normalization 26 May 2025 · 0 repositories · arXiv:2505.20215
-
Detection of Suicidal Risk on Social Media: A Hybrid Model 26 May 2025 · 0 repositories · arXiv:2505.23797
-
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems 26 May 2025 · 0 repositories · arXiv:2505.19847
-
DoctorRAG: Medical RAG Fusing Knowledge with Patient Analogy through Textual Gradients 26 May 2025 · 0 repositories · arXiv:2505.19538
-
Emotion Classification In-Context in Spanish 26 May 2025 · 0 repositories · arXiv:2505.20571
-
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining 26 May 2025 · 0 repositories · arXiv:2505.19893
-
GoLF-NRT: Integrating Global Context and Local Geometry for Few-Shot View Synthesis 26 May 2025 · 1 repository · arXiv:2505.19813Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Gradient Flow Matching for Learning Update Dynamics in Neural Network Training 26 May 2025 · 0 repositories · arXiv:2505.20221
-
Grokking ExPLAIND: Unifying Model, Data, and Training Attribution to Study Model Behavior 26 May 2025 · 1 repository · arXiv:2505.20076Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots 26 May 2025 · 1 repository · arXiv:2505.20288
-
KnowTrace: Bootstrapping Iterative Retrieval-Augmented Generation with Structured Knowledge Tracing 26 May 2025 · 1 repository · arXiv:2505.20245
-
Large Language Models' Reasoning Stalls: An Investigation into the Capabilities of Frontier Models 26 May 2025 · 0 repositories · arXiv:2505.19676
-
LeCoDe: A Benchmark Dataset for Interactive Legal Consultation Dialogue Evaluation 26 May 2025 · 0 repositories · arXiv:2505.19667
-
LlamaSeg: Image Segmentation via Autoregressive Mask Generation 26 May 2025 · 0 repositories · arXiv:2505.19422
-
MA-RAG: Multi-Agent Retrieval-Augmented Generation via Collaborative Chain-of-Thought Reasoning 26 May 2025 · 0 repositories · arXiv:2505.20096
-
Minimalist Softmax Attention Provably Learns Constrained Boolean Functions 26 May 2025 · 0 repositories · arXiv:2505.19531
-
Multi-modal brain encoding models for multi-modal stimuli 26 May 2025 · 1 repository · arXiv:2505.20027Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning 26 May 2025 · 0 repositories · arXiv:2505.19938
-
NeuSym-RAG: Hybrid Neural Symbolic Retrieval with Multiview Structuring for PDF Question Answering 26 May 2025 · 1 repository · arXiv:2505.19754
-
One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP 26 May 2025 · 1 repository · arXiv:2505.19840
-
REARANK: Reasoning Re-ranking Agent via Reinforcement Learning 26 May 2025 · 1 repository · arXiv:2505.20046Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples) · 4 pointer-only (licence)
-
Structured Initialization for Vision Transformers 26 May 2025 · 0 repositories · arXiv:2505.19985
-
syftr: Pareto-Optimal Generative AI 26 May 2025 · 1 repository · arXiv:2505.20266
-
Synthetic Time Series Forecasting with Transformer Architectures: Extensive Simulation Benchmarks 26 May 2025 · 1 repository · arXiv:2505.20048
-
The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants 26 May 2025 · 1 repository · arXiv:2505.19797
-
The Missing Point in Vision Transformers for Universal Image Segmentation 26 May 2025 · 1 repository · arXiv:2505.19795
-
Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking 26 May 2025 · 0 repositories · arXiv:2505.20023
-
Transformers in Protein: A Survey 26 May 2025 · 0 repositories · arXiv:2505.20098
-
Understanding Transformer from the Perspective of Associative Memory 26 May 2025 · 0 repositories · arXiv:2505.19488
-
VADER: A Human-Evaluated Benchmark for Vulnerability Assessment, Detection, Explanation, and Remediation 26 May 2025 · 1 repository · arXiv:2505.19395
-
A Smart Healthcare System for Monkeypox Skin Lesion Detection and Tracking 25 May 2025 · 0 repositories · arXiv:2505.19023
-
AI4Math: A Native Spanish Benchmark for University-Level Mathematical Reasoning in Large Language Models 25 May 2025 · 0 repositories · arXiv:2505.18978
-
Assistant-Guided Mitigation of Teacher Preference Bias in LLM-as-a-Judge 25 May 2025 · 1 repository · arXiv:2505.19176
-
Benchmarking Large Language Models for Cyberbullying Detection in Real-World YouTube Comments 25 May 2025 · 0 repositories · arXiv:2505.18927
-
Communication-Efficient Multi-Device Inference Acceleration for Transformer Models 25 May 2025 · 1 repository · arXiv:2505.19342
-
Conventional Contrastive Learning Often Falls Short: Improving Dense Retrieval with Cross-Encoder Listwise Distillation and Synthetic Data 25 May 2025 · 1 repository · arXiv:2505.19274
-
Efficient Data Selection at Scale via Influence Distillation 25 May 2025 · 0 repositories · arXiv:2505.19051
-
Exploring Magnitude Preservation and Rotation Modulation in Diffusion Transformers 25 May 2025 · 0 repositories · arXiv:2505.19122
-
GhostPrompt: Jailbreaking Text-to-image Generative Models based on Dynamic Optimization 25 May 2025 · 0 repositories · arXiv:2505.18979
-
Hermes@DravidianLangTech 2025: Sentiment Analysis of Dravidian Languages using XLM-RoBERTa 25 May 2025 · 1 repository
-
Hypercube-RAG: Hypercube-Based Retrieval-Augmented Generation for In-domain Scientific Question-Answering 25 May 2025 · 1 repository · arXiv:2505.19288
-
Investigating Pedagogical Teacher and Student LLM Agents: Genetic Adaptation Meets Retrieval Augmented Generation Across Learning Style 25 May 2025 · 0 repositories · arXiv:2505.19173
-
Optimized Text Embedding Models and Benchmarks for Amharic Passage Retrieval 25 May 2025 · 1 repository · arXiv:2505.19356
-
POQD: Performance-Oriented Query Decomposer for Multi-vector retrieval 25 May 2025 · 1 repository · arXiv:2505.19189Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Retrieval-Augmented Generation for Service Discovery: Chunking Strategies and Benchmarking 25 May 2025 · 0 repositories · arXiv:2505.19310
-
Scaling Laws for Gradient Descent and Sign Descent for Linear Bigram Models under Zipf's Law 25 May 2025 · 0 repositories · arXiv:2505.19227
-
System-1.5 Reasoning: Traversal in Language and Latent Spaces with Dynamic Shortcuts 25 May 2025 · 0 repositories · arXiv:2505.18962
-
A Survey of LLM × DATA 24 May 2025 · 2 repositories · arXiv:2505.18458
-
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation 24 May 2025 · 0 repositories · arXiv:2505.18543
-
BRIT: Bidirectional Retrieval over Unified Image-Text Graph 24 May 2025 · 0 repositories · arXiv:2505.18450
-
Federated Retrieval-Augmented Generation: A Systematic Mapping Study 24 May 2025 · 0 repositories · arXiv:2505.18906
-
From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data 24 May 2025 · 0 repositories · arXiv:2505.18464