Methods › General › Stochastic Optimization › Adam › Papers, page 87
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 87 of 244: papers 8,601 to 8,700 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Is Mamba Capable of In-Context Learning? 5 Feb 2024 · 1 repository · arXiv:2402.03170Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
LB-KBQA: Large-language-model and BERT based Knowledge-Based Question and Answering System 5 Feb 2024 · 0 repositories · arXiv:2402.05130
-
LLM Agents in Interaction: Measuring Personality Consistency and Linguistic Alignment in Interacting Populations of Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.02896
-
MobilityGPT: Enhanced Human Mobility Modeling with a GPT model 5 Feb 2024 · 0 repositories · arXiv:2402.03264
-
Multi-Lingual Malaysian Embedding: Leveraging Large Language Models for Semantic Representations 5 Feb 2024 · 0 repositories · arXiv:2402.03053
-
SWAG: Storytelling With Action Guidance 5 Feb 2024 · 1 repository · arXiv:2402.03483
-
Toward Human-AI Alignment in Large-Scale Multi-Player Games 5 Feb 2024 · 0 repositories · arXiv:2402.03575
-
UniMem: Towards a Unified View of Long-Context Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03009
-
A flexible Bayesian g-formula for causal survival analyses with time-dependent confounding 4 Feb 2024 · 1 repository · arXiv:2402.02306
-
A Graph is Worth K Words: Euclideanizing Graph using Pure Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02464Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Aligner: Efficient Alignment by Learning to Correct 4 Feb 2024 · 0 repositories · arXiv:2402.02416
-
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models 4 Feb 2024 · 1 repository · arXiv:2402.02370
-
Breaking MLPerf Training: A Case Study on Optimizing BERT 4 Feb 2024 · 0 repositories · arXiv:2402.02447
-
Evaluating Large Language Models in Analysing Classroom Dialogue 4 Feb 2024 · 0 repositories · arXiv:2402.02380
-
GeReA: Question-Aware Prompt Captions for Knowledge-based Visual Question Answering 4 Feb 2024 · 1 repository · arXiv:2402.02503Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation 4 Feb 2024 · 0 repositories · arXiv:2402.14594
-
INViT: A Generalizable Routing Problem Solver with Invariant Nested View Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02317Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Key-Graph Transformer for Image Restoration 4 Feb 2024 · 0 repositories · arXiv:2402.02634
-
Minusformer: Improving Time Series Forecasting by Progressively Learning Residuals 4 Feb 2024 · 1 repository · arXiv:2402.02332
-
Pathformer: Multi-scale Transformers with Adaptive Pathways for Time Series Forecasting 4 Feb 2024 · 1 repository · arXiv:2402.05956Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
PROSAC: Provably Safe Certification for Machine Learning Models under Adversarial Attacks 4 Feb 2024 · 0 repositories · arXiv:2402.02629
-
SPCTNet: A Series-Parallel CNN and Transformer Network for 3D Medical Image Segmentation 4 Feb 2024 · 0 repositories
-
Spin: An Efficient Secure Computation Framework with GPU Acceleration 4 Feb 2024 · 0 repositories · arXiv:2402.02320
-
Timer: Generative Pre-trained Transformers Are Large Time Series Models 4 Feb 2024 · 1 repository · arXiv:2402.02368Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Unified Training of Universal Time Series Forecasting Transformers 4 Feb 2024 · 1 repository · arXiv:2402.02592
-
BetterV: Controlled Verilog Generation with Discriminative Guidance 3 Feb 2024 · 0 repositories · arXiv:2402.03375
-
Data Quality Matters: Suicide Intention Detection on Social Media Posts Using RoBERTa-CNN 3 Feb 2024 · 0 repositories · arXiv:2402.02262
-
DE³-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks 3 Feb 2024 · 0 repositories · arXiv:2402.05948
-
DiffVein: A Unified Diffusion Network for Finger Vein Segmentation and Authentication 3 Feb 2024 · 0 repositories · arXiv:2402.02060
-
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test 3 Feb 2024 · 0 repositories · arXiv:2402.02135
-
EffiBench: Benchmarking the Efficiency of Automatically Generated Code 3 Feb 2024 · 1 repository · arXiv:2402.02037Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
How well do LLMs cite relevant medical references? An evaluation framework and analyses 3 Feb 2024 · 1 repository · arXiv:2402.02008
-
IMUSE: IMU-based Facial Expression Capture 3 Feb 2024 · 0 repositories · arXiv:2402.03944
-
ParZC: Parametric Zero-Cost Proxies for Efficient NAS 3 Feb 2024 · 0 repositories · arXiv:2402.02105
-
ScribFormer: Transformer Makes CNN Work Better for Scribble-based Medical Image Segmentation 3 Feb 2024 · 1 repository · arXiv:2402.02029
-
TCI-Former: Thermal Conduction-Inspired Transformer for Infrared Small Target Detection 3 Feb 2024 · 0 repositories · arXiv:2402.02046
-
MinMaxMin Q-learning 3 Feb 2024 · 0 repositories · arXiv:2402.05951
-
SQT -- std Q-target 3 Feb 2024 · 0 repositories · arXiv:2402.05950
-
Topology-Informed Graph Transformer 3 Feb 2024 · 2 repositories · arXiv:2402.02005Syntology official (archive's flag): 8 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 2 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Hierarchical Structure Enhances the Convergence and Generalizability of Linear Molecular Representation 3 Feb 2024 · 1 repository · arXiv:2402.02164
-
A Data-Driven Analysis of Robust Automatic Piano Transcription 2 Feb 2024 · 0 repositories · arXiv:2402.01424
-
ALERT-Transformer: Bridging Asynchronous and Synchronous Machine Learning for Real-Time Event-based Spatio-Temporal Data 2 Feb 2024 · 0 repositories · arXiv:2402.01393
-
An introduction to graphical tensor notation for mechanistic interpretability 2 Feb 2024 · 0 repositories · arXiv:2402.01790
-
BAT: Learning to Reason about Spatial Sounds with Large Language Models 2 Feb 2024 · 0 repositories · arXiv:2402.01591
-
Challenges in Training PINNs: A Loss Landscape Perspective 2 Feb 2024 · 1 repository · arXiv:2402.01868Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Clarifying the Path to User Satisfaction: An Investigation into Clarification Usefulness 2 Feb 2024 · 1 repository · arXiv:2402.01934
-
COMET: Generating Commit Messages using Delta Graph Context Representation 2 Feb 2024 · 0 repositories · arXiv:2402.01841
-
Cross-view Masked Diffusion Transformers for Person Image Synthesis 2 Feb 2024 · 1 repository · arXiv:2402.01516Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Enhancing Stochastic Gradient Descent: A Unified Framework and Novel Acceleration Methods for Faster Convergence 2 Feb 2024 · 0 repositories · arXiv:2402.01515
-
Can LLMs perform structured graph reasoning? 2 Feb 2024 · 1 repository · arXiv:2402.01805
-
Faster Inference of Integer SWIN Transformer by Removing the GELU Activation 2 Feb 2024 · 0 repositories · arXiv:2402.01169
-
How Can Generative AI Enhance the Well-being of Blind? 2 Feb 2024 · 0 repositories · arXiv:2402.07919
-
Improving Sequential Recommendations with LLMs 2 Feb 2024 · 1 repository · arXiv:2402.01339Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Integrating Large Language Models in Causal Discovery: A Statistical Causal Approach 2 Feb 2024 · 2 repositories · arXiv:2402.01454
-
LLM-Detector: Improving AI-Generated Chinese Text Detection with Open-Source LLM Instruction Tuning 2 Feb 2024 · 1 repository · arXiv:2402.01158
-
LoTR: Low Tensor Rank Weight Adaptation 2 Feb 2024 · 0 repositories · arXiv:2402.01376
-
Predicting ATP binding sites in protein sequences using Deep Learning and Natural Language Processing 2 Feb 2024 · 0 repositories · arXiv:2402.01829
-
Retrieval Augmented End-to-End Spoken Dialog Models 2 Feb 2024 · 0 repositories · arXiv:2402.01828
-
Todyformer: Towards Holistic Dynamic Graph Transformers with Structure-Aware Tokenization 2 Feb 2024 · 0 repositories · arXiv:2402.05944
-
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks 2 Feb 2024 · 0 repositories · arXiv:2402.01176
-
Transformers Learn Nonlinear Features In Context: Nonconvex Mean-field Dynamics on the Attention Landscape 2 Feb 2024 · 0 repositories · arXiv:2402.01258
-
TravelPlanner: A Benchmark for Real-World Planning with Language Agents 2 Feb 2024 · 2 repositories · arXiv:2402.01622Syntology official (archive's flag): 13 ran · 17 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Understanding Adam Optimizer via Online Learning of Updates: Adam is FTRL in Disguise 2 Feb 2024 · 0 repositories · arXiv:2402.01567
-
What Will My Model Forget? Forecasting Forgotten Examples in Language Model Refinement 2 Feb 2024 · 0 repositories · arXiv:2402.01865
-
Dendritic Learning-incorporated Vision Transformer for Image Recognition 1 Feb 2024 · 1 repository
-
FuseFormer: A Transformer for Visual and Thermal Image Fusion 1 Feb 2024 · 0 repositories · arXiv:2402.00971
-
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model 1 Feb 2024 · 0 repositories · arXiv:2402.01051
-
Hierarchical Multi-Label Classification of Online Vaccine Concerns 1 Feb 2024 · 0 repositories · arXiv:2402.01783
-
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA 1 Feb 2024 · 0 repositories · arXiv:2402.01767
-
Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing 1 Feb 2024 · 0 repositories · arXiv:2402.00658
-
Merging Multi-Task Models via Weight-Ensembling Mixture of Experts 1 Feb 2024 · 1 repository · arXiv:2402.00433Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multivariate Probabilistic Time Series Forecasting with Correlated Errors 1 Feb 2024 · 1 repository · arXiv:2402.01000Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Neural Style Transfer with Twin-Delayed DDPG for Shared Control of Robotic Manipulators 1 Feb 2024 · 0 repositories · arXiv:2402.00722
-
Ocassionally Secure: A Comparative Analysis of Code Generation Assistants 1 Feb 2024 · 0 repositories · arXiv:2402.00689
-
On the Psychology of GPT-4: Moderately anxious, slightly masculine, honest, and humble 1 Feb 2024 · 0 repositories · arXiv:2402.01777
-
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models 1 Feb 2024 · 1 repository · arXiv:2402.00794
-
Investigating Recurrent Transformers with Dynamic Halt 1 Feb 2024 · 1 repository · arXiv:2402.00976
-
Self-Supervised Contrastive Pre-Training for Multivariate Point Processes 1 Feb 2024 · 0 repositories · arXiv:2402.00987
-
SPARQL Generation with Entity Pre-trained GPT for KG Question Answering 1 Feb 2024 · 1 repository · arXiv:2402.00969
-
Tiny Titans: Can Smaller Large Language Models Punch Above Their Weight in the Real World for Meeting Summarization? 1 Feb 2024 · 0 repositories · arXiv:2402.00841
-
Human-mediated Large Language Models for Robotic Intervention in Children with Autism Spectrum Disorders 1 Feb 2024 · 0 repositories · arXiv:2402.00260
-
Understanding the Expressive Power and Mechanisms of Transformer for Sequence Modeling 1 Feb 2024 · 0 repositories · arXiv:2402.00522
-
Code-Aware Prompting: A study of Coverage Guided Test Generation in Regression Setting using LLM 31 Jan 2024 · 0 repositories · arXiv:2402.00097
-
Computation and Parameter Efficient Multi-Modal Fusion Transformer for Cued Speech Recognition 31 Jan 2024 · 0 repositories · arXiv:2401.17604
-
ConSmax: Hardware-Friendly Alternative Softmax with Learnable Parameters 31 Jan 2024 · 1 repository · arXiv:2402.10930
-
Document Structure in Long Document Transformers 31 Jan 2024 · 0 repositories · arXiv:2401.17658
-
EnCLAP: Combining Neural Audio Codec and Audio-Text Joint Embedding for Automated Audio Captioning 31 Jan 2024 · 1 repository · arXiv:2401.17690
-
Exploring the limits of decoder-only models trained on public speech recognition corpora 31 Jan 2024 · 0 repositories · arXiv:2402.00235
-
Global-Liar: Factuality of LLMs over Time and Geographic Regions 31 Jan 2024 · 0 repositories · arXiv:2401.17839
-
Graph Transformers without Positional Encodings 31 Jan 2024 · 0 repositories · arXiv:2401.17791
-
Head and Neck Tumor Segmentation from [18F]F-FDG PET/CT Images Based on 3D Diffusion Model 31 Jan 2024 · 0 repositories · arXiv:2401.17593
-
Leveraging Swin Transformer for Local-to-Global Weakly Supervised Semantic Segmentation 31 Jan 2024 · 1 repository · arXiv:2401.17828
-
LLM Voting: Human Choices and AI Collective Decision Making 31 Jan 2024 · 1 repository · arXiv:2402.01766Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Local Feature Matching Using Deep Learning: A Survey 31 Jan 2024 · 1 repository · arXiv:2401.17592
-
Making a Long Story Short in Conversation Modeling 31 Jan 2024 · 0 repositories · arXiv:2402.00143
-
Mitigating the Influence of Distractor Tasks in LMs with Prior-Aware Decoding 31 Jan 2024 · 0 repositories · arXiv:2401.17692
-
Paramanu: A Family of Novel Efficient Generative Foundation Language Models for Indian Languages 31 Jan 2024 · 0 repositories · arXiv:2401.18034
-
Positional Encoding Helps Recurrent Neural Networks Handle a Large Vocabulary 31 Jan 2024 · 1 repository · arXiv:2402.00236
-
RAG-Fusion: a New Take on Retrieval-Augmented Generation 31 Jan 2024 · 0 repositories · arXiv:2402.03367
-
RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval 31 Jan 2024 · 3 repositories · arXiv:2401.18059Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)