Methods › General › Stochastic Optimization › Adam › Papers, page 65
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 65 of 244: papers 6,401 to 6,500 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
KerasCV and KerasNLP: Vision and Language Power-Ups 30 May 2024 · 0 repositories · arXiv:2405.20247
-
Knowledge Graph Tuning: Real-time Large Language Model Personalization based on Human Feedback 30 May 2024 · 0 repositories · arXiv:2405.19686
-
LLaMEA: A Large Language Model Evolutionary Algorithm for Automatically Generating Metaheuristics 30 May 2024 · 2 repositories · arXiv:2405.20132Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
One Token Can Help! Learning Scalable and Pluggable Virtual Tokens for Retrieval-Augmented Large Language Models 30 May 2024 · 2 repositories · arXiv:2405.19670
-
PATIENT-Ψ: Using Large Language Models to Simulate Patients for Training Mental Health Professionals 30 May 2024 · 1 repository · arXiv:2405.19660Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations 30 May 2024 · 1 repository · arXiv:2405.19740Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Phantom: General Trigger Attacks on Retrieval Augmented Language Generation 30 May 2024 · 0 repositories · arXiv:2405.20485
-
Preference Alignment with Flow Matching 30 May 2024 · 1 repository · arXiv:2405.19806Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
QClusformer: A Quantum Transformer-based Framework for Unsupervised Visual Clustering 30 May 2024 · 0 repositories · arXiv:2405.19722
-
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs 30 May 2024 · 0 repositories · arXiv:2405.20179
-
Robust Image Semantic Coding with Learnable CSI Fusion Masking over MIMO Fading Channels 30 May 2024 · 0 repositories · arXiv:2406.07389
-
Sharing Key Semantics in Transformer Makes Efficient Image Restoration 30 May 2024 · 1 repository · arXiv:2405.20008
-
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation 30 May 2024 · 0 repositories · arXiv:2405.19701
-
Student Answer Forecasting: Transformer-Driven Answer Choice Prediction for Language Learning 30 May 2024 · 1 repository · arXiv:2405.20079
-
TAIA: Large Language Models are Out-of-Distribution Data Learners 30 May 2024 · 1 repository · arXiv:2405.20192Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Towards Ontology-Enhanced Representation Learning for Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20527
-
Towards RGB-NIR Cross-modality Image Registration and Beyond 30 May 2024 · 0 repositories · arXiv:2405.19914
-
Transformers and Slot Encoding for Sample Efficient Physical World Modelling 30 May 2024 · 1 repository · arXiv:2405.20180
-
Use of a Multiscale Vision Transformer to predict Nursing Activities Score from Low Resolution Thermal Videos in an Intensive Care Unit 30 May 2024 · 0 repositories · arXiv:2406.04364
-
WebUOT-1M: Advancing Deep Underwater Object Tracking with A Million-Scale Benchmark 30 May 2024 · 1 repository · arXiv:2405.19818
-
YotoR-You Only Transform One Representation 30 May 2024 · 0 repositories · arXiv:2405.19629
-
A Multi-Source Retrieval Question Answering Framework Based on RAG 29 May 2024 · 0 repositories · arXiv:2405.19207
-
Are You Sure? Rank Them Again: Repeated Ranking For Better Preference Datasets 29 May 2024 · 0 repositories · arXiv:2405.18952
-
Beyond Agreement: Diagnosing the Rationale Alignment of Automated Essay Scoring Methods based on Linguistically-informed Counterfactuals 29 May 2024 · 1 repository · arXiv:2405.19433
-
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension 29 May 2024 · 0 repositories · arXiv:2405.18682
-
Can We Enhance the Quality of Mobile Crowdsensing Data Without Ground Truth? 29 May 2024 · 1 repository · arXiv:2405.18725
-
Contextual Position Encoding: Learning to Count What's Important 29 May 2024 · 0 repositories · arXiv:2405.18719
-
CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control 29 May 2024 · 1 repository · arXiv:2405.18727Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
Deep Grokking: Would Deep Neural Networks Generalize Better? 29 May 2024 · 0 repositories · arXiv:2405.19454
-
Does learning the right latent variables necessarily improve in-context learning? 29 May 2024 · 1 repository · arXiv:2405.19162Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Efficient Model-agnostic Alignment via Bayesian Persuasion 29 May 2024 · 0 repositories · arXiv:2405.18718
-
Enhancing Vision-Language Model with Unmasked Token Alignment 29 May 2024 · 1 repository · arXiv:2405.19009
-
LLM-based Hierarchical Concept Decomposition for Interpretable Fine-Grained Image Classification 29 May 2024 · 0 repositories · arXiv:2405.18672
-
LLMs achieve adult human performance on higher-order theory of mind tasks 29 May 2024 · 0 repositories · arXiv:2405.18870
-
LMO-DP: Optimizing the Randomization Mechanism for Differentially Private Fine-Tuning (Large) Language Models 29 May 2024 · 0 repositories · arXiv:2405.18776
-
MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series 29 May 2024 · 1 repository · arXiv:2405.19327Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Matryoshka Query Transformer for Large Vision-Language Models 29 May 2024 · 1 repository · arXiv:2405.19315Syntology official (archive's flag): 8 ran · 9 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
MDS-ViTNet: Improving saliency prediction for Eye-Tracking with Vision Transformer 29 May 2024 · 1 repository · arXiv:2405.19501
-
MindSemantix: Deciphering Brain Visual Experiences with a Brain-Language Model 29 May 2024 · 0 repositories · arXiv:2405.18812
-
Multi-Channel Multi-Step Spectrum Prediction Using Transformer and Stacked Bi-LSTM 29 May 2024 · 0 repositories · arXiv:2405.19138
-
PediatricsGPT: Large Language Models as Chinese Medical Assistants for Pediatric Applications 29 May 2024 · 1 repository · arXiv:2405.19266Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Reverse Image Retrieval Cues Parametric Memory in Multimodal LLMs 29 May 2024 · 1 repository · arXiv:2405.18740
-
STAT: Shrinking Transformers After Training 29 May 2024 · 0 repositories · arXiv:2406.00061
-
Toward Conversational Agents with Context and Time Sensitive Long-term Memory 29 May 2024 · 1 repository · arXiv:2406.00057
-
Transcending Fusion: A Multi-Scale Alignment Method for Remote Sensing Image-Text Retrieval 29 May 2024 · 1 repository · arXiv:2405.18959
-
Two-Layer Retrieval-Augmented Generation Framework for Low-Resource Medical Question Answering Using Reddit Data: Proof-of-Concept Study 29 May 2024 · 0 repositories · arXiv:2405.19519
-
Understanding and Minimising Outlier Features in Neural Network Training 29 May 2024 · 1 repository · arXiv:2405.19279Syntology official: harvested, nothing ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples)
-
Adam with model exponential moving average is effective for nonconvex optimization 28 May 2024 · 0 repositories · arXiv:2405.18199
-
Aligning to Thousands of Preferences via System Message Generalization 28 May 2024 · 1 repository · arXiv:2405.17977Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
An Empirical Analysis on Large Language Models in Debate Evaluation 28 May 2024 · 1 repository · arXiv:2406.00050Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Are PPO-ed Language Models Hackable? 28 May 2024 · 0 repositories · arXiv:2406.02577
-
ATM: Adversarial Tuning Multi-agent System Makes a Robust Retrieval-Augmented Generator 28 May 2024 · 1 repository · arXiv:2405.18111Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Attention-based sequential recommendation system using multimodal data 28 May 2024 · 0 repositories · arXiv:2405.17959
-
Benchmarks Underestimate the Readiness of Multi-lingual Dialogue Agents 28 May 2024 · 0 repositories · arXiv:2405.17840
-
Delving into Differentially Private Transformer 28 May 2024 · 0 repositories · arXiv:2405.18194
-
Don't Forget to Connect! Improving RAG with Graph-based Reranking 28 May 2024 · 0 repositories · arXiv:2405.18414
-
Dual-Path Multi-Scale Transformer for High-Quality Image Deraining 28 May 2024 · 0 repositories · arXiv:2405.18124
-
Edinburgh Clinical NLP at MEDIQA-CORR 2024: Guiding Large Language Models with Hints 28 May 2024 · 0 repositories · arXiv:2405.18028
-
FASTopic: Pretrained Transformer is a Fast, Adaptive, Stable, and Transferable Topic Model 28 May 2024 · 2 repositories · arXiv:2405.17978Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
ForecastGrapher: Redefining Multivariate Time Series Forecasting with Graph Neural Networks 28 May 2024 · 0 repositories · arXiv:2405.18036
-
HarmoDT: Harmony Multi-Task Decision Transformer for Offline Reinforcement Learning 28 May 2024 · 1 repository · arXiv:2405.18080
-
IAPT: Instruction-Aware Prompt Tuning for Large Language Models 28 May 2024 · 0 repositories · arXiv:2405.18203
-
LDMol: Text-to-Molecule Diffusion Model with Structurally Informative Latent Space 28 May 2024 · 1 repository · arXiv:2405.17829Syntology official (archive's flag): 5 ran · 6 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
LLMs and Memorization: On Quality and Specificity of Copyright Compliance 28 May 2024 · 1 repository · arXiv:2405.18492
-
Modeling Long Sequences in Bladder Cancer Recurrence: A Comparative Evaluation of LSTM,Transformer,and Mamba 28 May 2024 · 0 repositories · arXiv:2405.18518
-
MindFormer: Semantic Alignment of Multi-Subject fMRI for Brain Decoding 28 May 2024 · 0 repositories · arXiv:2405.17720
-
Multi-objective Representation for Numbers in Clinical Narratives: A CamemBERT-Bio-Based Alternative to Large-Scale LLMs 28 May 2024 · 0 repositories · arXiv:2405.18448
-
Notes on Applicability of GPT-4 to Document Understanding 28 May 2024 · 0 repositories · arXiv:2405.18433
-
ORLM: A Customizable Framework in Training Large Models for Automated Optimization Modeling 28 May 2024 · 1 repository · arXiv:2405.17743Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
OV-DQUO: Open-Vocabulary DETR with Denoising Text Query Training and Open-World Unknown Objects Supervision 28 May 2024 · 1 repository · arXiv:2405.17913Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Peering into the Mind of Language Models: An Approach for Attribution in Contextual Question Answering 28 May 2024 · 1 repository · arXiv:2405.17980
-
Proof of Quality: A Costless Paradigm for Trustless Generative AI Model Inference on Blockchains 28 May 2024 · 0 repositories · arXiv:2405.17934
-
RealitySummary: Exploring On-Demand Mixed Reality Text Summarization and Question Answering using Large Language Models 28 May 2024 · 0 repositories · arXiv:2405.18620
-
Thai Winograd Schemas: A Benchmark for Thai Commonsense Reasoning 28 May 2024 · 1 repository · arXiv:2405.18375
-
The Battle of LLMs: A Comparative Study in Conversational QA Tasks 28 May 2024 · 0 repositories · arXiv:2405.18344
-
Towards Communication-efficient Federated Learning via Sparse and Aligned Adaptive Optimization 28 May 2024 · 0 repositories · arXiv:2405.17932
-
Understanding Intrinsic Socioeconomic Biases in Large Language Models 28 May 2024 · 0 repositories · arXiv:2405.18662
-
ViG: Linear-complexity Visual Sequence Learning with Gated Linear Attention 28 May 2024 · 1 repository · arXiv:2405.18425
-
Visual Anchors Are Strong Information Aggregators For Multimodal Large Language Model 28 May 2024 · 1 repository · arXiv:2405.17815
-
VITON-DiT: Learning In-the-Wild Video Try-On from Human Dance Videos via Diffusion Transformers 28 May 2024 · 0 repositories · arXiv:2405.18326
-
Wavelet-Based Image Tokenizer for Vision Transformers 28 May 2024 · 0 repositories · arXiv:2405.18616
-
WIDIn: Wording Image for Domain-Invariant Representation in Single-Source Domain Generalization 28 May 2024 · 0 repositories · arXiv:2405.18405
-
A One-Layer Decoder-Only Transformer is a Two-Layer RNN: With an Application to Certified Robustness 27 May 2024 · 0 repositories · arXiv:2405.17361
-
Advanced Language Model-based Translator for English-Vietnamese Translation 27 May 2024 · 1 repository
-
Are Self-Attentions Effective for Time Series Forecasting? 27 May 2024 · 1 repository · arXiv:2405.16877Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Assessing LLMs Suitability for Knowledge Graph Completion 27 May 2024 · 1 repository · arXiv:2405.17249
-
Augmenting Textual Generation via Topology Aware Retrieval 27 May 2024 · 0 repositories · arXiv:2405.17602
-
Autoformalizing Euclidean Geometry 27 May 2024 · 1 repository · arXiv:2405.17216Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 2 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Automatic Domain Adaptation by Transformers in In-Context Learning 27 May 2024 · 0 repositories · arXiv:2405.16819
-
BehaviorGPT: Smart Agent Simulation for Autonomous Driving with Next-Patch Prediction 27 May 2024 · 0 repositories · arXiv:2405.17372
-
CHESS: Contextual Harnessing for Efficient SQL Synthesis 27 May 2024 · 2 repositories · arXiv:2405.16755Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Cost-efficient Knowledge-based Question Answering with Large Language Models 27 May 2024 · 0 repositories · arXiv:2405.17337
-
Deciphering Movement: Unified Trajectory Generation Model for Multi-Agent 27 May 2024 · 1 repository · arXiv:2405.17680Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
DeeperImpact: Optimizing Sparse Learned Index Structures 27 May 2024 · 1 repository · arXiv:2405.17093
-
Detecting Deceptive Dark Patterns in E-commerce Platforms 27 May 2024 · 0 repositories · arXiv:2406.01608
-
Exploiting the Layered Intrinsic Dimensionality of Deep Models for Practical Adversarial Training 27 May 2024 · 0 repositories · arXiv:2405.17130
-
Galaxy: A Resource-Efficient Collaborative Edge AI System for In-situ Transformer Inference 27 May 2024 · 0 repositories · arXiv:2405.17245
-
Interesting Scientific Idea Generation using Knowledge Graphs and LLMs: Evaluations with 100 Research Group Leaders 27 May 2024 · 1 repository · arXiv:2405.17044
-
How Do the Architecture and Optimizer Affect Representation Learning? On the Training Dynamics of Representations in Deep Neural Networks 27 May 2024 · 0 repositories · arXiv:2405.17377
-
InversionView: A General-Purpose Method for Reading Information from Neural Activations 27 May 2024 · 1 repository · arXiv:2405.17653Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; the one sample that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)