Methods › General › Stochastic Optimization › Adam › Papers, page 125
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 125 of 244: papers 12,401 to 12,500 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning 30 May 2023 · 1 repository · arXiv:2305.19426
-
Seeing Seeds Beyond Weeds: Green Teaming Generative AI for Beneficial Uses 30 May 2023 · 0 repositories · arXiv:2306.03097
-
Self-Verification Improves Few-Shot Clinical Information Extraction 30 May 2023 · 1 repository · arXiv:2306.00024Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Smooth, exact rotational symmetrization for deep learning on point clouds 30 May 2023 · 0 repositories · arXiv:2305.19302
-
Subequivariant Graph Reinforcement Learning in 3D Environments 30 May 2023 · 1 repository · arXiv:2305.18951Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 13 unverified (of 18 harvested samples)
-
Optimizing Attention and Cognitive Control Costs Using Temporally-Layered Architectures 30 May 2023 · 1 repository · arXiv:2305.18701
-
Abstractive Summarization as Augmentation for Document-Level Event Detection 29 May 2023 · 0 repositories · arXiv:2305.18023
-
Action valuation of on- and off-ball soccer players based on multi-agent deep reinforcement learning 29 May 2023 · 0 repositories · arXiv:2305.17886
-
Alignment-free HDR Deghosting with Semantics Consistent Transformer 29 May 2023 · 0 repositories · arXiv:2305.18135
-
Approximation Rate of the Transformer Architecture for Sequence Modeling 29 May 2023 · 0 repositories · arXiv:2305.18475
-
Attention Mechanisms in Medical Image Segmentation: A Survey 29 May 2023 · 0 repositories · arXiv:2305.17937
-
Chatbots to ChatGPT in a Cybersecurity Space: Evolution, Vulnerabilities, Attacks, Challenges, and Future Recommendations 29 May 2023 · 0 repositories · arXiv:2306.09255
-
Check-COVID: Fact-Checking COVID-19 News Claims with Scientific Evidence 29 May 2023 · 1 repository · arXiv:2305.18265
-
Coeditor: Leveraging Contextual Changes for Multi-round Code Auto-editing 29 May 2023 · 0 repositories · arXiv:2305.18584
-
Controllable Text-to-Image Generation with GPT-4 29 May 2023 · 0 repositories · arXiv:2305.18583
-
Diffusion Model is an Effective Planner and Data Synthesizer for Multi-Task Reinforcement Learning 29 May 2023 · 1 repository · arXiv:2305.18459Syntology official (archive's flag): 15 ran · 19 ran (of which 3 constructed an object rather than computing a result; 18 with no instrument failure: 5 honoured, 0 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 11 unverified (of 30 harvested samples) · 4 pointer-only (licence)
-
Do Language Models Know When They're Hallucinating References? 29 May 2023 · 1 repository · arXiv:2305.18248Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Do Large Language Models Know What They Don't Know? 29 May 2023 · 1 repository · arXiv:2305.18153Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Exploring Effectiveness of GPT-3 in Grammatical Error Correction: A Study on Performance and Controllability in Prompt-Based Methods 29 May 2023 · 0 repositories · arXiv:2305.18156
-
From Adversarial Arms Race to Model-centric Evaluation: Motivating a Unified Automatic Robustness Evaluation Framework 29 May 2023 · 1 repository · arXiv:2305.18503Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Game of Tones: Faculty detection of GPT-4 generated content in university assessments 29 May 2023 · 0 repositories · arXiv:2305.18081
-
Large Language Models are not Fair Evaluators 29 May 2023 · 1 repository · arXiv:2305.17926Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LM-CPPF: Paraphrasing-Guided Data Augmentation for Contrastive Prompt-Based Few-Shot Fine-Tuning 29 May 2023 · 1 repository · arXiv:2305.18169
-
Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models 29 May 2023 · 1 repository · arXiv:2305.18189Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ProcessGPT: Transforming Business Process Management with Generative Artificial Intelligence 29 May 2023 · 0 repositories · arXiv:2306.01771
-
Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo 29 May 2023 · 1 repository · arXiv:2305.18246Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
SlimFit: Memory-Efficient Fine-Tuning of Transformer-based Models Using Training Dynamics 29 May 2023 · 0 repositories · arXiv:2305.18513
-
Syntax and Semantics Meet in the "Middle": Probing the Syntax-Semantics Interface of LMs Through Agentivity 29 May 2023 · 1 repository · arXiv:2305.18185
-
Test-Time Training on Nearest Neighbors for Large Language Models 29 May 2023 · 1 repository · arXiv:2305.18466Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
The Rise of AI Language Pathologists: Exploring Two-level Prompt Learning for Few-shot Weakly-supervised Whole Slide Image Classification 29 May 2023 · 1 repository · arXiv:2305.17891Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Transformer Language Models Handle Word Frequency in Prediction Head 29 May 2023 · 0 repositories · arXiv:2305.18294
-
Geometric Algebra Transformer 28 May 2023 · 2 repositories · arXiv:2305.18415Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A New Deep Learning Architecture withInductive Bias Balance for Transformer Oil Temperature Forecasting 28 May 2023 · 1 repository
-
A Quantitative Review on Language Model Efficiency Research 28 May 2023 · 0 repositories · arXiv:2306.01768
-
ASU-CNN: An Efficient Deep Architecture for Image Classification and Feature Visualizations 28 May 2023 · 0 repositories · arXiv:2305.19146
-
Bridging the Language Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs 28 May 2023 · 0 repositories · arXiv:2305.17740
-
DPFormer: Learning Differentially Private Transformer on Long-Tailed Data 28 May 2023 · 0 repositories · arXiv:2305.17633
-
Evaluating GPT-3 Generated Explanations for Hateful Content Moderation 28 May 2023 · 1 repository · arXiv:2305.17680
-
Generating EDU Extracts for Plan-Guided Summary Re-Ranking 28 May 2023 · 1 repository · arXiv:2305.17779Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Knowledge-Augmented Reasoning Distillation for Small Language Models in Knowledge-Intensive Tasks 28 May 2023 · 1 repository · arXiv:2305.18395Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model Application 28 May 2023 · 1 repository · arXiv:2305.17701
-
Large Language Models, scientific knowledge and factuality: A framework to streamline human expert evaluation 28 May 2023 · 1 repository · arXiv:2305.17819
-
Mitigating Label Biases for In-context Learning 28 May 2023 · 1 repository · arXiv:2305.19148Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Reconstructing Sea Surface Temperature Images: A Masked Autoencoder Approach for Cloud Masking and Reconstruction 28 May 2023 · 0 repositories · arXiv:2306.00835
-
Rethinking Masked Language Modeling for Chinese Spelling Correction 28 May 2023 · 1 repository · arXiv:2305.17721
-
Reward Collapse in Aligning Large Language Models 28 May 2023 · 1 repository · arXiv:2305.17608
-
SQuARe: A Large-Scale Dataset of Sensitive Questions and Acceptable Responses Created Through Human-Machine Collaboration 28 May 2023 · 1 repository · arXiv:2305.17696
-
Transfer Learning for Power Outage Detection Task with Limited Training Data 28 May 2023 · 0 repositories · arXiv:2305.17817
-
Diagnosing Transformers: Illuminating Feature Spaces for Clinical Decision-Making 27 May 2023 · 1 repository · arXiv:2305.17588
-
Analysis over vision-based models for pedestrian action anticipation 27 May 2023 · 0 repositories · arXiv:2305.17451
-
Bridging the Granularity Gap for Acoustic Modeling 27 May 2023 · 1 repository · arXiv:2305.17356
-
Complementary and Integrative Health Lexicon (CIHLex) and Entity Recognition in the Literature 27 May 2023 · 0 repositories · arXiv:2305.17353
-
CrossGET: Cross-Guided Ensemble of Tokens for Accelerating Vision-Language Transformers 27 May 2023 · 1 repository · arXiv:2305.17455Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text 27 May 2023 · 1 repository · arXiv:2305.17359Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Graph Inductive Biases in Transformers without Message Passing 27 May 2023 · 2 repositories · arXiv:2305.17589Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
HyperFormer: Learning Expressive Sparse Feature Representations via Hypergraph Transformer 27 May 2023 · 0 repositories · arXiv:2305.17386
-
Is Centralized Training with Decentralized Execution Framework Centralized Enough for MARL? 27 May 2023 · 1 repository · arXiv:2305.17352Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Curse of Recursion: Training on Generated Data Makes Models Forget 27 May 2023 · 1 repository · arXiv:2305.17493
-
Modeling Adversarial Attack on Pre-trained Language Models as Sequential Decision Making 27 May 2023 · 1 repository · arXiv:2305.17440
-
Reinforcement Learning With Reward Machines in Stochastic Games 27 May 2023 · 0 repositories · arXiv:2305.17372
-
Scalable Transformer for PDE Surrogate Modeling 27 May 2023 · 1 repository · arXiv:2305.17560
-
SwiftSage: A Generative Agent with Fast and Slow Thinking for Complex Interactive Tasks 27 May 2023 · 2 repositories · arXiv:2305.17390Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Towards Explainable Conversational Recommender Systems 27 May 2023 · 1 repository · arXiv:2305.18363
-
What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks 27 May 2023 · 1 repository · arXiv:2305.18365
-
Zero-TPrune: Zero-Shot Token Pruning through Leveraging of the Attention Graph in Pre-Trained Transformers 27 May 2023 · 0 repositories · arXiv:2305.17328
-
AlignScore: Evaluating Factual Consistency with a Unified Alignment Function 26 May 2023 · 2 repositories · arXiv:2305.16739
-
Backpack Language Models 26 May 2023 · 1 repository · arXiv:2305.16765Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks 26 May 2023 · 1 repository · arXiv:2305.17100Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 7 pointer-only (licence)
-
Calibration of Transformer-based Models for Identifying Stress and Depression in Social Media 26 May 2023 · 0 repositories · arXiv:2305.16797
-
Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance 26 May 2023 · 1 repository · arXiv:2305.17306Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
ChatGPT: A Study on its Utility for Ubiquitous Software Engineering Tasks 26 May 2023 · 0 repositories · arXiv:2305.16837
-
COMCAT: Towards Efficient Compression and Customization of Attention-Based Vision Models 26 May 2023 · 1 repository · arXiv:2305.17235Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Counterfactual reasoning: Testing language models' understanding of hypothetical scenarios 26 May 2023 · 1 repository · arXiv:2305.16572
-
Distinguishing Human Generated Text From ChatGPT Generated Text Using Machine Learning 26 May 2023 · 0 repositories · arXiv:2306.01761
-
Do GPTs Produce Less Literal Translations? 26 May 2023 · 1 repository · arXiv:2305.16806
-
Emergent Agentic Transformer from Chain of Hindsight Experience 26 May 2023 · 0 repositories · arXiv:2305.16554
-
Evaluation of Question Generation Needs More References 26 May 2023 · 0 repositories · arXiv:2305.16626
-
Future-conditioned Unsupervised Pretraining for Decision Transformer 26 May 2023 · 1 repository · arXiv:2305.16683Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Generalizing Adam to Manifolds for Efficiently Training Transformers 26 May 2023 · 1 repository · arXiv:2305.16901
-
GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot Attention for Vision-and-Language Navigation 26 May 2023 · 1 repository · arXiv:2305.17102
-
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing 26 May 2023 · 0 repositories · arXiv:2305.16635
-
Improving accuracy of GPT-3/4 results on biomedical data using a retrieval-augmented language model 26 May 2023 · 0 repositories · arXiv:2305.17116
-
Improving Position Encoding of Transformers for Multivariate Time Series Classification 26 May 2023 · 1 repository · arXiv:2305.16642
-
KNSE: A Knowledge-aware Natural Language Inference Framework for Dialogue Symptom Status Recognition 26 May 2023 · 0 repositories · arXiv:2305.16833
-
Large Language Models as Tool Makers 26 May 2023 · 1 repository · arXiv:2305.17126
-
Learning and Leveraging Verifiers to Improve Planning Capabilities of Pre-trained Language Models 26 May 2023 · 0 repositories · arXiv:2305.17077
-
Learning to Imagine: Visually-Augmented Natural Language Generation 26 May 2023 · 1 repository · arXiv:2305.16944Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
LLMs and the Abstraction and Reasoning Corpus: Successes, Failures, and the Importance of Object-based Representations 26 May 2023 · 1 repository · arXiv:2305.18354
-
NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models 26 May 2023 · 2 repositories · arXiv:2305.16986Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Neural Task Synthesis for Visual Programming 26 May 2023 · 1 repository · arXiv:2305.18342
-
On Evaluating Adversarial Robustness of Large Vision-Language Models 26 May 2023 · 1 repository · arXiv:2305.16934Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 2 pointer-only (licence)
-
Playing repeated games with Large Language Models 26 May 2023 · 0 repositories · arXiv:2305.16867
-
Rotational Equilibrium: How Weight Decay Balances Learning Across Neural Networks 26 May 2023 · 2 repositories · arXiv:2305.17212Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Theoretical and Practical Perspectives on what Influence Functions Do 26 May 2023 · 0 repositories · arXiv:2305.16971
-
TranSFormer: Slow-Fast Transformer for Machine Translation 26 May 2023 · 0 repositories · arXiv:2305.16982
-
XGrad: Boosting Gradient-Based Optimizers With Weight Prediction 26 May 2023 · 1 repository · arXiv:2305.18240
-
Zero is Not Hero Yet: Benchmarking Zero-Shot Performance of LLMs for Financial Tasks 26 May 2023 · 1 repository · arXiv:2305.16633
-
A Survey on ChatGPT: AI-Generated Contents, Challenges, and Solutions 25 May 2023 · 0 repositories · arXiv:2305.18339
-
Asking Before Acting: Gather Information in Embodied Decision Making with Language Models 25 May 2023 · 0 repositories · arXiv:2305.15695
-
ChatGPT for PLC/DCS Control Logic Generation 25 May 2023 · 0 repositories · arXiv:2305.15809