Methods › General › Stochastic Optimization › Adam › Papers, page 126
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 126 of 244: papers 12,501 to 12,600 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Comparative Study of Pre-Trained BERT Models for Code-Mixed Hindi-English Data 25 May 2023 · 0 repositories · arXiv:2305.15722
-
Concept-Centric Transformers: Enhancing Model Interpretability through Object-Centric Concept Learning within a Shared Global Workspace 25 May 2023 · 2 repositories · arXiv:2305.15775
-
Context-aware attention layers coupled with optimal transport domain adaptation and multimodal fusion methods for recognizing dementia from spontaneous speech 25 May 2023 · 0 repositories · arXiv:2305.16406
-
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective 25 May 2023 · 0 repositories · arXiv:2305.15699
-
Exploiting Noise as a Resource for Computation and Learning in Spiking Neural Networks 25 May 2023 · 1 repository · arXiv:2305.16044Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Imitating Task and Motion Planning with Visuomotor Transformers 25 May 2023 · 0 repositories · arXiv:2305.16309
-
Landmark Attention: Random-Access Infinite Context Length for Transformers 25 May 2023 · 2 repositories · arXiv:2305.16300Syntology official (archive's flag): 1 ran · 11 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
Linguistic Properties of Truthful Response 25 May 2023 · 1 repository · arXiv:2305.15875
-
MERGE: Fast Private Text Generation 25 May 2023 · 1 repository · arXiv:2305.15769
-
Multi-scale Efficient Graph-Transformer for Whole Slide Image Classification 25 May 2023 · 0 repositories · arXiv:2305.15773
-
NexToU: Efficient Topology-Aware U-Net for Medical Image Segmentation 25 May 2023 · 2 repositories · arXiv:2305.15911Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
NODDLE: Node2vec based deep learning model for link prediction 25 May 2023 · 0 repositories · arXiv:2305.16421
-
Not wacky vs. definitely wacky: A study of scalar adverbs in pretrained language models 25 May 2023 · 0 repositories · arXiv:2305.16426
-
On the Tool Manipulation Capability of Open-source Large Language Models 25 May 2023 · 1 repository · arXiv:2305.16504Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation 25 May 2023 · 1 repository · arXiv:2305.15852Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Stecformer: Spatio-temporal Encoding Cascaded Transformer for Multivariate Long-term Time Series Forecasting 25 May 2023 · 0 repositories · arXiv:2305.16370
-
Text-to-Motion Retrieval: Towards Joint Understanding of Human Motion Data and Natural Language 25 May 2023 · 1 repository · arXiv:2305.15842
-
UMat: Uncertainty-Aware Single Image High Resolution Material Capture 25 May 2023 · 0 repositories · arXiv:2305.16312
-
Undetectable Watermarks for Language Models 25 May 2023 · 0 repositories · arXiv:2306.09194
-
UniTRec: A Unified Text-to-Text Transformer and Joint Contrastive Learning Framework for Text-based Recommendation 25 May 2023 · 1 repository · arXiv:2305.15756Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation 25 May 2023 · 0 repositories · arXiv:2305.16107
-
Voyager: An Open-Ended Embodied Agent with Large Language Models 25 May 2023 · 1 repository · arXiv:2305.16291Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
A Joint Time-frequency Domain Transformer for Multivariate Time Series Forecasting 24 May 2023 · 1 repository · arXiv:2305.14649
-
TriMLP: Revenge of a MLP-like Architecture in Sequential Recommendation 24 May 2023 · 1 repository · arXiv:2305.14675
-
A Causal View of Entity Bias in (Large) Language Models 24 May 2023 · 1 repository · arXiv:2305.14695Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A New Era in Software Security: Towards Self-Healing Software via Large Language Models and Formal Verification 24 May 2023 · 1 repository · arXiv:2305.14752
-
A RelEntLess Benchmark for Modelling Graded Relations between Named Entities 24 May 2023 · 0 repositories · arXiv:2305.15002
-
Adversarial Demonstration Attacks on Large Language Models 24 May 2023 · 0 repositories · arXiv:2305.14950
-
LAraBench: Benchmarking Arabic AI with Large Language Models 24 May 2023 · 0 repositories · arXiv:2305.14982
-
ByteSized32: A Corpus and Challenge Task for Generating Task-Specific World Models Expressed as Text Games 24 May 2023 · 1 repository · arXiv:2305.14879Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Chain-of-Questions Training with Latent Answers for Robust Multistep Question Answering 24 May 2023 · 0 repositories · arXiv:2305.14901
-
ChatAgri: Exploring Potentials of ChatGPT on Cross-linguistic Agricultural Text Classification 24 May 2023 · 1 repository · arXiv:2305.15024
-
Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models 24 May 2023 · 0 repositories · arXiv:2305.14763
-
Complex Mathematical Symbol Definition Structures: A Dataset and Model for Coordination Resolution in Definition Extraction 24 May 2023 · 1 repository · arXiv:2305.14660
-
Context-Aware Transformer Pre-Training for Answer Sentence Selection 24 May 2023 · 0 repositories · arXiv:2305.15358
-
A Survey of Diffusion Models in Natural Language Processing 24 May 2023 · 0 repositories · arXiv:2305.14671
-
Don't Take This Out of Context! On the Need for Contextual Models and Evaluations for Stylistic Rewriting 24 May 2023 · 0 repositories · arXiv:2305.14755
-
Don't Trust ChatGPT when Your Question is not in English: A Study of Multilingual Abilities and Types of LLMs 24 May 2023 · 0 repositories · arXiv:2305.16339
-
Dynamic Masking Rate Schedules for MLM Pretraining 24 May 2023 · 0 repositories · arXiv:2305.15096
-
Editing Common Sense in Transformers 24 May 2023 · 1 repository · arXiv:2305.14956Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
ExpertPrompting: Instructing Large Language Models to be Distinguished Experts 24 May 2023 · 2 repositories · arXiv:2305.14688Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Extracting Psychological Indicators Using Question Answering 24 May 2023 · 0 repositories · arXiv:2305.14891
-
Fourier Transformer: Fast Long Range Modeling by Removing Sequence Redundancy with FFT Operator 24 May 2023 · 1 repository · arXiv:2305.15099
-
From Words to Wires: Generating Functioning Electronic Devices from Natural Language Descriptions 24 May 2023 · 1 repository · arXiv:2305.14874Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Ghostbuster: Detecting Text Ghostwritten by Large Language Models 24 May 2023 · 2 repositories · arXiv:2305.15047
-
Gorilla: Large Language Model Connected with Massive APIs 24 May 2023 · 1 repository · arXiv:2305.15334
-
GPTAraEval: A Comprehensive Evaluation of ChatGPT on Arabic NLP 24 May 2023 · 0 repositories · arXiv:2305.14976
-
GTNet: Graph Transformer Network for 3D Point Cloud Classification and Semantic Segmentation 24 May 2023 · 0 repositories · arXiv:2305.15213
-
Harnessing the Power of Large Language Models for Natural Language to First-Order Logic Translation 24 May 2023 · 1 repository · arXiv:2305.15541Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Have LLMs Advanced Enough? A Challenging Problem Solving Benchmark For Large Language Models 24 May 2023 · 1 repository · arXiv:2305.15074Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
How to Distill your BERT: An Empirical Study on the Impact of Weight Initialisation and Distillation Objectives 24 May 2023 · 1 repository · arXiv:2305.15032Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
HuatuoGPT, towards Taming Language Model to Be a Doctor 24 May 2023 · 2 repositories · arXiv:2305.15075Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Psychological Metrics for Dialog System Evaluation 24 May 2023 · 0 repositories · arXiv:2305.14757
-
I Spy a Metaphor: Large Language Models and Diffusion Models Co-Create Visual Metaphors 24 May 2023 · 1 repository · arXiv:2305.14724
-
Inference-Time Policy Adapters (IPA): Tailoring Extreme-Scale LMs without Fine-tuning 24 May 2023 · 1 repository · arXiv:2305.15065Syntology official (archive's flag): 9 ran · 9 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
InterFormer: Interactive Local and Global Features Fusion for Automatic Speech Recognition 24 May 2023 · 0 repositories · arXiv:2305.16342
-
Is GPT-4 a Good Data Analyst? 24 May 2023 · 1 repository · arXiv:2305.15038
-
Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback 24 May 2023 · 1 repository · arXiv:2305.14975
-
KNN-LM Does Not Improve Open-ended Text Generation 24 May 2023 · 0 repositories · arXiv:2305.14625
-
Investigating Table-to-Text Generation Capabilities of LLMs in Real-World Information Seeking Scenarios 24 May 2023 · 2 repositories · arXiv:2305.14987
-
Leveraging GPT-4 for Automatic Translation Post-Editing 24 May 2023 · 0 repositories · arXiv:2305.14878
-
Enabling and Analyzing How to Efficiently Extract Information from Hybrid Long Documents with LLMs 24 May 2023 · 0 repositories · arXiv:2305.16344
-
Leveraging Pre-trained Large Language Models to Construct and Utilize World Models for Model-based Task Planning 24 May 2023 · 0 repositories · arXiv:2305.14909
-
LLMDet: A Third Party Large Language Models Generated Text Detection Tool 24 May 2023 · 1 repository · arXiv:2305.15004Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Mastering the ABCDs of Complex Questions: Answer-Based Claim Decomposition for Fine-grained Self-Evaluation 24 May 2023 · 0 repositories · arXiv:2305.14750
-
Meta-Learning Online Adaptation of Language Models 24 May 2023 · 1 repository · arXiv:2305.15076Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Multi-Modal Mutual Attention and Iterative Interaction for Referring Image Segmentation 24 May 2023 · 0 repositories · arXiv:2305.15302
-
Multiresolution Feature Guidance Based Transformer for Anomaly Detection 24 May 2023 · 0 repositories · arXiv:2305.14880
-
Neural Summarization of Electronic Health Records 24 May 2023 · 0 repositories · arXiv:2305.15222
-
P-vectors: A Parallel-Coupled TDNN/Transformer Network for Speaker Verification 24 May 2023 · 0 repositories · arXiv:2305.14778
-
Peek Across: Improving Multi-Document Modeling via Cross-Document Question-Answering 24 May 2023 · 1 repository · arXiv:2305.15387Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples)
-
Pre-RMSNorm and Pre-CRMSNorm Transformers: Equivalent and Efficient Pre-LN Transformers 24 May 2023 · 1 repository · arXiv:2305.14858Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Predicting Token Impact Towards Efficient Vision Transformer 24 May 2023 · 0 repositories · arXiv:2305.14840
-
AutoPlan: Automatic Planning of Interactive Decision-Making Tasks With Large Language Models 24 May 2023 · 1 repository · arXiv:2305.15064Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Reasoning with Language Model is Planning with World Model 24 May 2023 · 3 repositories · arXiv:2305.14992Syntology 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
RefGPT: Dialogue Generation of GPT, by GPT, and for GPT 24 May 2023 · 1 repository · arXiv:2305.14994
-
Learning UI-to-Code Reverse Generator Using Visual Critic Without Rendering 24 May 2023 · 0 repositories · arXiv:2305.14637
-
Revisiting Token Dropping Strategy in Efficient BERT Pretraining 24 May 2023 · 1 repository · arXiv:2305.15273
-
Segmented Recurrent Transformer: An Efficient Sequence-to-Sequence Model 24 May 2023 · 1 repository · arXiv:2305.16340
-
Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models 24 May 2023 · 1 repository · arXiv:2305.14623Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Semi-Supervised and Long-Tailed Object Detection with CascadeMatch 24 May 2023 · 1 repository · arXiv:2305.14813
-
SPRING: Studying the Paper and Reasoning to Play Games 24 May 2023 · 1 repository · arXiv:2305.15486
-
Testing Causal Models of Word Meaning in GPT-3 and -4 24 May 2023 · 1 repository · arXiv:2305.14630
-
ToMChallenges: A Principle-Guided Dataset and Diverse Evaluation Tasks for Exploring Theory of Mind 24 May 2023 · 1 repository · arXiv:2305.15068
-
Towards Adaptive Prefix Tuning for Parameter-Efficient Language Model Fine-tuning 24 May 2023 · 0 repositories · arXiv:2305.15212
-
Towards Reliable Misinformation Mitigation: Generalization, Uncertainty, and GPT-4 24 May 2023 · 1 repository · arXiv:2305.14928Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Tricking LLMs into Disobedience: Formalizing, Analyzing, and Detecting Jailbreaks 24 May 2023 · 1 repository · arXiv:2305.14965Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Trusting Your Evidence: Hallucinate Less with Context-aware Decoding 24 May 2023 · 3 repositories · arXiv:2305.14739Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Few-shot Adaptation to Distribution Shifts By Mixing Source and Target Embeddings 23 May 2023 · 0 repositories · arXiv:2305.14521
-
When should we prefer Decision Transformers for Offline Reinforcement Learning? 23 May 2023 · 1 repository · arXiv:2305.14550Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages 23 May 2023 · 1 repository · arXiv:2305.14012
-
A Trip Towards Fairness: Bias and De-Biasing in Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.13862
-
Active Learning Principles for In-Context Learning with Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.14264
-
Aligning Large Language Models through Synthetic Feedback 23 May 2023 · 1 repository · arXiv:2305.13735Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
All Roads Lead to Rome? Exploring the Invariance of Transformers' Representations 23 May 2023 · 1 repository · arXiv:2305.14555Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Assessing Linguistic Generalisation in Language Models: A Dataset for Brazilian Portuguese 23 May 2023 · 0 repositories · arXiv:2305.14070
-
Automatic Model Selection with Large Language Models for Reasoning 23 May 2023 · 1 repository · arXiv:2305.14333Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AxomiyaBERTa: A Phonologically-aware Transformer Model for Assamese 23 May 2023 · 1 repository · arXiv:2305.13641
-
Causal Intervention for Abstractive Related Work Generation 23 May 2023 · 0 repositories · arXiv:2305.13685
-
CGCE: A Chinese Generative Chat Evaluation Benchmark for General and Financial Domains 23 May 2023 · 1 repository · arXiv:2305.14471