Methods › General › Stochastic Optimization › Adam › Papers, page 144
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 144 of 244: papers 14,301 to 14,400 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Go-tuning: Improving Zero-shot Learning Abilities of Smaller Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10461
-
Hybrid Rule-Neural Coreference Resolution System based on Actor-Critic Learning 20 Dec 2022 · 0 repositories · arXiv:2212.10087
-
Identifying and Manipulating the Personality Traits of Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10276
-
In and Out-of-Domain Text Adversarial Robustness via Label Smoothing 20 Dec 2022 · 0 repositories · arXiv:2212.10258
-
Is GPT-3 a Good Data Annotator? 20 Dec 2022 · 1 repository · arXiv:2212.10450
-
Evaluating Psychological Safety of Large Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10529
-
Large Language Models Are Reasoning Teachers 20 Dec 2022 · 1 repository · arXiv:2212.10071Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
METEOR Guided Divergence for Video Captioning 20 Dec 2022 · 1 repository · arXiv:2212.10690
-
PairReranker: Pairwise Reranking for Natural Language Generation 20 Dec 2022 · 0 repositories · arXiv:2212.10555
-
Parameter-efficient Zero-shot Transfer for Cross-Language Dense Retrieval with Adapters 20 Dec 2022 · 0 repositories · arXiv:2212.10448
-
Pay Attention to Your Tone: Introducing a New Dataset for Polite Language Rewrite 20 Dec 2022 · 1 repository · arXiv:2212.10190
-
Pretraining Without Attention 20 Dec 2022 · 1 repository · arXiv:2212.10544Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
True Detective: A Deep Abductive Reasoning Benchmark Undoable for GPT-3 and Challenging for GPT-4 20 Dec 2022 · 0 repositories · arXiv:2212.10114
-
Why Can GPT Learn In-Context? Language Models Implicitly Perform Gradient Descent as Meta-Optimizers 20 Dec 2022 · 1 repository · arXiv:2212.10559
-
Denoising instrumented mouthguard measurements of head impact kinematics with a convolutional neural network 19 Dec 2022 · 0 repositories · arXiv:2212.09832
-
Empowering Diffusion Models on the Embedding Space for Text Generation 19 Dec 2022 · 1 repository · arXiv:2212.09412Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 3 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
Do CoNLL-2003 Named Entity Taggers Still Work Well in 2023? 19 Dec 2022 · 1 repository · arXiv:2212.09747
-
Emergent Analogical Reasoning in Large Language Models 19 Dec 2022 · 2 repositories · arXiv:2212.09196Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Enriching Relation Extraction with OpenIE 19 Dec 2022 · 0 repositories · arXiv:2212.09376
-
Evaluating Human-Language Model Interaction 19 Dec 2022 · 1 repository · arXiv:2212.09746
-
Large Language Models are Better Reasoners with Self-Verification 19 Dec 2022 · 1 repository · arXiv:2212.09561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
LENS: A Learnable Evaluation Metric for Text Simplification 19 Dec 2022 · 1 repository · arXiv:2212.09739
-
Less is More: Parameter-Free Text Classification with Gzip 19 Dec 2022 · 0 repositories · arXiv:2212.09410
-
MANTIS at TSAR-2022 Shared Task: Improved Unsupervised Lexical Simplification with Pretrained Encoders 19 Dec 2022 · 0 repositories · arXiv:2212.09855
-
MIST: Multi-modal Iterative Spatial-Temporal Transformer for Long-form Video Question Answering 19 Dec 2022 · 1 repository · arXiv:2212.09522
-
Reasoning with Language Model Prompting: A Survey 19 Dec 2022 · 2 repositories · arXiv:2212.09597
-
SrTR: Self-reasoning Transformer with Visual-linguistic Knowledge for Scene Graph Generation 19 Dec 2022 · 0 repositories · arXiv:2212.09329
-
The case for 4-bit precision: k-bit Inference Scaling Laws 19 Dec 2022 · 1 repository · arXiv:2212.09720
-
Tokenization Consistency Matters for Generative Models on Extractive NLP Tasks 19 Dec 2022 · 1 repository · arXiv:2212.09912
-
Bort: Towards Explainable Neural Networks with Bounded Orthogonal Constraint 18 Dec 2022 · 1 repository · arXiv:2212.09062Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can Retriever-Augmented Language Models Reason? The Blame Game Between the Retriever and the Language Model 18 Dec 2022 · 1 repository · arXiv:2212.09146
-
Fast FullSubNet: Accelerate Full-band and Sub-band Fusion Model for Single-channel Speech Enhancement 18 Dec 2022 · 1 repository · arXiv:2212.09019Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
Neural Coreference Resolution based on Reinforcement Learning 18 Dec 2022 · 0 repositories · arXiv:2212.09028
-
Neural Rankers for Effective Screening Prioritisation in Medical Systematic Review Literature Search 18 Dec 2022 · 0 repositories · arXiv:2212.09017
-
Style-Hallucinated Dual Consistency Learning: A Unified Framework for Visual Domain Generalization 18 Dec 2022 · 1 repository · arXiv:2212.09068
-
Claim Optimization in Computational Argumentation 17 Dec 2022 · 1 repository · arXiv:2212.08913
-
Improving Levenberg-Marquardt Algorithm for Neural Networks 17 Dec 2022 · 1 repository · arXiv:2212.08769
-
Leveraging Wastewater Monitoring for COVID-19 Forecasting in the US: a Deep Learning study 17 Dec 2022 · 1 repository · arXiv:2212.08798
-
Autoencoders as Cross-Modal Teachers: Can Pretrained 2D Image Transformers Help 3D Representation Learning? 16 Dec 2022 · 4 repositories · arXiv:2212.08320
-
Convolution-enhanced Evolving Attention Networks 16 Dec 2022 · 1 repository · arXiv:2212.08330
-
From Xception to NEXcepTion: New Design Decisions and Neural Architecture Search 16 Dec 2022 · 1 repository · arXiv:2212.08448
-
Homonymy Information for English WordNet 16 Dec 2022 · 1 repository · arXiv:2212.08388
-
LOANet: A Lightweight Network Using Object Attention for Extracting Buildings and Roads from UAV Aerial Remote Sensing Images 16 Dec 2022 · 1 repository · arXiv:2212.08490
-
LegalRelectra: Mixed-domain Language Modeling for Long-range Legal Text Comprehension 16 Dec 2022 · 0 repositories · arXiv:2212.08204
-
MURMUR: Modular Multi-Step Reasoning for Semi-Structured Data-to-Text Generation 16 Dec 2022 · 0 repositories · arXiv:2212.08607
-
Plansformer: Generating Symbolic Plans using Transformers 16 Dec 2022 · 0 repositories · arXiv:2212.08681
-
POIBERT: A Transformer-based Model for the Tour Recommendation Problem 16 Dec 2022 · 0 repositories · arXiv:2212.13900
-
ReCo: Reliable Causal Chain Reasoning via Structural Causal Recurrent Neural Networks 16 Dec 2022 · 1 repository · arXiv:2212.08322Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Assessing the Impact of Sequence Length Learning on Classification Tasks for Transformer Encoder Models 16 Dec 2022 · 0 repositories · arXiv:2212.08399
-
Rethinking Cooking State Recognition with Vision Transformers 16 Dec 2022 · 1 repository · arXiv:2212.08586
-
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA 16 Dec 2022 · 1 repository · arXiv:2212.08635
-
Utilizing distilBert transformer model for sentiment classification of COVID-19's Persian open-text responses 16 Dec 2022 · 0 repositories · arXiv:2212.08407
-
Efficient Long Sequence Modeling via State Space Augmented Transformer 15 Dec 2022 · 1 repository · arXiv:2212.08136Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Efficient Pre-training of Masked Language Model via Concept-based Curriculum Masking 15 Dec 2022 · 1 repository · arXiv:2212.07617Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Huber-energy measure quantization 15 Dec 2022 · 0 repositories · arXiv:2212.08162
-
HUM3DIL: Semi-supervised Multi-modal 3D Human Pose Estimation for Autonomous Driving 15 Dec 2022 · 0 repositories · arXiv:2212.07729
-
CLIPPO: Image-and-Language Understanding from Pixels Only 15 Dec 2022 · 1 repository · arXiv:2212.08045
-
Revisiting the Gold Standard: Grounding Summarization Evaluation with Robust Human Evaluation 15 Dec 2022 · 2 repositories · arXiv:2212.07981Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Scaffold-Based Multi-Objective Drug Candidate Optimization 15 Dec 2022 · 0 repositories · arXiv:2301.07175
-
Sim-to-Real Transfer for Quadrupedal Locomotion via Terrain Transformer 15 Dec 2022 · 0 repositories · arXiv:2212.07740
-
Visually-augmented pretrained language models for NLP tasks without images 15 Dec 2022 · 1 repository · arXiv:2212.07937
-
Most Important Person-guided Dual-branch Cross-Patch Attention for Group Affect Recognition 14 Dec 2022 · 0 repositories · arXiv:2212.07055
-
Efficient Exploration in Resource-Restricted Reinforcement Learning 14 Dec 2022 · 0 repositories · arXiv:2212.06988
-
Efficient Self-supervised Learning with Contextualized Target Representations for Vision, Speech and Language 14 Dec 2022 · 5 repositories · arXiv:2212.07525Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Explainability of Text Processing and Retrieval Methods: A Critical Survey 14 Dec 2022 · 0 repositories · arXiv:2212.07126
-
Hierarchical Strategies for Cooperative Multi-Agent Reinforcement Learning 14 Dec 2022 · 0 repositories · arXiv:2212.07397
-
VTCC-NLP at NL4Opt competition subtask 1: An Ensemble Pre-trained language models for Named Entity Recognition 14 Dec 2022 · 0 repositories · arXiv:2212.07219
-
CREPE: Can Vision-Language Foundation Models Reason Compositionally? 13 Dec 2022 · 1 repository · arXiv:2212.07796Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Generative artificial intelligence-enabled dynamic detection of nicotine-related circuits 13 Dec 2022 · 0 repositories · arXiv:2212.06330
-
GPViT: A High Resolution Non-Hierarchical Vision Transformer with Group Propagation 13 Dec 2022 · 2 repositories · arXiv:2212.06795
-
Paraphrase Identification with Deep Learning: A Review of Datasets and Methods 13 Dec 2022 · 0 repositories · arXiv:2212.06933
-
RT-1: Robotics Transformer for Real-World Control at Scale 13 Dec 2022 · 1 repository · arXiv:2212.06817Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Smart Journey in Istanbul: A Mobile Application in Smart Cities for Traffic Estimation by Harnessing Time Series 13 Dec 2022 · 0 repositories · arXiv:2212.09448
-
DexBERT: Effective, Task-Agnostic and Fine-grained Representation Learning of Android Bytecode 12 Dec 2022 · 1 repository · arXiv:2212.05976
-
Automated ICD Coding using Extreme Multi-label Long Text Transformer-based Models 12 Dec 2022 · 1 repository · arXiv:2212.05857
-
BeautyREC: Robust, Efficient, and Content-preserving Makeup Transfer 12 Dec 2022 · 0 repositories · arXiv:2212.05855
-
Classifying the Ideological Orientation of User-Submitted Texts in Social Media 12 Dec 2022 · 1 repository
-
CTT-Net: A Multi-view Cross-token Transformer for Cataract Postoperative Visual Acuity Prediction 12 Dec 2022 · 1 repository · arXiv:2212.05794
-
Deep learning approaches to building rooftop thermal bridge detection from aerial images 12 Dec 2022 · 1 repository
-
NMS Strikes Back 12 Dec 2022 · 1 repository · arXiv:2212.06137
-
P-Transformer: Towards Better Document-to-Document Neural Machine Translation 12 Dec 2022 · 1 repository · arXiv:2212.05830
-
ROIFormer: Semantic-Aware Region of Interest Transformer for Efficient Self-Supervised Monocular Depth Estimation 12 Dec 2022 · 0 repositories · arXiv:2212.05729
-
Video Prediction by Efficient Transformers 12 Dec 2022 · 1 repository · arXiv:2212.06026Syntology official: harvested, nothing ran · 0 ran · 7 unverified (of 7 harvested samples)
-
Off-Policy Deep Reinforcement Learning Algorithms for Handling Various Robotic Manipulator Tasks 11 Dec 2022 · 0 repositories · arXiv:2212.05572
-
Elixir: Train a Large Language Model on a Small GPU Cluster 10 Dec 2022 · 2 repositories · arXiv:2212.05339
-
Joint Spatio-Temporal Modeling for the Semantic Change Detection in Remote Sensing Images 10 Dec 2022 · 3 repositories · arXiv:2212.05245Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Thinking Fast and Slow in Large Language Models 10 Dec 2022 · 0 repositories · arXiv:2212.05206
-
MAGVIT: Masked Generative Video Transformer 10 Dec 2022 · 1 repository · arXiv:2212.05199Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Position Embedding Needs an Independent Layer Normalization 10 Dec 2022 · 1 repository · arXiv:2212.05262
-
Punctuation Restoration for Singaporean Spoken Languages: English, Malay, and Mandarin 10 Dec 2022 · 1 repository · arXiv:2212.05356
-
SMILE: Scaling Mixture-of-Experts with Efficient Bi-level Routing 10 Dec 2022 · 0 repositories · arXiv:2212.05191
-
Structured information extraction from complex scientific text with fine-tuned large language models 10 Dec 2022 · 0 repositories · arXiv:2212.05238
-
Dynamic Test-Time Augmentation via Differentiable Functions 9 Dec 2022 · 1 repository · arXiv:2212.04681
-
Incorporating Emotions into Health Mention Classification Task on Social Media 9 Dec 2022 · 1 repository · arXiv:2212.05039
-
Masked Lip-Sync Prediction by Audio-Visual Contextual Exploitation in Transformers 9 Dec 2022 · 0 repositories · arXiv:2212.04970
-
MIMO Is All You Need : A Strong Multi-In-Multi-Out Baseline for Video Prediction 9 Dec 2022 · 1 repository · arXiv:2212.04655
-
Cross-Domain Synthetic-to-Real In-the-Wild Depth and Normal Estimation for 3D Scene Understanding 9 Dec 2022 · 0 repositories · arXiv:2212.05040
-
RCDT: Relational Remote Sensing Change Detection with Transformer 9 Dec 2022 · 1 repository · arXiv:2212.04869
-
Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints 9 Dec 2022 · 1 repository · arXiv:2212.05055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Turing Deception 9 Dec 2022 · 0 repositories · arXiv:2212.06721