Methods › General › Stochastic Optimization › Adam › Papers, page 141
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 141 of 244: papers 14,001 to 14,100 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Guiding Online Reinforcement Learning with Action-Free Offline Pretraining 30 Jan 2023 · 1 repository · arXiv:2301.12876Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multi-modal Large Language Model Enhanced Pseudo 3D Perception Framework for Visual Commonsense Reasoning 30 Jan 2023 · 0 repositories · arXiv:2301.13335
-
REPLUG: Retrieval-Augmented Black-Box Language Models 30 Jan 2023 · 3 repositories · arXiv:2301.12652Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Representation biases in sentence transformers 30 Jan 2023 · 0 repositories · arXiv:2301.13039
-
Specializing Smaller Language Models towards Multi-Step Reasoning 30 Jan 2023 · 2 repositories · arXiv:2301.12726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
V2N Service Scaling with Deep Reinforcement Learning 30 Jan 2023 · 0 repositories · arXiv:2301.13324
-
A Discerning Several Thousand Judgments: GPT-3 Rates the Article + Adjective + Numeral + Noun Construction 29 Jan 2023 · 0 repositories · arXiv:2301.12564
-
BERT-based Authorship Attribution on the Romanian Dataset called ROST 29 Jan 2023 · 0 repositories · arXiv:2301.12500
-
Exploring Attention Map Reuse for Efficient Transformer Neural Networks 29 Jan 2023 · 0 repositories · arXiv:2301.12444
-
Global Flood Prediction: a Multimodal Machine Learning Approach 29 Jan 2023 · 0 repositories · arXiv:2301.12548
-
Graph Mixer Networks 29 Jan 2023 · 1 repository · arXiv:2301.12493
-
PhaVIP: Phage VIrion Protein classification based on chaos game representation and Vision Transformer 29 Jan 2023 · 1 repository · arXiv:2301.12422
-
Semantics-enhanced Temporal Graph Networks for Content Popularity Prediction 29 Jan 2023 · 0 repositories · arXiv:2301.12355
-
Towards Vision Transformer Unrolling Fixed-Point Algorithm: a Case Study on Image Restoration 29 Jan 2023 · 0 repositories · arXiv:2301.12332
-
Aerial Image Object Detection With Vision Transformer Detector (ViTDet) 28 Jan 2023 · 1 repository · arXiv:2301.12058
-
Multilingual Sentence Transformer as A Multilingual Word Aligner 28 Jan 2023 · 1 repository · arXiv:2301.12140
-
Predicting Visit Cost of Obstructive Sleep Apnea using Electronic Healthcare Records with Transformer 28 Jan 2023 · 1 repository · arXiv:2301.12289
-
Semantic Tagging with LSTM-CRF 28 Jan 2023 · 0 repositories · arXiv:2301.12206
-
CancerUniT: Towards a Single Unified Model for Effective Detection, Segmentation, and Diagnosis of Eight Major Cancers Using a Large Collection of CT Scans 28 Jan 2023 · 0 repositories · arXiv:2301.12291
-
Towards Equitable Representation in Text-to-Image Synthesis Models with the Cross-Cultural Understanding Benchmark (CCUB) Dataset 28 Jan 2023 · 1 repository · arXiv:2301.12073
-
A Multi-View Joint Learning Framework for Embedding Clinical Codes and Text Using Graph Neural Networks 27 Jan 2023 · 0 repositories · arXiv:2301.11608
-
Can We Use Probing to Better Understand Fine-tuning and Knowledge Distillation of the BERT NLU? 27 Jan 2023 · 0 repositories · arXiv:2301.11688
-
Context Matters: A Strategy to Pre-train Language Model for Science Education 27 Jan 2023 · 0 repositories · arXiv:2301.12031
-
Cross-Architectural Positive Pairs improve the effectiveness of Self-Supervised Learning 27 Jan 2023 · 1 repository · arXiv:2301.12025
-
Predicting Sentence-Level Factuality of News and Bias of Media Outlets 27 Jan 2023 · 1 repository · arXiv:2301.11850
-
The Exploration of Knowledge-Preserving Prompts for Document Summarisation 27 Jan 2023 · 0 repositories · arXiv:2301.11719
-
Large Language Models Are Latent Variable Models: Explaining and Finding Good Demonstrations for In-Context Learning 27 Jan 2023 · 1 repository · arXiv:2301.11916Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Large-Scale Traffic Data Imputation with Spatiotemporal Semantic Understanding 27 Jan 2023 · 0 repositories · arXiv:2301.11691
-
On the Connection Between MPNN and Graph Transformer 27 Jan 2023 · 1 repository · arXiv:2301.11956
-
Pre-training for Speech Translation: CTC Meets Optimal Transport 27 Jan 2023 · 1 repository · arXiv:2301.11716
-
Skeleton-based Action Recognition through Contrasting Two-Stream Spatial-Temporal Networks 27 Jan 2023 · 0 repositories · arXiv:2301.11495
-
SWARM Parallelism: Training Large Models Can Be Surprisingly Communication-Efficient 27 Jan 2023 · 2 repositories · arXiv:2301.11913
-
ThoughtSource: A central hub for large language model reasoning data 27 Jan 2023 · 1 repository · arXiv:2301.11596Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 7 harvested samples)
-
Understanding INT4 Quantization for Transformer Models: Latency Speedup, Composability, and Failure Cases 27 Jan 2023 · 1 repository · arXiv:2301.12017
-
Understanding the Effectiveness of Very Large Language Models on Dialog Evaluation 27 Jan 2023 · 0 repositories · arXiv:2301.12004
-
A benchmark for toxic comment classification on Civil Comments dataset 26 Jan 2023 · 1 repository · arXiv:2301.11125
-
BERT-Embedding and Citation Network Analysis based Query Expansion Technique for Scholarly Search 26 Jan 2023 · 0 repositories · arXiv:2301.11069
-
Causal Reasoning of Entities and Events in Procedural Texts 26 Jan 2023 · 1 repository · arXiv:2301.10896Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 15 harvested samples)
-
MusicLM: Generating Music From Text 26 Jan 2023 · 5 repositories · arXiv:2301.11325Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 3 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples) · 4 pointer-only (licence)
-
Potential Penetrative Pass (P3) 26 Jan 2023 · 0 repositories · arXiv:2302.10760
-
Semantic Segmentation Enhanced Transformer Model for Human Attention Prediction 26 Jan 2023 · 0 repositories · arXiv:2301.11022
-
Enhancing Medical Image Segmentation with TransCeption: A Multi-Scale Feature Fusion Approach 25 Jan 2023 · 1 repository · arXiv:2301.10847
-
ExaRanker: Explanation-Augmented Neural Ranker 25 Jan 2023 · 1 repository · arXiv:2301.10521
-
MLPGradientFlow: going with the flow of multilayer perceptrons (and finding minima fast and accurately) 25 Jan 2023 · 2 repositories · arXiv:2301.10638
-
Qualitative Analysis of a Graph Transformer Approach to Addressing Hate Speech: Adapting to Dynamically Changing Content 25 Jan 2023 · 0 repositories · arXiv:2301.10871
-
Transfer Learning in Deep Learning Models for Building Load Forecasting: Case of Limited Data 25 Jan 2023 · 2 repositories · arXiv:2301.10663
-
ViDeBERTa: A powerful pre-trained language model for Vietnamese 25 Jan 2023 · 1 repository · arXiv:2301.10439Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
A Novel Deep Reinforcement Learning-based Approach for Enhancing Spectral Efficiency of IRS-assisted Wireless Systems 24 Jan 2023 · 0 repositories · arXiv:2302.14706
-
A Stability Analysis of Fine-Tuning a Pre-Trained Model 24 Jan 2023 · 0 repositories · arXiv:2301.09820
-
A Watermark for Large Language Models 24 Jan 2023 · 8 repositories · arXiv:2301.10226Syntology official (archive's flag): 2 ran · 11 ran (of which 7 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 5 pointer-only (licence)
-
Audience-Centric Natural Language Generation via Style Infusion 24 Jan 2023 · 1 repository · arXiv:2301.10283
-
The Next Chapter: A Study of Large Language Models in Storytelling 24 Jan 2023 · 0 repositories · arXiv:2301.09790
-
ClimaX: A foundation model for weather and climate 24 Jan 2023 · 1 repository · arXiv:2301.10343
-
Event Detection in Football using Graph Convolutional Networks 24 Jan 2023 · 0 repositories · arXiv:2301.10052
-
Large Language Models as Fiduciaries: A Case Study Toward Robustly Communicating With Artificial Intelligence Through Legal Standards 24 Jan 2023 · 0 repositories · arXiv:2301.10095
-
Large language models can segment narrative events similarly to humans 24 Jan 2023 · 0 repositories · arXiv:2301.10297
-
Lightweight Neural Architecture Search for Temporal Convolutional Networks at the Edge 24 Jan 2023 · 1 repository · arXiv:2301.10281
-
Multitask Instruction-based Prompting for Fallacy Recognition 24 Jan 2023 · 0 repositories · arXiv:2301.09992
-
SMART: Self-supervised Multi-task pretrAining with contRol Transformers 24 Jan 2023 · 0 repositories · arXiv:2301.09816
-
Accelerating Fair Federated Learning: Adaptive Federated Adam 23 Jan 2023 · 0 repositories · arXiv:2301.09357
-
AI model GPT-3 (dis)informs us better than humans 23 Jan 2023 · 0 repositories · arXiv:2301.11924
-
Deep Learning Mental Health Dialogue System 23 Jan 2023 · 0 repositories · arXiv:2301.09412
-
Efficient Language Model Training through Cross-Lingual and Progressive Transfer Learning 23 Jan 2023 · 1 repository · arXiv:2301.09626
-
Fully transformer-based biomarker prediction from colorectal cancer histology: a large-scale multicentric study 23 Jan 2023 · 2 repositories · arXiv:2301.09617
-
Injecting the BM25 Score as Text Improves BERT-Based Re-rankers 23 Jan 2023 · 1 repository · arXiv:2301.09728
-
ISTVT: Interpretable Spatial-Temporal Video Transformer for Deepfake Detection 23 Jan 2023 · 1 repository
-
Learning to View: Decision Transformers for Active Object Detection 23 Jan 2023 · 0 repositories · arXiv:2301.09544
-
Local Window Attention Transformer for Polarimetric SAR Image Classification 23 Jan 2023 · 1 repository
-
Self-FuseNet: Data Free Unsupervised Remote Sensing Image Super-Resolution 23 Jan 2023 · 0 repositories
-
StockEmotions: Discover Investor Emotions for Financial Sentiment Analysis and Multivariate Time Series 23 Jan 2023 · 2 repositories · arXiv:2301.09279
-
Apples and Oranges? Assessing Image Quality over Content Recognition 22 Jan 2023 · 0 repositories · arXiv:2301.09190
-
Debiasing the Cloze Task in Sequential Recommendation with Bidirectional Transformers 22 Jan 2023 · 1 repository · arXiv:2301.09210
-
Exploring Methods for Building Dialects-Mandarin Code-Mixing Corpora: A Case Study in Taiwanese Hokkien 21 Jan 2023 · 1 repository · arXiv:2301.08937
-
Slice Transformer and Self-supervised Learning for 6DoF Localization in 3D Point Cloud Maps 21 Jan 2023 · 0 repositories · arXiv:2301.08957
-
Stress Test for BERT and Deep Models: Predicting Words from Italian Poetry 21 Jan 2023 · 0 repositories · arXiv:2302.09303
-
SuperScaler: Supporting Flexible DNN Parallelization via a Unified Abstraction 21 Jan 2023 · 0 repositories · arXiv:2301.08984
-
Time-Conditioned Generative Modeling of Object-Centric Representations for Video Decomposition and Prediction 21 Jan 2023 · 1 repository · arXiv:2301.08951
-
Accelerating Multi-Agent Planning Using Graph Transformers with Bounded Suboptimality 20 Jan 2023 · 0 repositories · arXiv:2301.08451
-
Is ChatGPT A Good Translator? Yes With GPT-4 As The Engine 20 Jan 2023 · 1 repository · arXiv:2301.08745
-
On Multi-Agent Deep Deterministic Policy Gradients and their Explainability for SMARTS Environment 20 Jan 2023 · 0 repositories · arXiv:2301.09420
-
Ontology Pre-training for Poison Prediction 20 Jan 2023 · 0 repositories · arXiv:2301.08577
-
Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions 20 Jan 2023 · 2 repositories · arXiv:2301.08810Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Revisiting Estimation Bias in Policy Gradients for Deep Reinforcement Learning 20 Jan 2023 · 0 repositories · arXiv:2301.08442
-
Which Features are Learned by CodeBert: An Empirical Study of the BERT-based Source Code Representation Learning 20 Jan 2023 · 0 repositories · arXiv:2301.08427
-
A Machine Learning Approach for Player and Position Adjusted Expected Goals in Football (Soccer) 19 Jan 2023 · 0 repositories · arXiv:2301.13052
-
Batch Prompting: Efficient Inference with Large Language Model APIs 19 Jan 2023 · 2 repositories · arXiv:2301.08721
-
Diagnose Like a Pathologist: Transformer-Enabled Hierarchical Attention-Guided Multiple Instance Learning for Whole Slide Image Classification 19 Jan 2023 · 1 repository · arXiv:2301.08125
-
FE-TCM: Filter-Enhanced Transformer Click Model for Web Search 19 Jan 2023 · 0 repositories · arXiv:2301.07854
-
MedSegDiff-V2: Diffusion based Medical Image Segmentation with Transformer 19 Jan 2023 · 2 repositories · arXiv:2301.11798Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 2 honoured, 0 violated, 10 with no contract checked; 5 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 10 pointer-only (licence)
-
Sentiment Analysis for Measuring Hope and Fear from Reddit Posts During the 2022 Russo-Ukrainian Conflict 19 Jan 2023 · 0 repositories · arXiv:2301.08347
-
Automated deep reinforcement learning for real-time scheduling strategy of multi-energy system integrated with post-carbon and direct-air carbon captured system 18 Jan 2023 · 0 repositories · arXiv:2301.07768
-
Development, Optimization, and Deployment of Thermal Forward Vision Systems for Advance Vehicular Applications on Edge Devices 18 Jan 2023 · 1 repository · arXiv:2301.07613
-
Impact of the Euro 2020 championship on the spread of COVID-19 18 Jan 2023 · 2 repositories · arXiv:2301.07659
-
Learning-Rate-Free Learning by D-Adaptation 18 Jan 2023 · 1 repository · arXiv:2301.07733Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Cooperation Learning Enhanced Colonic Polyp Segmentation Based on Transformer-CNN Fusion 17 Jan 2023 · 0 repositories · arXiv:2301.06892
-
Expected Gradients of Maxout Networks and Consequences to Parameter Initialization 17 Jan 2023 · 1 repository · arXiv:2301.06956
-
SAT: Size-Aware Transformer for 3D Point Cloud Semantic Segmentation 17 Jan 2023 · 0 repositories · arXiv:2301.06869
-
SwinDepth: Unsupervised Depth Estimation using Monocular Sequences via Swin Transformer and Densely Cascaded Network 17 Jan 2023 · 1 repository · arXiv:2301.06715
-
Tracing and Manipulating Intermediate Values in Neural Math Problem Solvers 17 Jan 2023 · 1 repository · arXiv:2301.06758
-
Transformer Based Implementation for Automatic Book Summarization 17 Jan 2023 · 0 repositories · arXiv:2301.07057