Methods › General › Stochastic Optimization › Adam › Papers, page 148
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 148 of 244: papers 14,701 to 14,800 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards Reasoning-Aware Explainable VQA 9 Nov 2022 · 0 repositories · arXiv:2211.05190
-
Training a Vision Transformer from scratch in less than 24 hours with 1 GPU 9 Nov 2022 · 1 repository · arXiv:2211.05187Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
ViTALiTy: Unifying Low-rank and Sparse Approximation for Vision Transformer Acceleration with a Linear Taylor Attention 9 Nov 2022 · 1 repository · arXiv:2211.05109Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
A Multimodal Approach for Dementia Detection from Spontaneous Speech with Tensor Fusion Layer 8 Nov 2022 · 0 repositories · arXiv:2211.04368
-
flexBART: Flexible Bayesian regression trees with categorical predictors 8 Nov 2022 · 1 repository · arXiv:2211.04459
-
Active Example Selection for In-Context Learning 8 Nov 2022 · 1 repository · arXiv:2211.04486Syntology official (archive's flag): 10 ran · 10 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples)
-
Discover, Explanation, Improvement: An Automatic Slice Detection Framework for Natural Language Processing 8 Nov 2022 · 0 repositories · arXiv:2211.04476
-
Going for GOAL: A Resource for Grounded Football Commentaries 8 Nov 2022 · 1 repository · arXiv:2211.04534
-
Linear Self-Attention Approximation via Trainable Feedforward Kernel 8 Nov 2022 · 0 repositories · arXiv:2211.04076
-
On the Algorithmic Stability and Generalization of Adaptive Optimization Methods 8 Nov 2022 · 0 repositories · arXiv:2211.03970
-
SimOn: A Simple Framework for Online Temporal Action Localization 8 Nov 2022 · 1 repository · arXiv:2211.04905
-
Splitting expands the application range of Vision Transformer -- variable Vision Transformer (vViT) 8 Nov 2022 · 0 repositories · arXiv:2211.03992
-
AD-BERT: Using Pre-trained contextualized embeddings to Predict the Progression from Mild Cognitive Impairment to Alzheimer's Disease 7 Nov 2022 · 0 repositories · arXiv:2212.06042
-
Retrieval augmentation of large language models for lay language generation 7 Nov 2022 · 1 repository · arXiv:2211.03818Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
CoNMix for Source-free Single and Multi-target Domain Adaptation 7 Nov 2022 · 1 repository · arXiv:2211.03876Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining 7 Nov 2022 · 0 repositories · arXiv:2211.03594
-
How Much Does Attention Actually Attend? Questioning the Importance of Attention in Pretrained Transformers 7 Nov 2022 · 1 repository · arXiv:2211.03495
-
Sequential Transformer for End-to-End Person Search 6 Nov 2022 · 0 repositories · arXiv:2211.04323
-
Suffix Retrieval-Augmented Language Modeling 6 Nov 2022 · 1 repository · arXiv:2211.03053
-
Wall Street Tree Search: Risk-Aware Planning for Offline Reinforcement Learning 6 Nov 2022 · 0 repositories · arXiv:2211.04583
-
Inductive Graph Transformer for Delivery Time Estimation 5 Nov 2022 · 1 repository · arXiv:2211.02863
-
Learning to Infer from Unlabeled Data: A Semi-supervised Learning Approach for Robust Natural Language Inference 5 Nov 2022 · 1 repository · arXiv:2211.02971
-
A Transformer Architecture for Online Gesture Recognition of Mathematical Expressions 4 Nov 2022 · 0 repositories · arXiv:2211.02643
-
A Weakly-Supervised Streaming Multilingual Speech Model with Truly Zero-Shot Capability 4 Nov 2022 · 0 repositories · arXiv:2211.02499
-
BERT-Deep CNN: State-of-the-Art for Sentiment Analysis of COVID-19 Tweets 4 Nov 2022 · 0 repositories · arXiv:2211.09733
-
BERT for Long Documents: A Case Study of Automated ICD Coding 4 Nov 2022 · 0 repositories · arXiv:2211.02519
-
CCATMos: Convolutional Context-aware Transformer Network for Non-intrusive Speech Quality Assessment 4 Nov 2022 · 0 repositories · arXiv:2211.02577
-
Continuous Prompt Tuning Based Textual Entailment Model for E-commerce Entity Typing 4 Nov 2022 · 1 repository · arXiv:2211.02483
-
Deep learning for structural health monitoring: An application to heritage structures 4 Nov 2022 · 0 repositories · arXiv:2211.10351
-
Fraudulent User Detection Via Behavior Information Aggregation Network (BIAN) On Large-Scale Financial Social Network 4 Nov 2022 · 0 repositories · arXiv:2211.06315
-
Generation of Chinese classical poetry based on pre-trained model 4 Nov 2022 · 0 repositories · arXiv:2211.02541
-
OSIC: A New One-Stage Image Captioner Coined 4 Nov 2022 · 0 repositories · arXiv:2211.02321
-
Patch DCT vs LeNet 4 Nov 2022 · 0 repositories · arXiv:2211.02392
-
Real-Time Target Sound Extraction 4 Nov 2022 · 1 repository · arXiv:2211.02250Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
SPEAKER VGG CCT: Cross-corpus Speech Emotion Recognition with Speaker Embedding and Vision Transformers 4 Nov 2022 · 1 repository · arXiv:2211.02366
-
Channel-Aware Pretraining of Joint Encoder-Decoder Self-Supervised Model for Telephonic-Speech ASR 3 Nov 2022 · 0 repositories · arXiv:2211.01669
-
Transformers on Multilingual Clause-Level Morphology 3 Nov 2022 · 1 repository · arXiv:2211.01736
-
FedTP: Federated Learning by Transformer Personalization 3 Nov 2022 · 1 repository · arXiv:2211.01572Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Fine-Tuning Language Models via Epistemic Neural Networks 3 Nov 2022 · 1 repository · arXiv:2211.01568Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Graph-Based Multi-Camera Soccer Player Tracker 3 Nov 2022 · 0 repositories · arXiv:2211.02125
-
Pangu-Weather: A 3D High-Resolution Model for Fast and Accurate Global Weather Forecast 3 Nov 2022 · 6 repositories · arXiv:2211.02556
-
PolyBuilding: Polygon Transformer for End-to-End Building Extraction 3 Nov 2022 · 0 repositories · arXiv:2211.01589
-
SAP-DETR: Bridging the Gap Between Salient Points and Queries-Based Transformer Detector for Fast Model Convergency 3 Nov 2022 · 1 repository · arXiv:2211.02006
-
Using Large Pre-Trained Language Model to Assist FDA in Premarket Medical Device 3 Nov 2022 · 0 repositories · arXiv:2212.01217
-
Attention-based Neural Cellular Automata 2 Nov 2022 · 0 repositories · arXiv:2211.01233
-
BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder 2 Nov 2022 · 0 repositories · arXiv:2211.00792
-
MAST: Multiscale Audio Spectrogram Transformers 2 Nov 2022 · 1 repository · arXiv:2211.01515
-
MPCFormer: fast, performant and private Transformer inference with MPC 2 Nov 2022 · 1 repository · arXiv:2211.01452
-
Multi-level Distillation of Semantic Knowledge for Pre-training Multilingual Language Model 2 Nov 2022 · 0 repositories · arXiv:2211.01200
-
Pop2Piano : Pop Audio-based Piano Cover Generation 2 Nov 2022 · 4 repositories · arXiv:2211.00895Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
RegCLR: A Self-Supervised Framework for Tabular Representation Learning in the Wild 2 Nov 2022 · 0 repositories · arXiv:2211.01165
-
Transformer-based encoder-encoder architecture for Spoken Term Detection 2 Nov 2022 · 0 repositories · arXiv:2211.01089
-
Wind Power Forecasting Considering Data Privacy Protection: A Federated Deep Reinforcement Learning Approach 2 Nov 2022 · 0 repositories · arXiv:2211.02674
-
WITT: A Wireless Image Transmission Transformer for Semantic Communications 2 Nov 2022 · 2 repositories · arXiv:2211.00937
-
FRSUM: Towards Faithful Abstractive Summarization via Enhancing Factual Robustness 1 Nov 2022 · 0 repositories · arXiv:2211.00294
-
Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small 1 Nov 2022 · 7 repositories · arXiv:2211.00593Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Investigating Content-Aware Neural Text-To-Speech MOS Prediction Using Prosodic and Linguistic Features 1 Nov 2022 · 0 repositories · arXiv:2211.00342
-
An Empirical Study on Data Leakage and Generalizability of Link Prediction Models for Issues and Commits 1 Nov 2022 · 0 repositories · arXiv:2211.00381
-
Reduce, Reuse, Recycle: Improving Training Efficiency with Distillation 1 Nov 2022 · 0 repositories · arXiv:2211.00683
-
Text-Only Training for Image Captioning using Noise-Injected CLIP 1 Nov 2022 · 4 repositories · arXiv:2211.00575Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
ViT-DeiT: An Ensemble Model for Breast Cancer Histopathological Images Classification 1 Nov 2022 · 0 repositories · arXiv:2211.00749
-
AdaMix: Mixture-of-Adaptations for Parameter-efficient Model Tuning 31 Oct 2022 · 1 repository · arXiv:2210.17451Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Efficient Document Retrieval by End-to-End Refining and Quantizing BERT Embedding with Contrastive Product Quantization 31 Oct 2022 · 1 repository · arXiv:2210.17170
-
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers 31 Oct 2022 · 17 repositories · arXiv:2210.17323Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Joint Audio/Text Training for Transformer Rescorer of Streaming Speech Recognition 31 Oct 2022 · 0 repositories · arXiv:2211.00174
-
Leveraging Pre-trained Models for Failure Analysis Triplets Generation 31 Oct 2022 · 0 repositories · arXiv:2210.17497
-
Multi-Camera Calibration Free BEV Representation for 3D Object Detection 31 Oct 2022 · 0 repositories · arXiv:2210.17252
-
Probabilistic Decomposition Transformer for Time Series Forecasting 31 Oct 2022 · 1 repository · arXiv:2210.17393
-
QNet: A Quantum-native Sequence Encoder Architecture 31 Oct 2022 · 0 repositories · arXiv:2210.17262
-
QuaLA-MiniLM: a Quantized Length Adaptive MiniLM 31 Oct 2022 · 2 repositories · arXiv:2210.17114
-
SDCL: Self-Distillation Contrastive Learning for Chinese Spell Checking 31 Oct 2022 · 0 repositories · arXiv:2210.17168
-
Spatial-Temporal Synchronous Graph Transformer network (STSGT) for COVID-19 forecasting 31 Oct 2022 · 1 repository · arXiv:2211.00082
-
SSD-LM: Semi-autoregressive Simplex-based Diffusion Language Model for Text Generation and Modular Control 31 Oct 2022 · 2 repositories · arXiv:2210.17432
-
Structured State Space Decoder for Speech Recognition and Synthesis 31 Oct 2022 · 0 repositories · arXiv:2210.17098
-
Teacher-student curriculum learning for reinforcement learning 31 Oct 2022 · 0 repositories · arXiv:2210.17368
-
Towards Zero-Shot and Few-Shot Table Question Answering using GPT-3 31 Oct 2022 · 0 repositories · arXiv:2210.17284
-
ViT-LSLA: Vision Transformer with Light Self-Limited-Attention 31 Oct 2022 · 0 repositories · arXiv:2210.17115
-
An Efficient Memory-Augmented Transformer for Knowledge-Intensive NLP Tasks 30 Oct 2022 · 1 repository · arXiv:2210.16773
-
Attention Swin U-Net: Cross-Contextual Attention Mechanism for Skin Lesion Segmentation 30 Oct 2022 · 1 repository · arXiv:2210.16898
-
Foreign Object Debris Detection for Airport Pavement Images based on Self-supervised Localization and Vision Transformer 30 Oct 2022 · 1 repository · arXiv:2210.16901
-
Learning to Decompose: Hypothetical Question Decomposition Based on Comparable Texts 30 Oct 2022 · 0 repositories · arXiv:2210.16865
-
Parameter-Efficient Tuning Makes a Good Classification Head 30 Oct 2022 · 1 repository · arXiv:2210.16771
-
QuEst: Graph Transformer for Quantum Circuit Reliability Estimation 30 Oct 2022 · 1 repository · arXiv:2210.16724
-
Time-rEversed diffusioN tEnsor Transformer: A new TENET of Few-Shot Object Detection 30 Oct 2022 · 0 repositories · arXiv:2210.16897
-
token2vec: A Joint Self-Supervised Pre-training Framework Using Unpaired Speech and Text 30 Oct 2022 · 0 repositories · arXiv:2210.16755
-
ViTASD: Robust Vision Transformer Baselines for Autism Spectrum Disorder Facial Diagnosis 30 Oct 2022 · 1 repository · arXiv:2210.16943
-
BERT Meets CTC: New Formulation of End-to-End Speech Recognition with Pre-trained Masked Language Model 29 Oct 2022 · 0 repositories · arXiv:2210.16663
-
Interpretable CNN-Multilevel Attention Transformer for Rapid Recognition of Pneumonia from Chest X-Ray Images 29 Oct 2022 · 0 repositories · arXiv:2210.16584
-
Empirical Evaluation of Post-Training Quantization Methods for Language Tasks 29 Oct 2022 · 0 repositories · arXiv:2210.16621
-
Exploiting prompt learning with pre-trained language models for Alzheimer's Disease detection 29 Oct 2022 · 1 repository · arXiv:2210.16539
-
Pair DETR: Contrastive Learning Speeds Up DETR Training 29 Oct 2022 · 0 repositories · arXiv:2210.16476
-
Recursive Reasoning in Minimax Games: A Level k Gradient Play Method 29 Oct 2022 · 1 repository · arXiv:2210.16482
-
A Long-term Dependent and Trustworthy Approach to Reactor Accident Prognosis based on Temporal Fusion Transformer 28 Oct 2022 · 0 repositories · arXiv:2210.17298
-
BEBERT: Efficient and Robust Binary Ensemble BERT 28 Oct 2022 · 1 repository · arXiv:2210.15976
-
Contextual Learning in Fourier Complex Field for VHR Remote Sensing Images 28 Oct 2022 · 3 repositories · arXiv:2210.15972
-
Dimensionality Reduced Antenna Array for Beamforming/steering 28 Oct 2022 · 0 repositories · arXiv:2210.16197
-
Efficient Speech Translation with Dynamic Latent Perceivers 28 Oct 2022 · 1 repository · arXiv:2210.16264
-
Exploring Spatial-Temporal Features for Deepfake Detection and Localization 28 Oct 2022 · 1 repository · arXiv:2210.15872
-
Feature Engineering vs BERT on Twitter Data 28 Oct 2022 · 0 repositories · arXiv:2210.16168
-
Grafting Vision Transformers 28 Oct 2022 · 0 repositories · arXiv:2210.15943