Methods › General › Stochastic Optimization › Adam › Papers, page 159
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 159 of 244: papers 15,801 to 15,900 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SimLM: Pre-training with Representation Bottleneck for Dense Passage Retrieval 6 Jul 2022 · 1 repository · arXiv:2207.02578
-
The Role of Complex NLP in Transformers for Text Ranking? 6 Jul 2022 · 0 repositories · arXiv:2207.02522
-
Transformers are Adaptable Task Planners 6 Jul 2022 · 0 repositories · arXiv:2207.02442
-
Betti numbers of attention graphs is all you really need 5 Jul 2022 · 1 repository · arXiv:2207.01903
-
CNN-based Local Vision Transformer for COVID-19 Diagnosis 5 Jul 2022 · 0 repositories · arXiv:2207.02027
-
CoBEVT: Cooperative Bird's Eye View Semantic Segmentation with Sparse Transformers 5 Jul 2022 · 2 repositories · arXiv:2207.02202Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Deep Learning Reveals Patterns of Diverse and Changing Sentiments Towards COVID-19 Vaccines Based on 11 Million Tweets 5 Jul 2022 · 0 repositories · arXiv:2207.10641
-
Detecting and Recovering Sequential DeepFake Manipulation 5 Jul 2022 · 1 repository · arXiv:2207.02204Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
FishFormer: Annulus Slicing-based Transformer for Fisheye Rectification with Efficacy Domain Exploration 5 Jul 2022 · 0 repositories · arXiv:2207.01925
-
Improving Semantic Segmentation in Transformers using Hierarchical Inter-Level Attention 5 Jul 2022 · 0 repositories · arXiv:2207.02126
-
Machine Learning Model Sizes and the Parameter Gap 5 Jul 2022 · 0 repositories · arXiv:2207.02852
-
TabPFN: A Transformer That Solves Small Tabular Classification Problems in a Second 5 Jul 2022 · 7 repositories · arXiv:2207.01848Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Multimodal Frame-Scoring Transformer for Video Summarization 5 Jul 2022 · 0 repositories · arXiv:2207.01814
-
Softmax-free Linear Transformers 5 Jul 2022 · 1 repository · arXiv:2207.03341
-
Swin Deformable Attention U-Net Transformer (SDAUT) for Explainable Fast MRI 5 Jul 2022 · 1 repository · arXiv:2207.02390
-
Transformer based Models for Unsupervised Anomaly Segmentation in Brain MR Images 5 Jul 2022 · 0 repositories · arXiv:2207.02059
-
Ultra-Low-Bitrate Speech Coding with Pretrained Transformers 5 Jul 2022 · 0 repositories · arXiv:2207.02262
-
An adaptive music generation architecture for games based on the deep learning Transformer mode 4 Jul 2022 · 0 repositories · arXiv:2207.01698
-
BERT, can HE predict contrastive focus? Predicting and controlling prominence in neural TTS using a language model 4 Jul 2022 · 0 repositories · arXiv:2207.01718
-
Interaction Transformer for Human Reaction Generation 4 Jul 2022 · 1 repository · arXiv:2207.01685
-
Large-scale Robustness Analysis of Video Action Recognition Models 4 Jul 2022 · 1 repository · arXiv:2207.01398
-
One Model is Not Enough: Ensembles for Isolated Sign Language Recognition 4 Jul 2022 · 0 repositories
-
Towards Real-World Video Denosing: A Practical Video Denosing Dataset and Network 4 Jul 2022 · 0 repositories · arXiv:2207.01356
-
TANet: Transformer-based Asymmetric Network for RGB-D Salient Object Detection 4 Jul 2022 · 1 repository · arXiv:2207.01172
-
Understanding Performance of Long-Document Ranking Models through Comprehensive Evaluation and Leaderboarding 4 Jul 2022 · 3 repositories · arXiv:2207.01262
-
Using contextual sentence analysis models to recognize ESG concepts 4 Jul 2022 · 0 repositories · arXiv:2207.01402
-
Divert More Attention to Vision-Language Tracking 3 Jul 2022 · 1 repository · arXiv:2207.01076
-
GUIM -- General User and Item Embedding with Mixture of Representation in E-commerce 2 Jul 2022 · 0 repositories · arXiv:2207.00750
-
Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism 2 Jul 2022 · 0 repositories · arXiv:2207.00883
-
A Polyphone BERT for Polyphone Disambiguation in Mandarin Chinese 1 Jul 2022 · 0 repositories · arXiv:2207.12089
-
A Temporal Fusion Transformer for Long-term Explainable Prediction of Emergency Department Overcrowding 1 Jul 2022 · 0 repositories · arXiv:2207.00610
-
DALG: Deep Attentive Local and Global Modeling for Image Retrieval 1 Jul 2022 · 0 repositories · arXiv:2207.00287
-
Improving Low-Resource Speech Recognition with Pretrained Speech Models: Continued Pretraining vs. Semi-Supervised Training 1 Jul 2022 · 0 repositories · arXiv:2207.00659
-
Masked Autoencoder for Self-Supervised Pre-training on Lidar Point Clouds 1 Jul 2022 · 1 repository · arXiv:2207.00531Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Polarized Color Image Denoising using Pocoformer 1 Jul 2022 · 0 repositories · arXiv:2207.00215
-
Time-aware Dynamic Graph Embedding for Asynchronous Structural Evolution 1 Jul 2022 · 0 repositories · arXiv:2207.00594
-
TopicFM: Robust and Interpretable Topic-Assisted Feature Matching 1 Jul 2022 · 1 repository · arXiv:2207.00328
-
AnoShift: A Distribution Shift Benchmark for Unsupervised Anomaly Detection 30 Jun 2022 · 1 repository · arXiv:2206.15476Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Classical and learned MR to pseudo-CT mappings for accurate transcranial ultrasound simulation 30 Jun 2022 · 0 repositories · arXiv:2206.15441
-
Compressing Pre-trained Transformers via Low-Bit NxM Sparsity for Natural Language Understanding 30 Jun 2022 · 0 repositories · arXiv:2206.15014
-
CTrGAN: Cycle Transformers GAN for Gait Transfer 30 Jun 2022 · 0 repositories · arXiv:2206.15248
-
Deep Reinforcement Learning with Swin Transformers 30 Jun 2022 · 1 repository · arXiv:2206.15269
-
DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale 30 Jun 2022 · 2 repositories · arXiv:2207.00032
-
FL-Tuning: Layer Tuning for Feed-Forward Network in Transformer 30 Jun 2022 · 1 repository · arXiv:2206.15312
-
GaitForeMer: Self-Supervised Pre-Training of Transformers via Human Motion Forecasting for Few-Shot Gait Impairment Severity Estimation 30 Jun 2022 · 1 repository · arXiv:2207.00106
-
ListBERT: Learning to Rank E-commerce products with Listwise BERT 30 Jun 2022 · 0 repositories · arXiv:2206.15198
-
PolarFormer: Multi-camera 3D Object Detection with Polar Transformer 30 Jun 2022 · 1 repository · arXiv:2206.15398Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
PVT-COV19D: Pyramid Vision Transformer for COVID-19 Diagnosis 30 Jun 2022 · 0 repositories · arXiv:2206.15069
-
Rethinking Surgical Captioning: End-to-End Window-Based MLP Transformer Using Patches 30 Jun 2022 · 1 repository · arXiv:2207.00113
-
TENET: Transformer Encoding Network for Effective Temporal Flow on Motion Prediction 30 Jun 2022 · 0 repositories · arXiv:2207.00170
-
The Topological BERT: Transforming Attention into Topology for Natural Language Processing 30 Jun 2022 · 0 repositories · arXiv:2206.15195
-
Two-Stage Classifier for COVID-19 Misinformation Detection Using BERT: a Study on Indonesian Tweets 30 Jun 2022 · 2 repositories · arXiv:2206.15359
-
BATFormer: Towards Boundary-Aware Lightweight Transformer for Efficient Medical Image Segmentation 29 Jun 2022 · 2 repositories · arXiv:2206.14409
-
Chinese Word Sense Embedding with SememeWSD and Synonym Set 29 Jun 2022 · 2 repositories · arXiv:2206.14388
-
Deformable Graph Transformer 29 Jun 2022 · 0 repositories · arXiv:2206.14337
-
Momentum Diminishes the Effect of Spectral Bias in Physics-Informed Neural Networks 29 Jun 2022 · 0 repositories · arXiv:2206.14862
-
Multi-Channel Vision Transformer for Epileptic Seizure Prediction 29 Jun 2022 · 0 repositories
-
On the Prediction Network Architecture in RNN-T for ASR 29 Jun 2022 · 0 repositories · arXiv:2206.14618
-
Simple and Effective Multi-sentence TTS with Expressive and Coherent Prosody 29 Jun 2022 · 0 repositories · arXiv:2206.14643
-
Solving Quantitative Reasoning Problems with Language Models 29 Jun 2022 · 1 repository · arXiv:2206.14858
-
Summarizing Videos using Concentrated Attention and Considering the Uniqueness and Diversity of the Video Frames 29 Jun 2022 · 1 repository
-
Two-Stage COVID19 Classification Using BERT Features 29 Jun 2022 · 0 repositories · arXiv:2206.14861
-
Cross-Forgery Analysis of Vision Transformers and CNNs for Deepfake Image Detection 28 Jun 2022 · 2 repositories · arXiv:2206.13829
-
Exploring linguistic feature and model combination for speech recognition based automatic AD detection 28 Jun 2022 · 0 repositories · arXiv:2206.13758
-
Robustifying Vision Transformer without Retraining from Scratch by Test-Time Class-Conditional Feature Alignment 28 Jun 2022 · 1 repository · arXiv:2206.13951Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ZoDIAC: Zoneout Dropout Injection Attention Calculation 28 Jun 2022 · 1 repository · arXiv:2206.14263
-
Context-Aware Transformers For Spinal Cancer Detection and Radiological Grading 27 Jun 2022 · 0 repositories · arXiv:2206.13173
-
Kernel Attention Transformer (KAT) for Histopathology Whole Slide Image Classification 27 Jun 2022 · 1 repository · arXiv:2206.13156
-
Materials Transformers Language Models for Generative Materials Design: a benchmark study 27 Jun 2022 · 1 repository · arXiv:2206.13578
-
Prompting Decision Transformer for Few-Shot Policy Generalization 27 Jun 2022 · 0 repositories · arXiv:2206.13499
-
Theoretical analysis of Adam using hyperparameters close to one without Lipschitz smoothness 27 Jun 2022 · 0 repositories · arXiv:2206.13290
-
Multiple Instance Learning with Mixed Supervision in Gleason Grading 26 Jun 2022 · 1 repository · arXiv:2206.12798
-
On Comparison of Encoders for Attention based End to End Speech Recognition in Standalone and Rescoring Mode 26 Jun 2022 · 0 repositories · arXiv:2206.12829
-
Vision Transformer for Contrastive Clustering 26 Jun 2022 · 1 repository · arXiv:2206.12925
-
Adversarial Self-Attention for Language Understanding 25 Jun 2022 · 1 repository · arXiv:2206.12608
-
Construct a Sentence with Multiple Specified Words 25 Jun 2022 · 0 repositories · arXiv:2206.12565
-
SC-Transformer++: Structured Context Transformer for Generic Event Boundary Detection 25 Jun 2022 · 1 repository · arXiv:2206.12634
-
Self-supervised Context-aware Style Representation for Expressive Speech Synthesis 25 Jun 2022 · 0 repositories · arXiv:2206.12559
-
A multi-model-based deep learning framework for short text multiclass classification with the imbalanced and extremely small data set 24 Jun 2022 · 0 repositories · arXiv:2206.12027
-
A Test for Evaluating Performance in Human-Computer Systems 24 Jun 2022 · 0 repositories · arXiv:2206.12390
-
Capture Salient Historical Information: A Fast and Accurate Non-Autoregressive Model for Multi-turn Spoken Language Understanding 24 Jun 2022 · 0 repositories · arXiv:2206.12209
-
Confidence Score Based Conformer Speaker Adaptation for Speech Recognition 24 Jun 2022 · 0 repositories · arXiv:2206.12045
-
MVP: Multi-task Supervised Pre-training for Natural Language Generation 24 Jun 2022 · 4 repositories · arXiv:2206.12131
-
Text and author-level political inference using heterogeneous knowledge representations 24 Jun 2022 · 0 repositories · arXiv:2206.12293
-
The Second Place Solution for The 4th Large-scale Video Object Segmentation Challenge--Track 3: Referring Video Object Segmentation 24 Jun 2022 · 0 repositories · arXiv:2206.12035
-
Unified BERT for Few-shot Natural Language Understanding 24 Jun 2022 · 0 repositories · arXiv:2206.12094
-
Using BERT Embeddings to Model Word Importance in Conversational Transcripts for Deaf and Hard of Hearing Users 24 Jun 2022 · 0 repositories · arXiv:2206.12368
-
A Disability Lens towards Biases in GPT-3 Generated Open-Ended Languages 23 Jun 2022 · 0 repositories · arXiv:2206.11993
-
BERT Rankers are Brittle: a Study using Adversarial Document Perturbations 23 Jun 2022 · 1 repository · arXiv:2206.11724
-
Equiformer: Equivariant Graph Attention Transformer for 3D Atomistic Graphs 23 Jun 2022 · 4 repositories · arXiv:2206.11990Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
ICOS Protein Expression Segmentation: Can Transformer Networks Give Better Results? 23 Jun 2022 · 1 repository · arXiv:2206.11520
-
Learning Viewpoint-Agnostic Visual Representations by Recovering Tokens in 3D Space 23 Jun 2022 · 1 repository · arXiv:2206.11895Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Revisiting Orthogonality Regularization: A Study for Convolutional Neural Networks in Image Classification 23 Jun 2022 · 1 repository
-
Set Norm and Equivariant Skip Connections: Putting the Deep in Deep Sets 23 Jun 2022 · 1 repository · arXiv:2206.11925Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Toward Clinically Assisted Colorectal Polyp Recognition via Structured Cross-modal Representation Consistency 23 Jun 2022 · 1 repository · arXiv:2206.11826
-
Towards WinoQueer: Developing a Benchmark for Anti-Queer Bias in Large Language Models 23 Jun 2022 · 0 repositories · arXiv:2206.11484
-
Answer Fast: Accelerating BERT on the Tensor Streaming Processor 22 Jun 2022 · 0 repositories · arXiv:2206.11062
-
Behavior Transformers: Cloning k modes with one stone 22 Jun 2022 · 2 repositories · arXiv:2206.11251Syntology official (archive's flag): 5 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
Efficient and effective training of language and graph neural network models 22 Jun 2022 · 0 repositories · arXiv:2206.10781
-
Feature Re-calibration based Multiple Instance Learning for Whole Slide Image Classification 22 Jun 2022 · 1 repository · arXiv:2206.10878