Methods › General › Stochastic Optimization › Adam › Papers, page 183
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 183 of 244: papers 18,201 to 18,300 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Offensive Language Detection with BERT-based models, By Customizing Attention Probabilities 11 Oct 2021 · 0 repositories · arXiv:2110.05133
-
SRU++: Pioneering Fast Recurrence with Attention for Speech Recognition 11 Oct 2021 · 0 repositories · arXiv:2110.05571
-
Topic Modeling, Clade-assisted Sentiment Analysis, and Vaccine Brand Reputation Analysis of COVID-19 Vaccine-related Facebook Comments in the Philippines 11 Oct 2021 · 1 repository · arXiv:2111.04416
-
Unsupervised Source Separation via Bayesian Inference in the Latent Domain 11 Oct 2021 · 1 repository · arXiv:2110.05313
-
DCT: Dynamic Compressive Transformer for Modeling Unbounded Sequence 10 Oct 2021 · 0 repositories · arXiv:2110.04821
-
Multi-Channel End-to-End Neural Diarization with Distributed Microphones 10 Oct 2021 · 0 repositories · arXiv:2110.04694
-
Global Vision Transformer Pruning with Hessian-Aware Saliency 10 Oct 2021 · 1 repository · arXiv:2110.04869Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Automatic Text Extractive Summarization Based on Graph and Pre-trained Language Model Attention 10 Oct 2021 · 0 repositories · arXiv:2110.04878
-
PASTE: A Tagging-Free Decoding Framework Using Pointer Networks for Aspect Sentiment Triplet Extraction 10 Oct 2021 · 1 repository · arXiv:2110.04794
-
SP-GPT2: Semantics Improvement in Vietnamese Poetry Generation 10 Oct 2021 · 2 repositories · arXiv:2110.15723
-
SuperShaper: Task-Agnostic Super Pre-training of BERT Models with Variable Hidden Dimensions 10 Oct 2021 · 0 repositories · arXiv:2110.04711
-
Yuan 1.0: Large-Scale Pre-trained Language Model in Zero-Shot and Few-Shot Learning 10 Oct 2021 · 1 repository · arXiv:2110.04725
-
An Isotropy Analysis in the Multilingual BERT Embedding Space 9 Oct 2021 · 1 repository · arXiv:2110.04504
-
Gated recurrent units and temporal convolutional network for multilabel classification 9 Oct 2021 · 0 repositories · arXiv:2110.04414
-
Leveraging recent advances in Pre-Trained Language Models forEye-Tracking Prediction 9 Oct 2021 · 1 repository · arXiv:2110.04475
-
Vector-quantized Image Modeling with Improved VQGAN 9 Oct 2021 · 5 repositories · arXiv:2110.04627Syntology 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
The Layout Generation Algorithm of Graphic Design Based on Transformer-CVAE 8 Oct 2021 · 0 repositories · arXiv:2110.06794
-
Adversarial Token Attacks on Vision Transformers 8 Oct 2021 · 0 repositories · arXiv:2110.04337
-
ALL-IN-ONE: Multi-Task Learning BERT models for Evaluating Peer Assessments 8 Oct 2021 · 0 repositories · arXiv:2110.03895
-
Context-LGM: Leveraging Object-Context Relation for Context-Aware Object Recognition 8 Oct 2021 · 0 repositories · arXiv:2110.04042
-
Towards Learning (Dis)-Similarity of Source Code from Program Contrasts 8 Oct 2021 · 0 repositories · arXiv:2110.03868
-
Development of an Extractive Title Generation System Using Titles of Papers of Top Conferences for Intermediate English Students 8 Oct 2021 · 0 repositories · arXiv:2110.04204
-
HydraSum: Disentangling Stylistic Features in Text Summarization using Multi-Decoder Models 8 Oct 2021 · 1 repository · arXiv:2110.04400Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples)
-
Knowledge-Enhanced Hierarchical Graph Transformer Network for Multi-Behavior Recommendation 8 Oct 2021 · 1 repository · arXiv:2110.04000
-
Learning Adaptive Control Flow in Transformers for Improved Systematic Generalization 8 Oct 2021 · 0 repositories
-
Local and Global Context-Based Pairwise Models for Sentence Ordering 8 Oct 2021 · 1 repository · arXiv:2110.04291
-
M6-10T: A Sharing-Delinking Paradigm for Efficient Multi-Trillion Parameter Pretraining 8 Oct 2021 · 0 repositories · arXiv:2110.03888
-
Does Momentum Change the Implicit Regularization on Separable Data? 8 Oct 2021 · 0 repositories · arXiv:2110.03891
-
Multiplex Behavioral Relation Learning for Recommendation via Memory Augmented Transformer Network 8 Oct 2021 · 1 repository · arXiv:2110.04002
-
RPT: Toward Transferable Model on Heterogeneous Researcher Data via Pre-Training 8 Oct 2021 · 1 repository · arXiv:2110.07336
-
Speeding up Deep Model Training by Sharing Weights and Then Unsharing 8 Oct 2021 · 0 repositories · arXiv:2110.03848
-
Taming Sparsely Activated Transformer with Stochastic Experts 8 Oct 2021 · 1 repository · arXiv:2110.04260Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Text analysis and deep learning: A network approach 8 Oct 2021 · 0 repositories · arXiv:2110.04151
-
ViDT: An Efficient and Effective Fully Transformer-based Object Detector 8 Oct 2021 · 1 repository · arXiv:2110.03921
-
A Comparative Study of Transformer-Based Language Models on Extractive Question Answering 7 Oct 2021 · 0 repositories · arXiv:2110.03142
-
Attention is All You Need? Good Embeddings with Statistics are enough:Large Scale Audio Understanding without Transformers/ Convolutions/ BERTs/ Mixers/ Attention/ RNNs or .... 7 Oct 2021 · 0 repositories · arXiv:2110.03183
-
Cross-Language Learning for Entity Matching 7 Oct 2021 · 1 repository · arXiv:2110.03338
-
End-to-End Supermask Pruning: Learning to Prune Image Captioning Models 7 Oct 2021 · 1 repository · arXiv:2110.03298Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Generative Pre-Trained Transformer for Cardiac Abnormality Detection 7 Oct 2021 · 0 repositories · arXiv:2110.04071
-
Investigating Growth at Risk Using a Multi-country Non-parametric Quantile Factor Model 7 Oct 2021 · 0 repositories · arXiv:2110.03411
-
Layer-wise Pruning of Transformer Attention Heads for Efficient Language Modeling 7 Oct 2021 · 1 repository · arXiv:2110.03252
-
Minimum word error training for non-autoregressive Transformer-based code-switching ASR 7 Oct 2021 · 0 repositories · arXiv:2110.03573
-
Universality of Winning Tickets: A Renormalization Group Perspective 7 Oct 2021 · 0 repositories · arXiv:2110.03210
-
8-bit Optimizers via Block-wise Quantization 6 Oct 2021 · 3 repositories · arXiv:2110.02861
-
Adversarial Robustness Comparison of Vision Transformer and MLP-Mixer to CNNs 6 Oct 2021 · 1 repository · arXiv:2110.02797
-
Anomaly Transformer: Time Series Anomaly Detection with Association Discrepancy 6 Oct 2021 · 3 repositories · arXiv:2110.02642Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Dynamically Decoding Source Domain Knowledge for Domain Generalization 6 Oct 2021 · 0 repositories · arXiv:2110.03027
-
Geometric Transformers for Protein Interface Contact Prediction 6 Oct 2021 · 2 repositories · arXiv:2110.02423
-
How BPE Affects Memorization in Transformers 6 Oct 2021 · 0 repositories · arXiv:2110.02782
-
Learning to Iteratively Solve Routing Problems with Dual-Aspect Collaborative Transformer 6 Oct 2021 · 2 repositories · arXiv:2110.02544Syntology official (archive's flag): 7 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples)
-
NUS-IDS at FinCausal 2021: Dependency Tree in Graph Neural Network for Better Cause-Effect Span Detection 6 Oct 2021 · 1 repository · arXiv:2110.02991
-
On Neurons Invariant to Sentence Structural Changes in Neural Machine Translation 6 Oct 2021 · 1 repository · arXiv:2110.03067
-
PoNet: Pooling Network for Efficient Token Mixing in Long Sequences 6 Oct 2021 · 1 repository · arXiv:2110.02442Syntology official: no sample here; runs from other or unrecorded repositories · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Pretrained Transformers for Offensive Language Identification in Tanglish 6 Oct 2021 · 1 repository · arXiv:2110.02852
-
Semantic Prediction: Which One Should Come First, Recognition or Prediction? 6 Oct 2021 · 1 repository · arXiv:2110.02829
-
Analyzing the Impact of COVID-19 on Economy from the Perspective of Users Reviews 5 Oct 2021 · 0 repositories · arXiv:2110.02198
-
ASR Rescoring and Confidence Estimation with ELECTRA 5 Oct 2021 · 0 repositories · arXiv:2110.01857
-
BERT Attends the Conversation: Improving Low-Resource Conversational ASR 5 Oct 2021 · 1 repository · arXiv:2110.02267
-
DistilHuBERT: Speech Representation Learning by Layer-wise Distillation of Hidden-unit BERT 5 Oct 2021 · 1 repository · arXiv:2110.01900
-
Exploiting Twitter as Source of Large Corpora of Weakly Similar Pairs for Semantic Sentence Embeddings 5 Oct 2021 · 1 repository · arXiv:2110.02030
-
FoodChem: A food-chemical relation extraction model 5 Oct 2021 · 1 repository · arXiv:2110.02019
-
KKT Conditions, First-Order and Second-Order Optimization, and Distributed Optimization: Tutorial and Survey 5 Oct 2021 · 0 repositories · arXiv:2110.01858
-
Learning Sense-Specific Static Embeddings using Contextualised Word Embeddings as a Proxy 5 Oct 2021 · 0 repositories · arXiv:2110.02204
-
Leveraging the Inductive Bias of Large Language Models for Abstract Textual Reasoning 5 Oct 2021 · 0 repositories · arXiv:2110.02370
-
Sicilian Translator: A Recipe for Low-Resource NMT 5 Oct 2021 · 1 repository · arXiv:2110.01938
-
Sound Event Detection Transformer: An Event-based End-to-End Model for Sound Event Detection 5 Oct 2021 · 1 repository · arXiv:2110.02011
-
Top-N: Equivariant set and graph generation without exchangeability 5 Oct 2021 · 1 repository · arXiv:2110.02096Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
ur-iw-hnt at GermEval 2021: An Ensembling Strategy with Multiple BERT Models 5 Oct 2021 · 0 repositories · arXiv:2110.02042
-
Word Acquisition in Neural Language Models 5 Oct 2021 · 1 repository · arXiv:2110.02406
-
Molformer: Motif-based Transformer on 3D Heterogeneous Molecular Graphs 4 Oct 2021 · 2 repositories · arXiv:2110.01191Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
A free lunch from ViT:Adaptive Attention Multi-scale Fusion Transformer for Fine-grained Visual Recognition 4 Oct 2021 · 0 repositories · arXiv:2110.01240
-
Effectiveness of Optimization Algorithms in Deep Image Classification 4 Oct 2021 · 1 repository · arXiv:2110.01598
-
Exploiting Pre-Trained ASR Models for Alzheimer's Disease Recognition Through Spontaneous Speech 4 Oct 2021 · 0 repositories · arXiv:2110.01493
-
JuriBERT: A Masked-Language Model Adaptation for French Legal Text 4 Oct 2021 · 1 repository · arXiv:2110.01485
-
VTAMIQ: Transformers for Attention Modulated Image Quality Assessment 4 Oct 2021 · 1 repository · arXiv:2110.01655
-
Adversarial Examples Generation for Reducing Implicit Gender Bias in Pre-trained Models 3 Oct 2021 · 0 repositories · arXiv:2110.01094
-
Music Playlist Title Generation: A Machine-Translation Approach 3 Oct 2021 · 0 repositories · arXiv:2110.07354
-
Parallel Actors and Learners: A Framework for Generating Scalable RL Implementations 3 Oct 2021 · 0 repositories · arXiv:2110.01101
-
Unsupervised paradigm for information extraction from transcripts using BERT 3 Oct 2021 · 0 repositories · arXiv:2110.00949
-
Artificial intelligence for Sustainable Energy: A Contextual Topic Modeling and Content Analysis 2 Oct 2021 · 0 repositories · arXiv:2110.00828
-
Implicit and Explicit Attention for Zero-Shot Learning 2 Oct 2021 · 1 repository · arXiv:2110.00860
-
ProTo: Program-Guided Transformer for Program-Guided Tasks 2 Oct 2021 · 1 repository · arXiv:2110.00804Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Swiss-Judgment-Prediction: A Multilingual Legal Judgment Prediction Benchmark 2 Oct 2021 · 1 repository · arXiv:2110.00806
-
BERT4GCN: Using BERT Intermediate Layers to Augment GCN for Aspect-based Sentiment Classification 1 Oct 2021 · 0 repositories · arXiv:2110.00171
-
3rd Place Scheme on Instance Segmentation Track of ICCV 2021 VIPriors Challenges 1 Oct 2021 · 0 repositories · arXiv:2110.00242
-
Layer-wise and Dimension-wise Locally Adaptive Federated Learning 1 Oct 2021 · 0 repositories · arXiv:2110.00532
-
Geometry Attention Transformer with Position-aware LSTMs for Image Captioning 1 Oct 2021 · 1 repository · arXiv:2110.00335
-
Improving Punctuation Restoration for Speech Transcripts via External Data 1 Oct 2021 · 0 repositories · arXiv:2110.00560
-
Low Frequency Names Exhibit Bias and Overfitting in Contextualizing Language Models 1 Oct 2021 · 0 repositories · arXiv:2110.00672
-
Span Labeling Approach for Vietnamese and Chinese Word Segmentation 1 Oct 2021 · 0 repositories · arXiv:2110.00156
-
Unpacking the Interdependent Systems of Discrimination: Ableist Bias in NLP Systems through an Intersectional Lens 1 Oct 2021 · 0 repositories · arXiv:2110.00521
-
BERT got a Date: Introducing Transformers to Temporal Tagging 30 Sep 2021 · 1 repository · arXiv:2109.14927
-
Bitcoin Transaction Strategy Construction Based on Deep Reinforcement Learning 30 Sep 2021 · 0 repositories · arXiv:2109.14789
-
COVID-19 Fake News Detection Using Bidirectional Encoder Representations from Transformers Based Models 30 Sep 2021 · 1 repository · arXiv:2109.14816
-
GT U-Net: A U-Net Like Group Transformer Network for Tooth Root Segmentation 30 Sep 2021 · 1 repository · arXiv:2109.14813
-
Inducing Transformer's Compositional Generalization Ability via Auxiliary Sequence Prediction Tasks 30 Sep 2021 · 1 repository · arXiv:2109.15256
-
Learning to Predict Trustworthiness with Steep Slope Loss 30 Sep 2021 · 1 repository · arXiv:2110.00054Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
MobTCast: Leveraging Auxiliary Trajectory Forecasting for Human Mobility Prediction 30 Sep 2021 · 0 repositories · arXiv:2110.01401
-
Prose2Poem: The Blessing of Transformers in Translating Prose to Persian Poetry 30 Sep 2021 · 1 repository · arXiv:2109.14934
-
Redesigning the Transformer Architecture with Insights from Multi-particle Dynamical Systems 30 Sep 2021 · 1 repository · arXiv:2109.15142