Methods › General › Stochastic Optimization › Adam › Papers, page 205
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 205 of 244: papers 20,401 to 20,500 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Automatic punctuation restoration with BERT models 18 Jan 2021 · 1 repository · arXiv:2101.07343
-
Can a Fruit Fly Learn Word Embeddings? 18 Jan 2021 · 2 repositories · arXiv:2101.06887Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Inference for BART with Multinomial Outcomes 18 Jan 2021 · 1 repository · arXiv:2101.06823
-
Dual-Level Collaborative Transformer for Image Captioning 16 Jan 2021 · 1 repository · arXiv:2101.06462
-
Match-Ignition: Plugging PageRank into Transformer for Long-form Text Matching 16 Jan 2021 · 1 repository · arXiv:2101.06423
-
Transformer-Based Models for Question Answering on COVID19 16 Jan 2021 · 0 repositories · arXiv:2101.11432
-
Grid Search Hyperparameter Benchmarking of BERT, ALBERT, and LongFormer on DuoRC 15 Jan 2021 · 0 repositories · arXiv:2101.06326
-
Hostility Detection and Covid-19 Fake News Detection in Social Media 15 Jan 2021 · 0 repositories · arXiv:2101.05953
-
KDLSQ-BERT: A Quantized Bert Combining Knowledge Distillation with Learned Step Size Quantization 15 Jan 2021 · 0 repositories · arXiv:2101.05938
-
ECOL: Early Detection of COVID Lies Using Content, Prior Knowledge and Source Information 14 Jan 2021 · 1 repository · arXiv:2101.05499
-
Exploration of Visual Features and their weighted-additive fusion for Video Captioning 14 Jan 2021 · 0 repositories · arXiv:2101.05806
-
GAN Inversion: A Survey 14 Jan 2021 · 1 repository · arXiv:2101.05278
-
Persistent Anti-Muslim Bias in Large Language Models 14 Jan 2021 · 1 repository · arXiv:2101.05783
-
Training Data Leakage Analysis in Language Models 14 Jan 2021 · 0 repositories · arXiv:2101.05405
-
Towards Practical Adam: Non-Convexity, Convergence Theory, and Mini-Batch Acceleration 14 Jan 2021 · 0 repositories · arXiv:2101.05471
-
Transformer-based Language Model Fine-tuning Methods for COVID-19 Fake News Detection 14 Jan 2021 · 0 repositories · arXiv:2101.05509
-
WER-BERT: Automatic WER Estimation with BERT in a Balanced Ordinal Classification Paradigm 14 Jan 2021 · 0 repositories · arXiv:2101.05478
-
Coarse and Fine-Grained Hostility Detection in Hindi Posts using Fine Tuned Multilingual Embeddings 13 Jan 2021 · 1 repository · arXiv:2101.04998
-
Experimental Evaluation of Deep Learning models for Marathi Text Classification 13 Jan 2021 · 0 repositories · arXiv:2101.04899
-
Heterogeneous Network Embedding for Deep Semantic Relevance Match in E-commerce Search 13 Jan 2021 · 0 repositories · arXiv:2101.04850
-
LaDiff ULMFiT: A Layer Differentiated training approach for ULMFiT 13 Jan 2021 · 1 repository · arXiv:2101.04965
-
Fake News Detection System using XLNet model with Topic Distributions: CONSTRAINT@AAAI2021 Shared Task 12 Jan 2021 · 0 repositories · arXiv:2101.11425
-
Neural Contract Element Extraction Revisited: Letters from Sesame Street 12 Jan 2021 · 0 repositories · arXiv:2101.04355
-
Neural News Recommendation with Negative Feedback 12 Jan 2021 · 0 repositories · arXiv:2101.04328
-
Of Non-Linearity and Commutativity in BERT 12 Jan 2021 · 1 repository · arXiv:2101.04547
-
A More Efficient Chinese Named Entity Recognition base on BERT and Syntactic Analysis 11 Jan 2021 · 0 repositories · arXiv:2101.11423
-
AT-BERT: Adversarial Training BERT for Acronym Identification Winning Solution for SDU@AAAI-21 11 Jan 2021 · 0 repositories · arXiv:2101.03700
-
BERT-GT: Cross-sentence n-ary relation extraction with BERT and Graph Transformer 11 Jan 2021 · 0 repositories · arXiv:2101.04158
-
Evaluation of Deep Learning Models for Hostility Detection in Hindi Text 11 Jan 2021 · 0 repositories · arXiv:2101.04144
-
Investigating the Vision Transformer Model for Image Retrieval Tasks 11 Jan 2021 · 0 repositories · arXiv:2101.03771
-
Revisiting Mahalanobis Distance for Transformer-Based Out-of-Domain Detection 11 Jan 2021 · 1 repository · arXiv:2101.03778
-
Spherical Transformer: Adapting Spherical Signal to CNNs 11 Jan 2021 · 0 repositories · arXiv:2101.03848
-
BERT & Family Eat Word Salad: Experiments with Text Understanding 10 Jan 2021 · 1 repository · arXiv:2101.03453
-
Channel Boosting Feature Ensemble for Radar-based Object Detection 10 Jan 2021 · 0 repositories · arXiv:2101.03531
-
Cisco at AAAI-CAD21 shared task: Predicting Emphasis in Presentation Slides using Contextualized Embeddings 10 Jan 2021 · 1 repository · arXiv:2101.11422
-
Deep Reinforcement Learning with Function Properties in Mean Reversion Strategies 9 Jan 2021 · 1 repository · arXiv:2101.03418
-
Learning Better Sentence Representation with Syntax Information 9 Jan 2021 · 0 repositories · arXiv:2101.03343
-
Trankit: A Light-Weight Transformer-based Toolkit for Multilingual Natural Language Processing 9 Jan 2021 · 1 repository · arXiv:2101.03289
-
Contextual Non-Local Alignment over Full-Scale Representation for Text-Based Person Search 8 Jan 2021 · 2 repositories · arXiv:2101.03036
-
Leveraging Multilingual Transformers for Hate Speech Detection 8 Jan 2021 · 1 repository · arXiv:2101.03207
-
Misspelling Correction with Pre-trained Contextual Language Model 8 Jan 2021 · 0 repositories · arXiv:2101.03204
-
Applying Transfer Learning for Improving Domain-Specific Search Experience Using Query to Question Similarity 7 Jan 2021 · 0 repositories · arXiv:2101.02351
-
Compound Word Transformer: Learning to Compose Full-Song Music over Dynamic Directed Hypergraphs 7 Jan 2021 · 5 repositories · arXiv:2101.02402Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples)
-
Exploring Text-transformers in AAAI 2021 Shared Task: COVID-19 Fake News Detection in English 7 Jan 2021 · 1 repository · arXiv:2101.02359
-
Homonym Identification using BERT -- Using a Clustering Approach 7 Jan 2021 · 0 repositories · arXiv:2101.02398
-
TrackFormer: Multi-Object Tracking with Transformers 7 Jan 2021 · 2 repositories · arXiv:2101.02702
-
Transformer-based approach towards music emotion recognition from lyrics 6 Jan 2021 · 1 repository · arXiv:2101.02051
-
AutoDropout: Learning Dropout Patterns to Regularize Deep Networks 5 Jan 2021 · 1 repository · arXiv:2101.01761
-
COVID-19: Comparative Analysis of Methods for Identifying Articles Related to Therapeutics and Vaccines without Using Labeled Data 5 Jan 2021 · 0 repositories · arXiv:2101.02017
-
I-BERT: Integer-only BERT Quantization 5 Jan 2021 · 7 repositories · arXiv:2101.01321
-
Generalized Latency Performance Estimation for Once-For-All Neural Architecture Search 4 Jan 2021 · 2 repositories · arXiv:2101.00732
-
Improving reference mining in patents with BERT 4 Jan 2021 · 1 repository · arXiv:2101.01039
-
Improving Portuguese Semantic Role Labeling with Transformers and Transfer Learning 4 Jan 2021 · 1 repository · arXiv:2101.01213
-
Transformers in Vision: A Survey 4 Jan 2021 · 0 repositories · arXiv:2101.01169
-
Using BART to Perform Pareto Optimization and Quantify its Uncertainties 4 Jan 2021 · 0 repositories · arXiv:2101.02558
-
An Efficient Transformer Decoder with Compressed Sub-layers 3 Jan 2021 · 0 repositories · arXiv:2101.00542
-
A Robust and Domain-Adaptive Approach for Low-Resource Named Entity Recognition 2 Jan 2021 · 1 repository · arXiv:2101.00388
-
End-to-End Training of Neural Retrievers for Open-Domain Question Answering 2 Jan 2021 · 2 repositories · arXiv:2101.00408
-
KM-BART: Knowledge Enhanced Multimodal BART for Visual Commonsense Generation 2 Jan 2021 · 1 repository · arXiv:2101.00419
-
Lex-BERT: Enhancing BERT based NER with lexicons 2 Jan 2021 · 0 repositories · arXiv:2101.00396
-
Superbizarre Is Not Superb: Derivational Morphology Improves BERT's Interpretation of Complex Words 2 Jan 2021 · 1 repository · arXiv:2101.00403Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 2 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
What all do audio transformer models hear? Probing Acoustic Representations for Language Delivery and its Structure 2 Jan 2021 · 0 repositories · arXiv:2101.00387
-
Learning to Generate Task-Specific Adapters from Task Description 2 Jan 2021 · 1 repository · arXiv:2101.00420
-
A Geometric Analysis of Deep Generative Image Models and Its Applications 1 Jan 2021 · 1 repository
-
A New Variant of Stochastic Heavy ball Optimization Method for Deep Learning 1 Jan 2021 · 0 repositories
-
Adam revisited: a weighted past gradients perspective 1 Jan 2021 · 0 repositories · arXiv:2101.00238
-
Adding Recurrence to Pretrained Transformers 1 Jan 2021 · 0 repositories
-
Analogical Reasoning for Visually Grounded Compositional Generalization 1 Jan 2021 · 0 repositories
-
Apollo: An Adaptive Parameter-wised Diagonal Quasi-Newton Method for Nonconvex Stochastic Optimization 1 Jan 2021 · 0 repositories
-
AriEL: Volume Coding for Sentence Generation Comparisons 1 Jan 2021 · 0 repositories
-
Attention Is Not Enough: Mitigating the Distribution Discrepancy in Asynchronous Multimodal Sequence Fusion 1 Jan 2021 · 0 repositories
-
Block Skim Transformer for Efficient Question Answering 1 Jan 2021 · 0 repositories
-
BROS: A Pre-trained Language Model for Understanding Texts in Document 1 Jan 2021 · 0 repositories
-
Cluster-Former: Clustering-based Sparse Transformer for Question Answering 1 Jan 2021 · 0 repositories
-
Cluster & Tune: Enhance BERT Performance in Low Resource Text Classification 1 Jan 2021 · 0 repositories
-
Convergent Adaptive Gradient Methods in Decentralized Optimization 1 Jan 2021 · 0 repositories
-
CrackFormer: Transformer Network for Fine-Grained Crack Detection 1 Jan 2021 · 1 repository
-
Cross-Probe BERT for Efficient and Effective Cross-Modal Search 1 Jan 2021 · 0 repositories
-
DACT-BERT: Increasing the efficiency and interpretability of BERT by using adaptive computation time. 1 Jan 2021 · 0 repositories
-
Data-aware Low-Rank Compression for Large NLP Models 1 Jan 2021 · 0 repositories
-
Deep Learning Proteins using a Triplet-BERT network 1 Jan 2021 · 0 repositories
-
Deep Representational Re-tuning using Contrastive Tension 1 Jan 2021 · 1 repository
-
Discovering Human Interactions With Large-Vocabulary Objects via Query and Multi-Scale Detection 1 Jan 2021 · 0 repositories
-
Do Transformers Understand Polynomial Simplification? 1 Jan 2021 · 0 repositories
-
Domain-slot Relationship Modeling using a Pre-trained Language Encoder for Multi-Domain Dialogue State Tracking 1 Jan 2021 · 0 repositories
-
Dynamic DETR: End-to-End Object Detection With Dynamic Attention 1 Jan 2021 · 0 repositories
-
Erasure for Advancing: Dynamic Self-Supervised Learning for Commonsense Reasoning 1 Jan 2021 · 0 repositories
-
Event-Based Video Reconstruction Using Transformer 1 Jan 2021 · 1 repository
-
Exploring Routing Strategies for Multilingual Mixture-of-Experts Models 1 Jan 2021 · 0 repositories
-
EXPLORING VULNERABILITIES OF BERT-BASED APIS 1 Jan 2021 · 0 repositories
-
Factored Action Spaces in Deep Reinforcement Learning 1 Jan 2021 · 0 repositories
-
Frequency-Aware Spatiotemporal Transformers for Video Inpainting Detection 1 Jan 2021 · 0 repositories
-
Generalizing Tree Models for Improving Prediction Accuracy 1 Jan 2021 · 0 repositories
-
High-Performance Discriminative Tracking With Transformers 1 Jan 2021 · 0 repositories
-
How Multipurpose Are Language Models? 1 Jan 2021 · 0 repositories
-
HyperGrid Transformers: Towards A Single Model for Multiple Tasks 1 Jan 2021 · 0 repositories
-
Image Harmonization With Transformer 1 Jan 2021 · 1 repository
-
Improving Generalizability of Protein Sequence Models via Data Augmentations 1 Jan 2021 · 0 repositories
-
Improving Machine Translation by Searching Skip Connections Efficiently 1 Jan 2021 · 0 repositories
-
Isotropy in the Contextual Embedding Space: Clusters and Manifolds 1 Jan 2021 · 0 repositories