Methods › General › Stochastic Optimization › Adam › Papers, page 186
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 186 of 244: papers 18,501 to 18,600 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Language Models are Few-shot Multilingual Learners 16 Sep 2021 · 1 repository · arXiv:2109.07684Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Let the CAT out of the bag: Contrastive Attributed explanations for Text 16 Sep 2021 · 0 repositories · arXiv:2109.07983
-
MeLT: Message-Level Transformer with Masked Document Representations as Pre-Training for Stance Detection 16 Sep 2021 · 1 repository · arXiv:2109.08113
-
MOVER: Mask, Over-generate and Rank for Hyperbole Generation 16 Sep 2021 · 1 repository · arXiv:2109.07726
-
RetrievalSum: A Retrieval Enhanced Framework for Abstractive Summarization 16 Sep 2021 · 0 repositories · arXiv:2109.07943
-
Revisiting Tri-training of Dependency Parsers 16 Sep 2021 · 2 repositories · arXiv:2109.08122
-
Scaling Laws for Neural Machine Translation 16 Sep 2021 · 0 repositories · arXiv:2109.07740
-
Sparse Factorization of Large Square Matrices 16 Sep 2021 · 1 repository · arXiv:2109.08184
-
TANet: A new Paradigm for Global Face Super-resolution via Transformer-CNN Aggregation Network 16 Sep 2021 · 0 repositories · arXiv:2109.08174
-
The NiuTrans System for the WMT21 Efficiency Task 16 Sep 2021 · 1 repository · arXiv:2109.08003
-
The NiuTrans System for WNGT 2020 Efficiency Task 16 Sep 2021 · 2 repositories · arXiv:2109.08008Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Utterance-level neural confidence measure for end-to-end children speech recognition 16 Sep 2021 · 0 repositories · arXiv:2109.07750
-
Anchor DETR: Query Design for Transformer-Based Object Detection 15 Sep 2021 · 2 repositories · arXiv:2109.07107
-
Attention Is Indeed All You Need: Semantically Attention-Guided Decoding for Data-to-Text NLG 15 Sep 2021 · 1 repository · arXiv:2109.07043
-
BERT is Robust! A Case Against Synonym-Based Adversarial Examples in Text Classification 15 Sep 2021 · 0 repositories · arXiv:2109.07403
-
EfficientBERT: Progressively Searching Multilayer Perceptron via Warm-up Knowledge Distillation 15 Sep 2021 · 1 repository · arXiv:2109.07222Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Efficient Domain Adaptation of Language Models via Adaptive Tokenization 15 Sep 2021 · 0 repositories · arXiv:2109.07460
-
Enhancing Clinical Information Extraction with Transferred Contextual Embeddings 15 Sep 2021 · 0 repositories · arXiv:2109.07243
-
Complementary Feature Enhanced Network with Vision Transformer for Image Dehazing 15 Sep 2021 · 1 repository · arXiv:2109.07100
-
Improving Text Auto-Completion with Next Phrase Prediction 15 Sep 2021 · 0 repositories · arXiv:2109.07067
-
Incorporating Residual and Normalization Layers into Analysis of Masked Language Models 15 Sep 2021 · 2 repositories · arXiv:2109.07152Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Learning to Match Job Candidates Using Multilingual Bi-Encoder BERT 15 Sep 2021 · 0 repositories · arXiv:2109.07157
-
MISSFormer: An Effective Medical Image Segmentation Transformer 15 Sep 2021 · 1 repository · arXiv:2109.07162
-
On the Universality of Deep Contextual Language Models 15 Sep 2021 · 0 repositories · arXiv:2109.07140
-
PnP-DETR: Towards Efficient Visual Analysis with Transformers 15 Sep 2021 · 1 repository · arXiv:2109.07036Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Pose Transformers (POTR): Human Motion Prediction with Non-Autoregressive Transformers 15 Sep 2021 · 1 repository · arXiv:2109.07531
-
RetroPrime: A Diverse, plausible and Transformer-based method for Single-Step retrosynthesis predictions 15 Sep 2021 · 1 repository
-
Sequence Length is a Domain: Length-based Overfitting in Transformer Models 15 Sep 2021 · 1 repository · arXiv:2109.07276
-
SupCL-Seq: Supervised Contrastive Learning for Downstream Optimized Sequence Representations 15 Sep 2021 · 1 repository · arXiv:2109.07424
-
The Unreasonable Effectiveness of the Baseline: Discussing SVMs in Legal Text Classification 15 Sep 2021 · 0 repositories · arXiv:2109.07234
-
Topic Transferable Table Question Answering 15 Sep 2021 · 1 repository · arXiv:2109.07377
-
Towards Incremental Transformers: An Empirical Analysis of Transformer Models for Incremental NLU 15 Sep 2021 · 1 repository · arXiv:2109.07364
-
Transformer-based Language Models for Factoid Question Answering at BioASQ9b 15 Sep 2021 · 1 repository · arXiv:2109.07185
-
Transformer-based Lexically Constrained Headline Generation 15 Sep 2021 · 1 repository · arXiv:2109.07080
-
A pragmatic approach to estimating average treatment effects from EHR data: the effect of prone positioning on mechanically ventilated COVID-19 patients 14 Sep 2021 · 1 repository · arXiv:2109.06707
-
A Temporal Variational Model for Story Generation 14 Sep 2021 · 3 repositories · arXiv:2109.06807
-
A Three Step Training Approach with Data Augmentation for Morphological Inflection 14 Sep 2021 · 0 repositories · arXiv:2109.07006
-
conSultantBERT: Fine-tuned Siamese Sentence-BERT for Matching Jobs and Job Seekers 14 Sep 2021 · 0 repositories · arXiv:2109.06501
-
Deep learning-based NLP Data Pipeline for EHR Scanned Document Information Extraction 14 Sep 2021 · 0 repositories · arXiv:2110.11864
-
Evaluating Biomedical BERT Models for Vocabulary Alignment at Scale in the UMLS Metathesaurus 14 Sep 2021 · 0 repositories · arXiv:2109.13348
-
Explainable Identification of Dementia from Transcripts using Transformer Networks 14 Sep 2021 · 0 repositories · arXiv:2109.06980
-
Exploring Personality and Online Social Engagement: An Investigation of MBTI Users on Twitter 14 Sep 2021 · 0 repositories · arXiv:2109.06402
-
Frequency Effects on Syntactic Rule Learning in Transformers 14 Sep 2021 · 1 repository · arXiv:2109.07020
-
Learning Bill Similarity with Annotated and Augmented Corpora of Bills 14 Sep 2021 · 1 repository · arXiv:2109.06527
-
Legal Transformer Models May Not Always Help 14 Sep 2021 · 0 repositories · arXiv:2109.06862
-
On the Language-specificity of Multilingual BERT and the Impact of Fine-tuning 14 Sep 2021 · 1 repository · arXiv:2109.06935
-
Semantic Answer Type Prediction using BERT: IAI at the ISWC SMART Task 2020 14 Sep 2021 · 0 repositories · arXiv:2109.06714
-
Structure-Enhanced Pop Music Generation via Harmony-Aware Learning 14 Sep 2021 · 1 repository · arXiv:2109.06441
-
Tribrid: Stance Classification with Neural Inconsistency Detection 14 Sep 2021 · 1 repository · arXiv:2109.06508
-
Vision Transformer for Learning Driving Policies in Complex Multi-Agent Environments 14 Sep 2021 · 0 repositories · arXiv:2109.06514
-
YES SIR!Optimizing Semantic Space of Negatives with Self-Involvement Ranker 14 Sep 2021 · 0 repositories · arXiv:2109.06436
-
Attention Weights in Transformer NMT Fail Aligning Words Between Sequences but Largely Explain Model Predictions 13 Sep 2021 · 0 repositories · arXiv:2109.05853
-
CDTrans: Cross-domain Transformer for Unsupervised Domain Adaptation 13 Sep 2021 · 2 repositories · arXiv:2109.06165Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
CPT: A Pre-Trained Unbalanced Transformer for Both Chinese Language Understanding and Generation 13 Sep 2021 · 1 repository · arXiv:2109.05729
-
Effectiveness of Pre-training for Few-shot Intent Classification 13 Sep 2021 · 0 repositories · arXiv:2109.05782
-
Evaluating Transferability of BERT Models on Uralic Languages 13 Sep 2021 · 1 repository · arXiv:2109.06327
-
Exploring a Unified Sequence-To-Sequence Transformer for Medical Product Safety Monitoring in Social Media 13 Sep 2021 · 1 repository · arXiv:2109.05815
-
Keyword Extraction for Improved Document Retrieval in Conversational Search 13 Sep 2021 · 0 repositories · arXiv:2109.05979
-
KroneckerBERT: Learning Kronecker Decomposition for Pre-trained Language Models via Knowledge Distillation 13 Sep 2021 · 0 repositories · arXiv:2109.06243
-
Mitigating Language-Dependent Ethnic Bias in BERT 13 Sep 2021 · 1 repository · arXiv:2109.05704Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Not All Models Localize Linguistic Knowledge in the Same Place: A Layer-wise Probing on BERToids' Representations 13 Sep 2021 · 0 repositories · arXiv:2109.05958
-
Connecting degree and polarity: An artificial language learning study 13 Sep 2021 · 1 repository · arXiv:2109.06333
-
On Pursuit of Designing Multi-modal Transformer for Video Grounding 13 Sep 2021 · 0 repositories · arXiv:2109.06085
-
Phrase-BERT: Improved Phrase Embeddings from BERT with an Application to Corpus Exploration 13 Sep 2021 · 2 repositories · arXiv:2109.06304Syntology official (archive's flag): 2 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 1 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Question Answering over Electronic Devices: A New Benchmark Dataset and a Multi-Task Learning based QA Framework 13 Sep 2021 · 1 repository · arXiv:2109.05897Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
UniMS: A Unified Framework for Multimodal Summarization with Knowledge Distillation 13 Sep 2021 · 0 repositories · arXiv:2109.05812
-
ArtiBoost: Boosting Articulated 3D Hand-Object Pose Estimation via Online Exploration and Synthesis 12 Sep 2021 · 2 repositories · arXiv:2109.05488
-
Constructing Phrase-level Semantic Labels to Form Multi-Grained Supervision for Image-Text Retrieval 12 Sep 2021 · 0 repositories · arXiv:2109.05523
-
FLiText: A Faster and Lighter Semi-Supervised Text Classification with Convolution Networks 12 Sep 2021 · 1 repository · arXiv:2110.11869
-
Levenshtein Training for Word-level Quality Estimation 12 Sep 2021 · 1 repository · arXiv:2109.05611
-
Single-Read Reconstruction for DNA Data Storage Using Transformers 12 Sep 2021 · 0 repositories · arXiv:2109.05478
-
Sparse MLP for Image Recognition: Is Self-Attention Really Necessary? 12 Sep 2021 · 2 repositories · arXiv:2109.05422
-
TEASEL: A Transformer-Based Speech-Prefixed Language Model 12 Sep 2021 · 1 repository · arXiv:2109.05522Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Adaptive network reliability analysis: Methodology and applications to power grid 11 Sep 2021 · 0 repositories · arXiv:2109.05360
-
Bornon: Bengali Image Captioning with Transformer-based Deep learning approach 11 Sep 2021 · 0 repositories · arXiv:2109.05218
-
Clinical Trial Information Extraction with BERT 11 Sep 2021 · 0 repositories · arXiv:2110.10027
-
Empirical Analysis of Training Strategies of Transformer-based Japanese Chit-chat Systems 11 Sep 2021 · 1 repository · arXiv:2109.05217
-
Multilingual Translation via Grafting Pre-trained Language Models 11 Sep 2021 · 1 repository · arXiv:2109.05256
-
TopicRefine: Joint Topic Prediction and Dialogue Response Generation for Multi-turn End-to-End Dialogue System 11 Sep 2021 · 0 repositories · arXiv:2109.05187
-
An Empirical Study of GPT-3 for Few-Shot Knowledge-Based VQA 10 Sep 2021 · 1 repository · arXiv:2109.05014Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Artificial Text Detection via Examining the Topology of Attention Maps 10 Sep 2021 · 2 repositories · arXiv:2109.04825Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 8 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Block Pruning For Faster Transformers 10 Sep 2021 · 1 repository · arXiv:2109.04838
-
D-REX: Dialogue Relation Extraction with Explanations 10 Sep 2021 · 1 repository · arXiv:2109.05126
-
Enhancing Self-Disclosure In Neural Dialog Models By Candidate Re-ranking 10 Sep 2021 · 0 repositories · arXiv:2109.05090
-
FBERT: A Neural Transformer for Identifying Offensive Content 10 Sep 2021 · 0 repositories · arXiv:2109.05074
-
How May I Help You? Using Neural Text Simplification to Improve Downstream NLP Tasks 10 Sep 2021 · 1 repository · arXiv:2109.04604
-
IndoBERTweet: A Pretrained Language Model for Indonesian Twitter with Effective Domain-Specific Vocabulary Initialization 10 Sep 2021 · 1 repository · arXiv:2109.04607
-
Instance-Conditioned GAN 10 Sep 2021 · 1 repository · arXiv:2109.05070
-
Mixture-of-Partitions: Infusing Large Biomedical Knowledge Graphs into BERT 10 Sep 2021 · 1 repository · arXiv:2109.04810Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
On the validity of pre-trained transformers for natural language processing in the software engineering domain 10 Sep 2021 · 0 repositories · arXiv:2109.04738
-
Real-time multimodal image registration with partial intraoperative point-set data 10 Sep 2021 · 0 repositories · arXiv:2109.05023
-
RoR: Read-over-Read for Long Document Machine Reading Comprehension 10 Sep 2021 · 1 repository · arXiv:2109.04780
-
Temporal Pyramid Transformer with Multimodal Interaction for Video Question Answering 10 Sep 2021 · 1 repository · arXiv:2109.04735
-
What Changes Can Large-scale Language Models Bring? Intensive Study on HyperCLOVA: Billions-scale Korean Generative Pretrained Transformers 10 Sep 2021 · 2 repositories · arXiv:2109.04650Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
A Three-Stage Learning Framework for Low-Resource Knowledge-Grounded Dialogue Generation 9 Sep 2021 · 1 repository · arXiv:2109.04096Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 27 harvested samples) · 5 pointer-only (licence)
-
All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational Quality 9 Sep 2021 · 1 repository · arXiv:2109.04404
-
Bag of Tricks for Optimizing Transformer Efficiency 9 Sep 2021 · 1 repository · arXiv:2109.04030
-
BERT, mBERT, or BiBERT? A Study on Contextualized Embeddings for Neural Machine Translation 9 Sep 2021 · 2 repositories · arXiv:2109.04588Syntology official (archive's flag): 1 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
DAN: Decentralized Attention-based Neural Network for the MinMax Multiple Traveling Salesman Problem 9 Sep 2021 · 0 repositories · arXiv:2109.04205
-
ESimCSE: Enhanced Sample Building Method for Contrastive Learning of Unsupervised Sentence Embedding 9 Sep 2021 · 2 repositories · arXiv:2109.04380