Methods › General › Stochastic Optimization › Adam › Papers, page 136
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 136 of 244: papers 13,501 to 13,600 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TRON: Transformer Neural Network Acceleration with Non-Coherent Silicon Photonics 22 Mar 2023 · 0 repositories · arXiv:2303.12914
-
A Complete Survey on Generative AI (AIGC): Is ChatGPT from GPT-4 to GPT-5 All You Need? 21 Mar 2023 · 0 repositories · arXiv:2303.11717
-
Artificial muses: Generative Artificial Intelligence Chatbots Have Risen to Human-Level Creativity 21 Mar 2023 · 0 repositories · arXiv:2303.12003
-
ChatGPT and a New Academic Reality: Artificial Intelligence-Written Research Papers and the Ethics of the Large Language Models in Scholarly Publishing 21 Mar 2023 · 0 repositories · arXiv:2303.13367
-
cTBLS: Augmenting Large Language Models with Conversational Tables 21 Mar 2023 · 1 repository · arXiv:2303.12024
-
Debiased Contrastive Learning for Sequential Recommendation 21 Mar 2023 · 1 repository · arXiv:2303.11780Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Difficulty in chirality recognition for Transformer architectures learning chemical structures from string 21 Mar 2023 · 1 repository · arXiv:2303.11593
-
Fine-tuning ClimateBert transformer with ClimaText for the disclosure analysis of climate-related financial risks 21 Mar 2023 · 0 repositories · arXiv:2303.13373
-
Is BERT Blind? Exploring the Effect of Vision-and-Language Pretraining on Visual Language Understanding 21 Mar 2023 · 1 repository · arXiv:2303.12513
-
Learning A Sparse Transformer Network for Effective Image Deraining 21 Mar 2023 · 1 repository · arXiv:2303.11950Syntology official (archive's flag): 8 ran · 8 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 8 samples that ran constructed an object rather than computing a result (of 11 harvested samples) · 11 pointer-only (licence)
-
Machine Learning for Brain Disorders: Transformers and Visual Transformers 21 Mar 2023 · 0 repositories · arXiv:2303.12068
-
MSTFormer: Motion Inspired Spatial-temporal Transformer with Dynamic-aware Attention for long-term Vessel Trajectory Prediction 21 Mar 2023 · 1 repository · arXiv:2303.11540
-
Multimodal Pre-training Framework for Sequential Recommendation via Contrastive Learning 21 Mar 2023 · 0 repositories · arXiv:2303.11879
-
Robust Table Structure Recognition with Dynamic Queries Enhanced Detection Transformer 21 Mar 2023 · 0 repositories · arXiv:2303.11615
-
Sparse-IFT: Sparse Iso-FLOP Transformations for Maximizing Training Efficiency 21 Mar 2023 · 2 repositories · arXiv:2303.11525Syntology official (archive's flag): 21 ran · 21 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 24 harvested samples) · 2 pointer-only (licence)
-
The Multiscale Surface Vision Transformer 21 Mar 2023 · 1 repository · arXiv:2303.11909
-
Capabilities of GPT-4 on Medical Challenge Problems 20 Mar 2023 · 1 repository · arXiv:2303.13375
-
Character, Word, or Both? Revisiting the Segmentation Granularity for Chinese Pre-trained Language Models 20 Mar 2023 · 1 repository · arXiv:2303.10893
-
DeID-GPT: Zero-shot Medical Text De-Identification by GPT-4 20 Mar 2023 · 1 repository · arXiv:2303.11032
-
EVA-02: A Visual Representation for Neon Genesis 20 Mar 2023 · 6 repositories · arXiv:2303.11331
-
FF-Former: Swin Fourier Transformer for Nighttime Flare Removal 20 Mar 2023 · 0 repositories
-
Mind meets machine: Unravelling GPT-4's cognitive psychology 20 Mar 2023 · 0 repositories · arXiv:2303.11436
-
Multi-task Transformer with Relation-attention and Type-attention for Named Entity Recognition 20 Mar 2023 · 0 repositories · arXiv:2303.10870
-
PanGu-Σ: Towards Trillion Parameter Language Model with Sparse Heterogeneous Computing 20 Mar 2023 · 0 repositories · arXiv:2303.10845
-
Reflexion: Language Agents with Verbal Reinforcement Learning 20 Mar 2023 · 5 repositories · arXiv:2303.11366Syntology community repositories only · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
SGFormer: Semantic Graph Transformer for Point Cloud-based 3D Scene Graph Generation 20 Mar 2023 · 1 repository · arXiv:2303.11048
-
Sparse Distributed Memory is a Continual Learner 20 Mar 2023 · 1 repository · arXiv:2303.11934
-
Towards End-to-End Generative Modeling of Long Videos with Memory-Efficient Bidirectional Transformers 20 Mar 2023 · 1 repository · arXiv:2303.11251
-
Bangla Grammatical Error Detection Using T5 Transformer Model 19 Mar 2023 · 2 repositories · arXiv:2303.10612
-
CCTV-Gun: Benchmarking Handgun Detection in CCTV Images 19 Mar 2023 · 1 repository · arXiv:2303.10703
-
CTRAN: CNN-Transformer-based Network for Natural Language Understanding 19 Mar 2023 · 1 repository · arXiv:2303.10606
-
Multiscale Audio Spectrogram Transformer for Efficient Audio Classification 19 Mar 2023 · 0 repositories · arXiv:2303.10757
-
PACO: Provocation Involving Action, Culture, and Oppression 19 Mar 2023 · 0 repositories · arXiv:2303.12808
-
Spatial-temporal Transformer for Affective Behavior Analysis 19 Mar 2023 · 0 repositories · arXiv:2303.10561
-
A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models 18 Mar 2023 · 0 repositories · arXiv:2303.10420
-
An Empirical Study of Pre-trained Language Models in Simple Knowledge Graph Question Answering 18 Mar 2023 · 1 repository · arXiv:2303.10368
-
Channel-Aware Distillation Transformer for Depth Estimation on Nano Drones 18 Mar 2023 · 1 repository · arXiv:2303.10386
-
Discovering Predictable Latent Factors for Time Series Forecasting 18 Mar 2023 · 1 repository · arXiv:2303.10426
-
NoisyHate: Mining Online Human-Written Perturbations for Realistic Robustness Benchmarking of Content Moderation Models 18 Mar 2023 · 0 repositories · arXiv:2303.10430
-
SPDF: Sparse Pre-training and Dense Fine-tuning for Large Language Models 18 Mar 2023 · 0 repositories · arXiv:2303.10464
-
Vision Transformer-based Model for Severity Quantification of Lung Pneumonia Using Chest X-ray Images 18 Mar 2023 · 1 repository · arXiv:2303.11935
-
CerviFormer: A Pap-smear based cervical cancer classification method using cross attention and latent transformer 17 Mar 2023 · 0 repositories · arXiv:2303.10222
-
CoLT5: Faster Long-Range Transformers with Conditional Computation 17 Mar 2023 · 0 repositories · arXiv:2303.09752Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
GADformer: A Transparent Transformer Model for Group Anomaly Detection on Trajectories 17 Mar 2023 · 1 repository · arXiv:2303.09841
-
GNNFormer: A Graph-based Framework for Cytopathology Report Generation 17 Mar 2023 · 0 repositories · arXiv:2303.09956
-
GPTs are GPTs: An Early Look at the Labor Market Impact Potential of Large Language Models 17 Mar 2023 · 0 repositories · arXiv:2303.10130
-
HDformer: A Higher Dimensional Transformer for Diabetes Detection Utilizing Long Range Vascular Signals 17 Mar 2023 · 0 repositories · arXiv:2303.11340
-
LSwinSR: UAV Imagery Super-Resolution based on Linear Swin Transformer 17 Mar 2023 · 1 repository · arXiv:2303.10232
-
Pedestrain detection for low-light vision proposal 17 Mar 2023 · 0 repositories · arXiv:2303.12725
-
Star-Net: Improving Single Image Desnowing Model With More Efficient Connection and Diverse Feature Interaction 17 Mar 2023 · 0 repositories · arXiv:2303.09988
-
Trained on 100 million words and still in shape: BERT meets British National Corpus 17 Mar 2023 · 2 repositories · arXiv:2303.09859Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Resolution Enhancement Processing on Low Quality Images Using Swin Transformer Based on Interval Dense Connection Strategy 16 Mar 2023 · 2 repositories · arXiv:2303.09190
-
Block-wise Bit-Compression of Transformer-based Models 16 Mar 2023 · 0 repositories · arXiv:2303.09184
-
Can Generative Pre-trained Transformers (GPT) Pass Assessments in Higher Education Programming Courses? 16 Mar 2023 · 0 repositories · arXiv:2303.09325
-
Emotional Reaction Intensity Estimation Based on Multimodal Data 16 Mar 2023 · 0 repositories · arXiv:2303.09167
-
Energy Management of Multi-mode Plug-in Hybrid Electric Vehicle using Multi-agent Deep Reinforcement Learning 16 Mar 2023 · 0 repositories · arXiv:2303.09658
-
Facial Affect Recognition based on Transformer Encoder and Audiovisual Fusion for the ABAW5 Challenge 16 Mar 2023 · 0 repositories · arXiv:2303.09158
-
Highly Accurate Quantum Chemical Property Prediction with Uni-Mol+ 16 Mar 2023 · 2 repositories · arXiv:2303.16982Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
How well do Large Language Models perform in Arithmetic tasks? 16 Mar 2023 · 1 repository · arXiv:2304.02015
-
Hybrid Spectral Denoising Transformer with Guided Attention 16 Mar 2023 · 1 repository · arXiv:2303.09040Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
InCrowdFormer: On-Ground Pedestrian World Model From Egocentric Views 16 Mar 2023 · 0 repositories · arXiv:2303.09534
-
Jump to Conclusions: Short-Cutting Transformers With Linear Transformations 16 Mar 2023 · 2 repositories · arXiv:2303.09435Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Measuring Improvement of F₁-Scores in Detection of Self-Admitted Technical Debt 16 Mar 2023 · 0 repositories · arXiv:2303.09617
-
Unifying Top-down and Bottom-up Scanpath Prediction Using Transformers 16 Mar 2023 · 1 repository · arXiv:2303.09383
-
PSVT: End-to-End Multi-person 3D Pose and Shape Estimation with Progressive Video Transformers 16 Mar 2023 · 0 repositories · arXiv:2303.09187
-
Rehearsal-Free Domain Continual Face Anti-Spoofing: Generalize More and Forget Less 16 Mar 2023 · 0 repositories · arXiv:2303.09914Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
SmartBERT: A Promotion of Dynamic Early Exiting Mechanism for Accelerating BERT Inference 16 Mar 2023 · 0 repositories · arXiv:2303.09266Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
Steering Prototypes with Prompt-tuning for Rehearsal-free Continual Learning 16 Mar 2023 · 2 repositories · arXiv:2303.09447Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Taming Diffusion Models for Audio-Driven Co-Speech Gesture Generation 16 Mar 2023 · 1 repository · arXiv:2303.09119
-
Towards the Scalable Evaluation of Cooperativeness in Language Models 16 Mar 2023 · 0 repositories · arXiv:2303.13360
-
Translating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and Potential 16 Mar 2023 · 0 repositories · arXiv:2303.09038
-
Automated Interactive Domain-Specific Conversational Agents that Understand Human Dialogs 15 Mar 2023 · 0 repositories · arXiv:2303.08941
-
Leveraging TCN and Transformer for effective visual-audio fusion in continuous emotion recognition 15 Mar 2023 · 1 repository · arXiv:2303.08356
-
DeepMIM: Deep Supervision for Masked Image Modeling 15 Mar 2023 · 1 repository · arXiv:2303.08817
-
Efficient Uncertainty Estimation with Gaussian Process for Reliable Dialog Response Retrieval 15 Mar 2023 · 0 repositories · arXiv:2303.08599
-
FastInst: A Simple Query-Based Model for Real-Time Instance Segmentation 15 Mar 2023 · 1 repository · arXiv:2303.08594Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
GCRE-GPT: A Generative Model for Comparative Relation Extraction 15 Mar 2023 · 0 repositories · arXiv:2303.08601
-
GPT-4 Technical Report 15 Mar 2023 · 11 repositories · arXiv:2303.08774Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 2 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Implicit Ray-Transformers for Multi-view Remote Sensing Image Segmentation 15 Mar 2023 · 0 repositories · arXiv:2303.08401
-
LEP-AD: Language Embedding of Proteins and Attention to Drugs predicts drug target interactions 15 Mar 2023 · 1 repository
-
Multi-Exposure HDR Composition by Gated Swin Transformer 15 Mar 2023 · 0 repositories · arXiv:2303.08704
-
Multi Modal Facial Expression Recognition with Transformer-Based Fusion Networks and Dynamic Sampling 15 Mar 2023 · 0 repositories · arXiv:2303.08419
-
SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models 15 Mar 2023 · 1 repository · arXiv:2303.08896Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
SpatialFormer: Semantic and Target Aware Attentions for Few-Shot Learning 15 Mar 2023 · 1 repository · arXiv:2303.09281
-
Transformer Models for Type Inference in the Simply Typed Lambda Calculus: A Case Study in Deep Learning for Code 15 Mar 2023 · 0 repositories · arXiv:2304.10500
-
VDDT: Improving Vessel Detection with Deformable Transfomer 15 Mar 2023 · 0 repositories
-
Do Transformers Parse while Predicting the Masked Word? 14 Mar 2023 · 0 repositories · arXiv:2303.08117
-
Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family 14 Mar 2023 · 2 repositories · arXiv:2303.07992
-
Features matching using natural language processing 14 Mar 2023 · 0 repositories · arXiv:2303.12804
-
Finding the Needle in a Haystack: Unsupervised Rationale Extraction from Long Text Classifiers 14 Mar 2023 · 0 repositories · arXiv:2303.07991
-
FPTN: Fast Pure Transformer Network for Traffic Flow Forecasting 14 Mar 2023 · 0 repositories · arXiv:2303.07685
-
GeoSpark: Sparking up Point Cloud Segmentation with Geometry Clue 14 Mar 2023 · 0 repositories · arXiv:2303.08274
-
Graph Transformer GANs for Graph-Constrained House Generation 14 Mar 2023 · 0 repositories · arXiv:2303.08225
-
I3D: Transformer architectures with input-dependent dynamic depth for speech recognition 14 Mar 2023 · 1 repository · arXiv:2303.07624
-
Implant Global and Local Hierarchy Information to Sequence based Code Representation Models 14 Mar 2023 · 1 repository · arXiv:2303.07826
-
MEDBERT.de: A Comprehensive German BERT Model for the Medical Domain 14 Mar 2023 · 0 repositories · arXiv:2303.08179
-
Neuro-symbolic Commonsense Social Reasoning 14 Mar 2023 · 3 repositories · arXiv:2303.08264
-
Precise Facial Landmark Detection by Reference Heatmap Transformer 14 Mar 2023 · 0 repositories · arXiv:2303.07840
-
Quaternion Orthogonal Transformer for Facial Expression Recognition in the Wild 14 Mar 2023 · 1 repository · arXiv:2303.07831
-
RE-MOVE: An Adaptive Policy Design for Robotic Navigation Tasks in Dynamic Environments via Language-Based Feedback 14 Mar 2023 · 0 repositories · arXiv:2303.07622