Methods › General › Stochastic Optimization › Adam › Papers, page 129
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 129 of 244: papers 12,801 to 12,900 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CWTM: Leveraging Contextualized Word Embeddings from BERT for Neural Topic Modeling 16 May 2023 · 1 repository · arXiv:2305.09329
-
Blind Image Quality Assessment via Transformer Predicted Error Map and Perceptual Quality Token 16 May 2023 · 1 repository · arXiv:2305.09353
-
CB-HVTNet: A channel-boosted hybrid vision transformer network for lymphocyte assessment in histopathological images 16 May 2023 · 0 repositories · arXiv:2305.09211
-
Cooperation Is All You Need 16 May 2023 · 0 repositories · arXiv:2305.10449
-
Exploring the Impact of Layer Normalization for Zero-shot Neural Machine Translation 16 May 2023 · 0 repositories · arXiv:2305.09312
-
Generative Table Pre-training Empowers Models for Tabular Prediction 16 May 2023 · 1 repository · arXiv:2305.09696
-
GIFT: Graph-Induced Fine-Tuning for Multi-Party Conversation Understanding 16 May 2023 · 1 repository · arXiv:2305.09360
-
Life of PII -- A PII Obfuscation Transformer 16 May 2023 · 0 repositories · arXiv:2305.09550
-
Measuring Dimensions of Self-Presentation in Twitter Bios and their Links to Misinformation Sharing 16 May 2023 · 1 repository · arXiv:2305.09548
-
PanelNet: Understanding 360 Indoor Environment via Panel Representation 16 May 2023 · 0 repositories · arXiv:2305.09078
-
Weight-Inherited Distillation for Task-Agnostic BERT Compression 16 May 2023 · 1 repository · arXiv:2305.09098
-
AdamR at SemEval-2023 Task 10: Solving the Class Imbalance Problem in Sexism Detection with Ensemble Learning 15 May 2023 · 0 repositories · arXiv:2305.08636
-
C-Eval: A Multi-Level Multi-Discipline Chinese Evaluation Suite for Foundation Models 15 May 2023 · 1 repository · arXiv:2305.08322Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Continual Multimodal Knowledge Graph Construction 15 May 2023 · 1 repository · arXiv:2305.08698
-
Coreference-aware Double-channel Attention Network for Multi-party Dialogue Reading Comprehension 15 May 2023 · 1 repository · arXiv:2305.08348
-
GeoMAE: Masked Geometric Target Prediction for Self-supervised Point Cloud Pre-Training 15 May 2023 · 1 repository · arXiv:2305.08808
-
Keras GPT Copilot: Integrating the Power of Large Language Models in Deep Learning Model Development 15 May 2023 · 1 repository
-
Knowledge Rumination for Pre-trained Language Models 15 May 2023 · 1 repository · arXiv:2305.08732
-
LoViT: Long Video Transformer for Surgical Phase Recognition 15 May 2023 · 1 repository · arXiv:2305.08989
-
Masked Collaborative Contrast for Weakly Supervised Semantic Segmentation 15 May 2023 · 1 repository · arXiv:2305.08491
-
MaxViT-UNet: Multi-Axis Attention for Medical Image Segmentation 15 May 2023 · 2 repositories · arXiv:2305.08396
-
Physics Informed Token Transformer for Solving Partial Differential Equations 15 May 2023 · 1 repository · arXiv:2305.08757
-
Private Training Set Inspection in MLaaS 15 May 2023 · 0 repositories · arXiv:2305.09058
-
RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs 15 May 2023 · 1 repository · arXiv:2305.08844Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 6 pointer-only (licence)
-
Schema-adaptable Knowledge Graph Construction 15 May 2023 · 1 repository · arXiv:2305.08703
-
Sensitivity and Robustness of Large Language Models to Prompt Template in Japanese Text Classification Tasks 15 May 2023 · 0 repositories · arXiv:2305.08714
-
Similarity-weighted Construction of Contextualized Commonsense Knowledge Graphs for Knowledge-intense Argumentation Tasks 15 May 2023 · 1 repository · arXiv:2305.08495
-
Small Models are Valuable Plug-ins for Large Language Models 15 May 2023 · 1 repository · arXiv:2305.08848Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Text Classification via Large Language Models 15 May 2023 · 1 repository · arXiv:2305.08377Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Text2Gender: A Deep Learning Architecture for Analysis of Blogger's Age and Gender 15 May 2023 · 0 repositories · arXiv:2305.08633
-
Assessing the potential of LLM-assisted annotation for corpus-based pragmatics and discourse analysis: The case of apology 15 May 2023 · 0 repositories · arXiv:2305.08339
-
MatSci-NLP: Evaluating Scientific Language Models on Materials Science Language Tasks Using Text-to-Schema Modeling 14 May 2023 · 1 repository · arXiv:2305.08264Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction 14 May 2023 · 2 repositories · arXiv:2305.08144Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 17 harvested samples)
-
Self-supervised Neural Factor Analysis for Disentangling Utterance-level Speech Representations 14 May 2023 · 0 repositories · arXiv:2305.08099
-
TSGN: Temporal Scene Graph Neural Networks with Projected Vectorized Representation for Multi-Agent Motion Prediction 14 May 2023 · 0 repositories · arXiv:2305.08190
-
Bridging History with AI A Comparative Evaluation of GPT 3.5, GPT4, and GoogleBARD in Predictive Accuracy and Fact Checking 13 May 2023 · 0 repositories · arXiv:2305.07868
-
CEMFormer: Learning to Predict Driver Intentions from In-Cabin and External Cameras via Spatial-Temporal Transformers 13 May 2023 · 0 repositories · arXiv:2305.07840
-
GPT-Sentinel: Distinguishing Human and ChatGPT Generated Content 13 May 2023 · 2 repositories · arXiv:2305.07969
-
GSB: Group Superposition Binarization for Vision Transformer with Limited Training Samples 13 May 2023 · 1 repository · arXiv:2305.07931
-
The Machine Psychology of Cooperation: Can GPT models operationalise prompts for altruism, cooperation, competitiveness and selfishness in economic games? 13 May 2023 · 2 repositories · arXiv:2305.07970Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Meta-Polyp: a baseline for efficient Polyp segmentation 13 May 2023 · 2 repositories · arXiv:2305.07848
-
PESTS: Persian_English Cross Lingual Corpus for Semantic Textual Similarity 13 May 2023 · 0 repositories · arXiv:2305.07893
-
Stackelberg Decision Transformer for Asynchronous Action Coordination in Multi-Agent Systems 13 May 2023 · 0 repositories · arXiv:2305.07856
-
AGFormer: Efficient Graph Representation with Anchor-Graph Transformer 12 May 2023 · 0 repositories · arXiv:2305.07521
-
ArtGPT-4: Towards Artistic-understanding Large Vision-Language Models with Enhanced Adapter 12 May 2023 · 1 repository · arXiv:2305.07490
-
CLIP-Count: Towards Text-Guided Zero-Shot Object Counting 12 May 2023 · 1 repository · arXiv:2305.07304Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Improving Small Language Models on PubMedQA via Generative Data Augmentation 12 May 2023 · 0 repositories · arXiv:2305.07804
-
Learning to Reason over Scene Graphs: A Case Study of Finetuning GPT-2 into a Robot Language Model for Grounded Task Planning 12 May 2023 · 0 repositories · arXiv:2305.07716
-
MoMo: Momentum Models for Adaptive Learning Rates 12 May 2023 · 1 repository · arXiv:2305.07583Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Hausdorff Distance Matching with Adaptive Query Denoising for Rotated Detection Transformer 12 May 2023 · 1 repository · arXiv:2305.07598
-
TinyStories: How Small Can Language Models Be and Still Speak Coherent English? 12 May 2023 · 8 repositories · arXiv:2305.07759Syntology 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples)
-
When Giant Language Brains Just Aren't Enough! Domain Pizzazz with Knowledge Sparkle Dust 12 May 2023 · 0 repositories · arXiv:2305.07230
-
COCKATIEL: COntinuous Concept ranKed ATtribution with Interpretable ELements for explaining neural net classifiers on NLP tasks 11 May 2023 · 1 repository · arXiv:2305.06754
-
Generative Pre-trained Transformer: A Comprehensive Review on Enabling Technologies, Potential Applications, Emerging Challenges, and Future Directions 11 May 2023 · 0 repositories · arXiv:2305.10435
-
InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning 11 May 2023 · 4 repositories · arXiv:2305.06500
-
Spear Phishing With Large Language Models 11 May 2023 · 0 repositories · arXiv:2305.06972
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm 11 May 2023 · 0 repositories · arXiv:2305.06657
-
Overinformative Question Answering by Humans and Machines 11 May 2023 · 0 repositories · arXiv:2305.07151
-
PVT-SSD: Single-Stage 3D Object Detector with Point-Voxel Transformer 11 May 2023 · 1 repository · arXiv:2305.06621
-
Recommendation as Instruction Following: A Large Language Model Empowered Recommendation Approach 11 May 2023 · 0 repositories · arXiv:2305.07001
-
Salient Mask-Guided Vision Transformer for Fine-Grained Classification 11 May 2023 · 1 repository · arXiv:2305.07102
-
The ConceptARC Benchmark: Evaluating Understanding and Generalization in the ARC Domain 11 May 2023 · 1 repository · arXiv:2305.07141
-
Transformers for CT Reconstruction From Monoplanar and Biplanar Radiographs 11 May 2023 · 0 repositories · arXiv:2305.06965
-
Undercover Deepfakes: Detecting Fake Segments in Videos 11 May 2023 · 2 repositories · arXiv:2305.06564
-
A Method to Automate the Discharge Summary Hospital Course for Neurology Patients 10 May 2023 · 0 repositories · arXiv:2305.06416
-
Adapter-TST: A Parameter Efficient Method for Multiple-Attribute Text Style Transfer 10 May 2023 · 0 repositories · arXiv:2305.05945
-
Alternating Gradient Descent and Mixture-of-Experts for Integrated Multimodal Perception 10 May 2023 · 0 repositories · arXiv:2305.06324
-
Are ChatGPT and GPT-4 General-Purpose Solvers for Financial Text Analytics? A Study on Several Typical Tasks 10 May 2023 · 0 repositories · arXiv:2305.05862
-
Autonomous GIS: the next-generation AI-powered GIS 10 May 2023 · 1 repository · arXiv:2305.06453
-
BIOT: Cross-data Biosignal Learning in the Wild 10 May 2023 · 1 repository · arXiv:2305.10351
-
Bits of Grass: Does GPT already know how to write like Whitman? 10 May 2023 · 0 repositories · arXiv:2305.11064
-
Bot or Human? Detecting ChatGPT Imposters with A Single Question 10 May 2023 · 1 repository · arXiv:2305.06424
-
Davinci the Dualist: the mind-body divide in large language models and in human learners 10 May 2023 · 0 repositories · arXiv:2305.07667
-
Enriching language models with graph-based context information to better understand textual data 10 May 2023 · 1 repository · arXiv:2305.11070
-
Generating medically-accurate summaries of patient-provider dialogue: A multi-stage approach using large language models 10 May 2023 · 0 repositories · arXiv:2305.05982
-
LACoS-BLOOM: Low-rank Adaptation with Contrastive objective on 8 bits Siamese-BLOOM 10 May 2023 · 0 repositories · arXiv:2305.06404
-
Benchmarking large language models for biomedical natural language processing applications and recommendations 10 May 2023 · 1 repository · arXiv:2305.16326
-
MMoT: Mixture-of-Modality-Tokens Transformer for Composed Multimodal Conditional Image Synthesis 10 May 2023 · 0 repositories · arXiv:2305.05992
-
Multi-Path Transformer is Better: A Case Study on Neural Machine Translation 10 May 2023 · 0 repositories · arXiv:2305.05948
-
Summarizing, Simplifying, and Synthesizing Medical Evidence Using GPT-3 (with Varying Success) 10 May 2023 · 1 repository · arXiv:2305.06299
-
VTPNet for 3D deep learning on point cloud 10 May 2023 · 0 repositories · arXiv:2305.06115
-
A Review of Vision-Language Models and their Performance on the Hateful Memes Challenge 9 May 2023 · 1 repository · arXiv:2305.06159
-
Alleviating Over-smoothing for Unsupervised Sentence Representation 9 May 2023 · 1 repository · arXiv:2305.06154Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples)
-
Attack Named Entity Recognition by Entity Boundary Interference 9 May 2023 · 0 repositories · arXiv:2305.05253
-
Tomography of Quantum States from Structured Measurements via quantum-aware transformer 9 May 2023 · 0 repositories · arXiv:2305.05433
-
AudioSlots: A slot-centric generative model for audio separation 9 May 2023 · 0 repositories · arXiv:2305.05591
-
CodeIE: Large Code Generation Models are Better Few-Shot Information Extractors 9 May 2023 · 1 repository · arXiv:2305.05711Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Detection of depression on social networks using transformers and ensembles 9 May 2023 · 1 repository · arXiv:2305.05325
-
FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance 9 May 2023 · 1 repository · arXiv:2305.05176Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
GPT in Game Theory Experiments 9 May 2023 · 0 repositories · arXiv:2305.05516
-
GPT-NAS: Evolutionary Neural Architecture Search with the Generative Pre-Trained Model 9 May 2023 · 0 repositories · arXiv:2305.05351
-
Hybrid Transformer and CNN Attention Network for Stereo Image Super-resolution 9 May 2023 · 0 repositories · arXiv:2305.05177
-
InternGPT: Solving Vision-Centric Tasks by Interacting with ChatGPT Beyond Language 9 May 2023 · 2 repositories · arXiv:2305.05662Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Effects of sub-word segmentation on performance of transformer language models 9 May 2023 · 0 repositories · arXiv:2305.05480
-
Simplicial Hopfield networks 9 May 2023 · 1 repository · arXiv:2305.05179
-
StrAE: Autoencoding for Pre-Trained Embeddings using Explicit Structure 9 May 2023 · 0 repositories · arXiv:2305.05588
-
Towards an Automatic Optimisation Model Generator Assisted with Generative Pre-trained Transformer 9 May 2023 · 0 repositories · arXiv:2305.05811
-
Towards Building the Federated GPT: Federated Instruction Tuning 9 May 2023 · 1 repository · arXiv:2305.05644Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
UAdam: Unified Adam-Type Algorithmic Framework for Non-Convex Stochastic Optimization 9 May 2023 · 0 repositories · arXiv:2305.05675
-
Vision-Language Models in Remote Sensing: Current Progress and Future Trends 9 May 2023 · 3 repositories · arXiv:2305.05726