Methods › General › Stochastic Optimization › Adam › Papers, page 111
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 111 of 244: papers 11,001 to 11,100 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Generative AI 13 Sep 2023 · 0 repositories · arXiv:2309.07930
-
In-Contextual Gender Bias Suppression for Large Language Models 13 Sep 2023 · 1 repository · arXiv:2309.07251
-
Large Language Models Can Infer Psychological Dispositions of Social Media Users 13 Sep 2023 · 0 repositories · arXiv:2309.08631
-
RAIN: Your Language Models Can Align Themselves without Finetuning 13 Sep 2023 · 1 repository · arXiv:2309.07124Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SafetyBench: Evaluating the Safety of Large Language Models 13 Sep 2023 · 1 repository · arXiv:2309.07045Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
ShaDocFormer: A Shadow-Attentive Threshold Detector With Cascaded Fusion Refiner for Document Shadow Removal 13 Sep 2023 · 1 repository · arXiv:2309.06670
-
Sudden Drops in the Loss: Syntax Acquisition, Phase Transitions, and Simplicity Bias in MLMs 13 Sep 2023 · 1 repository · arXiv:2309.07311Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Traveling Words: A Geometric Interpretation of Transformers 13 Sep 2023 · 1 repository · arXiv:2309.07315
-
Balanced and Explainable Social Media Analysis for Public Health with Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05951
-
Neural Network Layer Matrix Decomposition reveals Latent Manifold Encoding and Memory Capacity 12 Sep 2023 · 0 repositories · arXiv:2309.05968
-
Circuit Breaking: Removing Model Behaviors with Targeted Ablation 12 Sep 2023 · 1 repository · arXiv:2309.05973Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
BHASA: A Holistic Southeast Asian Linguistic and Cultural Evaluation Suite for Large Language Models 12 Sep 2023 · 3 repositories · arXiv:2309.06085
-
Characterizing Latent Perspectives of Media Houses Towards Public Figures 12 Sep 2023 · 0 repositories · arXiv:2309.06112
-
Long-term drought prediction using deep neural networks based on geospatial weather data 12 Sep 2023 · 1 repository · arXiv:2309.06212
-
A 3M-Hybrid Model for the Restoration of Unique Giant Murals: A Case Study on the Murals of Yongle Palace 12 Sep 2023 · 0 repositories · arXiv:2309.06194
-
ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning 12 Sep 2023 · 1 repository · arXiv:2309.05915Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
PRESTI: Predicting Repayment Effort of Self-Admitted Technical Debt Using Textual Information 12 Sep 2023 · 0 repositories · arXiv:2309.06020
-
Comparing Llama-2 and GPT-3 LLMs for HPC kernels generation 12 Sep 2023 · 0 repositories · arXiv:2309.07103
-
ELRA: Exponential learning rate adaption gradient descent optimization method 12 Sep 2023 · 0 repositories · arXiv:2309.06274
-
Exploring Large Language Models for Ontology Alignment 12 Sep 2023 · 1 repository · arXiv:2309.07172
-
Exploring the Benefits of Differentially Private Pre-training and Parameter-Efficient Fine-tuning for Table Transformers 12 Sep 2023 · 1 repository · arXiv:2309.06526
-
Feature Aggregation Network for Building Extraction from High-resolution Remote Sensing Images 12 Sep 2023 · 0 repositories · arXiv:2309.06017
-
FLDNet: A Foreground-Aware Network for Polyp Segmentation Leveraging Long-Distance Dependencies 12 Sep 2023 · 0 repositories · arXiv:2309.05987
-
Hierarchical Multi-Task Learning Framework for Session-based Recommendations 12 Sep 2023 · 0 repositories · arXiv:2309.06533
-
Breaking through the learning plateaus of in-context learning in Transformer 12 Sep 2023 · 0 repositories · arXiv:2309.06054
-
IBAFormer: Intra-batch Attention Transformer for Domain Generalized Semantic Segmentation 12 Sep 2023 · 0 repositories · arXiv:2309.06282
-
Jersey Number Recognition using Keyframe Identification from Low-Resolution Broadcast Videos 12 Sep 2023 · 0 repositories · arXiv:2309.06285
-
Leveraging Large Language Models and Weak Supervision for Social Media data annotation: an evaluation using COVID-19 self-reported vaccination tweets 12 Sep 2023 · 0 repositories · arXiv:2309.06503
-
Overview of Memotion 3: Sentiment and Emotion Analysis of Codemixed Hinglish Memes 12 Sep 2023 · 0 repositories · arXiv:2309.06517
-
Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing 12 Sep 2023 · 0 repositories · arXiv:2309.05898
-
The Moral Machine Experiment on Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05958
-
Unveiling the potential of large language models in generating semantic and cross-language clones 12 Sep 2023 · 0 repositories · arXiv:2309.06424
-
An Empirical Study of NetOps Capability of Pre-Trained Large Language Models 11 Sep 2023 · 0 repositories · arXiv:2309.05557
-
Applying BioBERT to Extract Germline Gene-Disease Associations for Building a Knowledge Graph from the Biomedical Literature 11 Sep 2023 · 1 repository · arXiv:2309.13061
-
Black-Box Analysis: GPTs Across Time in Legal Textual Entailment Task 11 Sep 2023 · 0 repositories · arXiv:2309.05501
-
Circle Feature Graphormer: Can Circle Features Stimulate Graph Transformer? 11 Sep 2023 · 1 repository · arXiv:2309.06574
-
Toward a Deeper Understanding: RetNet Viewed through Convolution 11 Sep 2023 · 1 repository · arXiv:2309.05375
-
CrisisTransformers: Pre-trained language models and sentence encoders for crisis-related social media texts 11 Sep 2023 · 0 repositories · arXiv:2309.05494
-
Detecting Natural Language Biases with Prompt-based Learning 11 Sep 2023 · 0 repositories · arXiv:2309.05227
-
HAT: Hybrid Attention Transformer for Image Restoration 11 Sep 2023 · 2 repositories · arXiv:2309.05239Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Large Language Model for Science: A Study on P vs. NP 11 Sep 2023 · 1 repository · arXiv:2309.05689
-
Long-Range Transformer Architectures for Document Understanding 11 Sep 2023 · 1 repository · arXiv:2309.05503
-
Memory Injections: Correcting Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models 11 Sep 2023 · 1 repository · arXiv:2309.05605Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
SparseSwin: Swin Transformer with Sparse Transformer Block 11 Sep 2023 · 1 repository · arXiv:2309.05224
-
Zero-shot Learning with Minimum Instruction to Extract Social Determinants and Family History from Clinical Notes using GPT Model 11 Sep 2023 · 0 repositories · arXiv:2309.05475
-
Implementing Learning Principles with a Personal AI Tutor: A Case Study 10 Sep 2023 · 0 repositories · arXiv:2309.13060
-
Learning Personalized User Preference from Cold Start in Multi-turn Conversations 10 Sep 2023 · 0 repositories · arXiv:2309.05127
-
Neural-Hidden-CRF: A Robust Weakly-Supervised Sequence Labeler 10 Sep 2023 · 1 repository · arXiv:2309.05086
-
RGAT: A Deeper Look into Syntactic Dependency Information for Coreference Resolution 10 Sep 2023 · 0 repositories · arXiv:2309.04977
-
Unified Contrastive Fusion Transformer for Multimodal Human Action Recognition 10 Sep 2023 · 0 repositories · arXiv:2309.05032
-
DeNoising-MOT: Towards Multiple Object Tracking with Severe Occlusions 9 Sep 2023 · 0 repositories · arXiv:2309.04682
-
Efficient Finetuning Large Language Models For Vietnamese Chatbot 9 Sep 2023 · 0 repositories · arXiv:2309.04646
-
Few-Shot Medical Image Segmentation via a Region-enhanced Prototypical Transformer 9 Sep 2023 · 1 repository · arXiv:2309.04825
-
How to Evaluate Semantic Communications for Images with ViTScore Metric? 9 Sep 2023 · 0 repositories · arXiv:2309.04891
-
Latent Spatiotemporal Adaptation for Generalized Face Forgery Video Detection 9 Sep 2023 · 0 repositories · arXiv:2309.04795
-
Transformer-Based Deep Learning Detector for Dual-Mode Index Modulation 3D-OFDM 9 Sep 2023 · 0 repositories · arXiv:2309.04764
-
Can NLP Models 'Identify', 'Distinguish', and 'Justify' Questions that Don't have a Definitive Answer? 8 Sep 2023 · 0 repositories · arXiv:2309.04635
-
CNN Injected Transformer for Image Exposure Correction 8 Sep 2023 · 1 repository · arXiv:2309.04366
-
Context-Aware Prompt Tuning for Vision-Language Model with Dual-Alignment 8 Sep 2023 · 0 repositories · arXiv:2309.04158
-
Curve Your Attention: Mixed-Curvature Transformers for Graph Representation Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04082
-
Encoding Multi-Domain Scientific Papers by Ensembling Multiple CLS Tokens 8 Sep 2023 · 1 repository · arXiv:2309.04333Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
FIMO: A Challenge Formal Dataset for Automated Theorem Proving 8 Sep 2023 · 1 repository · arXiv:2309.04295
-
From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting 8 Sep 2023 · 0 repositories · arXiv:2309.04269
-
Fuzzy Fingerprinting Transformer Language-Models for Emotion Recognition in Conversations 8 Sep 2023 · 0 repositories · arXiv:2309.04292
-
Language Prompt for Autonomous Driving 8 Sep 2023 · 1 repository · arXiv:2309.04379Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Leveraging Pretrained Image-text Models for Improving Audio-Visual Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04628
-
NESTLE: a No-Code Tool for Statistical Analysis of Legal Corpus 8 Sep 2023 · 1 repository · arXiv:2309.04146
-
UQ at #SMM4H 2023: ALEX for Public Health Analysis with Social Media 8 Sep 2023 · 1 repository · arXiv:2309.04213
-
Enhancing Pipeline-Based Conversational Agents with Large Language Models 7 Sep 2023 · 0 repositories · arXiv:2309.03748
-
Evaluating ChatGPT as a Recommender System: A Rigorous Approach 7 Sep 2023 · 1 repository · arXiv:2309.03613
-
Supervised Learning and Large Language Model Benchmarks on Mental Health Datasets: Cognitive Distortions and Suicidal Risks in Chinese Social Media 7 Sep 2023 · 2 repositories · arXiv:2309.03564
-
Evaluation of large language models for discovery of gene set function 7 Sep 2023 · 1 repository · arXiv:2309.04019
-
FLM-101B: An Open LLM and How to Train It with $100K Budget 7 Sep 2023 · 0 repositories · arXiv:2309.03852
-
Hybrid of representation learning and reinforcement learning for dynamic and complex robotic motion planning 7 Sep 2023 · 0 repositories · arXiv:2309.03758
-
MS-UNet-v2: Adaptive Denoising Method and Training Strategy for Medical Image Segmentation with Small Training Data 7 Sep 2023 · 0 repositories · arXiv:2309.03686
-
MMSFormer: Multimodal Transformer for Material and Semantic Segmentation 7 Sep 2023 · 1 repository · arXiv:2309.04001
-
ProPainter: Improving Propagation and Transformer for Video Inpainting 7 Sep 2023 · 3 repositories · arXiv:2309.03897Syntology official (archive's flag): 10 ran · 20 ran (of which 10 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 8 where Syntology's instrument failed) · 14 unverified (of 34 harvested samples) · 16 pointer-only (licence)
-
S-Adapter: Generalizing Vision Transformer for Face Anti-Spoofing with Statistical Tokens 7 Sep 2023 · 3 repositories · arXiv:2309.04038
-
Zero-Shot Audio Captioning via Audibility Guidance 7 Sep 2023 · 0 repositories · arXiv:2309.03884
-
Certifying LLM Safety against Adversarial Prompting 6 Sep 2023 · 1 repository · arXiv:2309.02705
-
Character Queries: A Transformer-based Approach to On-Line Handwritten Character Segmentation 6 Sep 2023 · 1 repository · arXiv:2309.03072
-
GPT Can Solve Mathematical Problems Without a Calculator 6 Sep 2023 · 1 repository · arXiv:2309.03241Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
HAE-RAE Bench: Evaluation of Korean Knowledge in Language Models 6 Sep 2023 · 1 repository · arXiv:2309.02706
-
Knowledge Solver: Teaching LLMs to Search for Domain Knowledge from Knowledge Graphs 6 Sep 2023 · 0 repositories · arXiv:2309.03118
-
Large Language Models for Automated Open-domain Scientific Hypotheses Discovery 6 Sep 2023 · 1 repository · arXiv:2309.02726Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Leave no Place Behind: Improved Geolocation in Humanitarian Documents 6 Sep 2023 · 0 repositories · arXiv:2309.02914
-
Offensive Hebrew Corpus and Detection using BERT 6 Sep 2023 · 1 repository · arXiv:2309.02724
-
Prompt-based Ingredient-Oriented All-in-One Image Restoration 6 Sep 2023 · 1 repository · arXiv:2309.03063Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Self-Supervised Masked Digital Elevation Models Encoding for Low-Resource Downstream Tasks 6 Sep 2023 · 0 repositories · arXiv:2309.03367
-
TFBEST: Dual-Aspect Transformer with Learnable Positional Encoding for Failure Prediction 6 Sep 2023 · 0 repositories · arXiv:2309.02641
-
A survey on efficient vision transformers: algorithms, techniques, and performance benchmarking 5 Sep 2023 · 0 repositories · arXiv:2309.02031
-
Asymmetric Momentum: A Rethinking of Gradient Descent 5 Sep 2023 · 1 repository · arXiv:2309.02130
-
BeeTLe: A Framework for Linear B-Cell Epitope Prediction and Classification 5 Sep 2023 · 1 repository · arXiv:2309.02071
-
CodeApex: A Bilingual Programming Evaluation Benchmark for Large Language Models 5 Sep 2023 · 1 repository · arXiv:2309.01940
-
Data-Juicer: A One-Stop Data Processing System for Large Language Models 5 Sep 2023 · 2 repositories · arXiv:2309.02033
-
Dense Object Grounding in 3D Scenes 5 Sep 2023 · 0 repositories · arXiv:2309.02224
-
Do You Trust ChatGPT? -- Perceived Credibility of Human and AI-Generated Content 5 Sep 2023 · 0 repositories · arXiv:2309.02524
-
Dynamic Brain Transformer with Multi-level Attention for Functional Brain Network Analysis 5 Sep 2023 · 1 repository · arXiv:2309.01941
-
Evaluation Kidney Layer Segmentation on Whole Slide Imaging using Convolutional Neural Networks and Transformers 5 Sep 2023 · 0 repositories · arXiv:2309.02563
-
Exchanging-based Multimodal Fusion with Transformer 5 Sep 2023 · 1 repository · arXiv:2309.02190Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)