Methods › General › Stochastic Optimization › Adam › Papers, page 61
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 61 of 244: papers 6,001 to 6,100 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Inpainting the Gaps: A Novel Framework for Evaluating Explanation Methods in Vision Transformers 17 Jun 2024 · 0 repositories · arXiv:2406.11534
-
Investigating Annotator Bias in Large Language Models for Hate Speech Detection 17 Jun 2024 · 3 repositories · arXiv:2406.11109
-
Iterative Length-Regularized Direct Preference Optimization: A Case Study on Improving 7B Language Models to GPT-4 Level 17 Jun 2024 · 0 repositories · arXiv:2406.11817
-
Iterative Utility Judgment Framework via LLMs Inspired by Relevance in Philosophy 17 Jun 2024 · 0 repositories · arXiv:2406.11290
-
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.15484
-
Large Language Models and Knowledge Graphs for Astronomical Entity Disambiguation 17 Jun 2024 · 0 repositories · arXiv:2406.11400
-
MetaGPT: Merging Large Language Models Using Model Exclusive Task Arithmetic 17 Jun 2024 · 0 repositories · arXiv:2406.11385
-
Prior Normality Prompt Transformer for Multi-class Industrial Image Anomaly Detection 17 Jun 2024 · 0 repositories · arXiv:2406.11507
-
Promises, Outlooks and Challenges of Diffusion Language Modeling 17 Jun 2024 · 0 repositories · arXiv:2406.11473
-
R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11681
-
Rethinking Spatio-Temporal Transformer for Traffic Prediction:Multi-level Multi-view Augmented Learning Framework 17 Jun 2024 · 0 repositories · arXiv:2406.11921
-
Satyrn: A Platform for Analytics Augmented Generation 17 Jun 2024 · 1 repository · arXiv:2406.12069
-
Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99% 17 Jun 2024 · 1 repository · arXiv:2406.11837Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
SEFraud: Graph-based Self-Explainable Fraud Detection via Interpretative Mask Learning 17 Jun 2024 · 0 repositories · arXiv:2406.11389
-
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment 17 Jun 2024 · 0 repositories · arXiv:2406.11285
-
Should AI Optimize Your Code? A Comparative Study of Classical Optimizing Compilers Versus Current Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.12146
-
Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers 17 Jun 2024 · 0 repositories · arXiv:2406.11274
-
Small Agent Can Also Rock! Empowering Small Language Models as Hallucination Detector 17 Jun 2024 · 1 repository · arXiv:2406.11277
-
SWCF-Net: Similarity-weighted Convolution and Local-global Fusion for Efficient Large-scale Point Cloud Semantic Segmentation 17 Jun 2024 · 1 repository · arXiv:2406.11441
-
Refiner: Restructure Retrieval Content Efficiently to Advance Question-Answering Capabilities 17 Jun 2024 · 1 repository · arXiv:2406.11357Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TRACE the Evidence: Constructing Knowledge-Grounded Reasoning Chains for Retrieval-Augmented Generation 17 Jun 2024 · 2 repositories · arXiv:2406.11460Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Unveiling and Mitigating Bias in Mental Health Analysis with Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.12033
-
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions 17 Jun 2024 · 1 repository · arXiv:2406.12058Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can LLMs Understand the Implication of Emphasized Sentences in Dialogue? 16 Jun 2024 · 1 repository · arXiv:2406.11065
-
Distilling Opinions at Scale: Incremental Opinion Summarization using XL-OPSUMM 16 Jun 2024 · 0 repositories · arXiv:2406.10886
-
Enhancing Supermarket Robot Interaction: A Multi-Level LLM Conversational Interface for Handling Diverse Customer Intents 16 Jun 2024 · 0 repositories · arXiv:2406.11047
-
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning 16 Jun 2024 · 0 repositories · arXiv:2406.10834
-
Generating Tables from the Parametric Knowledge of Language Models 16 Jun 2024 · 1 repository · arXiv:2406.10922
-
Geometric Distortion Guided Transformer for Omnidirectional Image Super-Resolution 16 Jun 2024 · 0 repositories · arXiv:2406.10869
-
Grading Massive Open Online Courses Using Large Language Models 16 Jun 2024 · 0 repositories · arXiv:2406.11102
-
KGPA: Robustness Evaluation for Large Language Models via Cross-Domain Knowledge Graphs 16 Jun 2024 · 1 repository · arXiv:2406.10802
-
Large Language Models for Automatic Milestone Detection in Group Discussions 16 Jun 2024 · 0 repositories · arXiv:2406.10842
-
Predicting the Understandability of Computational Notebooks through Code Metrics Analysis 16 Jun 2024 · 1 repository · arXiv:2406.10989
-
Scaling Synthetic Logical Reasoning Datasets with Context-Sensitive Declarative Grammars 16 Jun 2024 · 1 repository · arXiv:2406.11035
-
ShareLoRA: Parameter Efficient and Robust Large Language Model Fine-tuning via Shared Low-Rank Adaptation 16 Jun 2024 · 1 repository · arXiv:2406.10785Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
SPEAR: Receiver-to-Receiver Acoustic Neural Warping Field 16 Jun 2024 · 1 repository · arXiv:2406.11006
-
ViD-GPT: Introducing GPT-style Autoregressive Generation in Video Diffusion Models 16 Jun 2024 · 1 repository · arXiv:2406.10981
-
WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences 16 Jun 2024 · 0 repositories · arXiv:2406.11069
-
A Comprehensive Survey of Foundation Models in Medicine 15 Jun 2024 · 0 repositories · arXiv:2406.10729
-
Beyond Raw Videos: Understanding Edited Videos with Large Multimodal Model 15 Jun 2024 · 1 repository · arXiv:2406.10484
-
BlockPruner: Fine-grained Pruning for Large Language Models 15 Jun 2024 · 1 repository · arXiv:2406.10594Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Automating Pharmacovigilance Evidence Generation: Using Large Language Models to Produce Context-Aware SQL 15 Jun 2024 · 0 repositories · arXiv:2406.10690
-
Focused State Recognition Using EEG with Eye Movement-Assisted Annotation 15 Jun 2024 · 0 repositories · arXiv:2407.09508
-
MINT: a Multi-modal Image and Narrative Text Dubbing Dataset for Foley Audio Content Planning and Generation 15 Jun 2024 · 1 repository · arXiv:2406.10591
-
Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition 15 Jun 2024 · 0 repositories · arXiv:2406.10712
-
Self-Supervised Vision Transformer for Enhanced Virtual Clothes Try-On 15 Jun 2024 · 0 repositories · arXiv:2406.10539
-
SparseCL: Sparse Contrastive Learning for Contradiction Retrieval 15 Jun 2024 · 0 repositories · arXiv:2406.10746
-
The Implicit Bias of Adam on Separable Data 15 Jun 2024 · 0 repositories · arXiv:2406.10650
-
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation 15 Jun 2024 · 0 repositories · arXiv:2406.10561
-
Bag of Lies: Robustness in Continuous Pre-training BERT 14 Jun 2024 · 0 repositories · arXiv:2406.09967
-
BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages 14 Jun 2024 · 1 repository · arXiv:2406.09948Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Bootstrapping Language Models with DPO Implicit Rewards 14 Jun 2024 · 1 repository · arXiv:2406.09760Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Contrastive Imitation Learning for Language-guided Multi-Task Robotic Manipulation 14 Jun 2024 · 0 repositories · arXiv:2406.09738
-
Datasets for Multilingual Answer Sentence Selection 14 Jun 2024 · 0 repositories · arXiv:2406.10172
-
Domain-Specific Shorthand for Generation Based on Context-Free Grammar 14 Jun 2024 · 0 repositories · arXiv:2406.10442
-
Evaluating LLM-driven User-Intent Formalization for Verification-Aware Languages 14 Jun 2024 · 0 repositories · arXiv:2406.09757
-
Evolving Self-Assembling Neural Networks: From Spontaneous Activity to Experience-Dependent Learning 14 Jun 2024 · 1 repository · arXiv:2406.09787
-
Exploring the Correlation between Human and Machine Evaluation of Simultaneous Speech Translation 14 Jun 2024 · 0 repositories · arXiv:2406.10091
-
GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding 14 Jun 2024 · 0 repositories · arXiv:2406.09781
-
Memory-Efficient Optimization with Factorized Hamiltonian Descent 14 Jun 2024 · 0 repositories · arXiv:2406.09958
-
HIRO: Hierarchical Information Retrieval Optimization 14 Jun 2024 · 1 repository · arXiv:2406.09979
-
Know the Unknown: An Uncertainty-Sensitive Method for LLM Instruction Tuning 14 Jun 2024 · 1 repository · arXiv:2406.10099Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Neural Concept Binder 14 Jun 2024 · 1 repository · arXiv:2406.09949Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Precipitation Nowcasting Using Physics Informed Discriminator Generative Models 14 Jun 2024 · 0 repositories · arXiv:2406.10108
-
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models 14 Jun 2024 · 1 repository · arXiv:2406.10130
-
Towards Efficient Pareto Set Approximation via Mixture of Experts Based Model Fusion 14 Jun 2024 · 1 repository · arXiv:2406.09770
-
When Will Gradient Regularization Be Harmful? 14 Jun 2024 · 1 repository · arXiv:2406.09723Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A More Practical Approach to Machine Unlearning 13 Jun 2024 · 0 repositories · arXiv:2406.09391
-
Alleviating Distortion in Image Generation via Multi-Resolution Diffusion Models and Time-Dependent Layer Normalization 13 Jun 2024 · 1 repository · arXiv:2406.09416Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples) · 6 pointer-only (licence)
-
An Image is Worth More Than 16x16 Patches: Exploring Transformers on Individual Pixels 13 Jun 2024 · 0 repositories · arXiv:2406.09415
-
Analyzing Gender Polarity in Short Social Media Texts with BERT: The Role of Emojis and Emoticons 13 Jun 2024 · 0 repositories · arXiv:2406.09573
-
BPE-knockout: Pruning Pre-existing BPE Tokenisers with Backwards-compatible Morphological Semi-supervision 13 Jun 2024 · 1 repository
-
Cognitively Inspired Energy-Based World Models 13 Jun 2024 · 0 repositories · arXiv:2406.08862
-
Cross-Modal Learning for Anomaly Detection in Complex Industrial Process: Methodology and Benchmark 13 Jun 2024 · 1 repository · arXiv:2406.09016Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Deep Transformer Network for Monocular Pose Estimation of Ship-Based UAV 13 Jun 2024 · 1 repository · arXiv:2406.09260
-
DenoiseRep: Denoising Model for Representation Learning 13 Jun 2024 · 1 repository · arXiv:2406.08773Syntology official (archive's flag): 17 ran · 17 ran (of which 2 constructed an object rather than computing a result; 17 with no instrument failure: 2 honoured, 1 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 6 pointer-only (licence)
-
Efficient Multi-View Fusion and Flexible Adaptation to View Missing in Cardiovascular System Signals 13 Jun 2024 · 1 repository · arXiv:2406.08930
-
Fredformer: Frequency Debiased Transformer for Time Series Forecasting 13 Jun 2024 · 1 repository · arXiv:2406.09009Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Generative Inverse Design of Crystal Structures via Diffusion Models with Transformers 13 Jun 2024 · 0 repositories · arXiv:2406.09263
-
Hybrid Spatial-spectral Neural Network for Hyperspectral Image Denoising 13 Jun 2024 · 0 repositories · arXiv:2406.08782
-
JailbreakEval: An Integrated Toolkit for Evaluating Jailbreak Attempts Against Large Language Models 13 Jun 2024 · 1 repository · arXiv:2406.09321Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination 13 Jun 2024 · 0 repositories · arXiv:2406.08818
-
Low-Overhead Channel Estimation via 3D Extrapolation for TDD mmWave Massive MIMO Systems Under High-Mobility Scenarios 13 Jun 2024 · 0 repositories · arXiv:2406.08887
-
MGRQ: Post-Training Quantization For Vision Transformer With Mixed Granularity Reconstruction 13 Jun 2024 · 0 repositories · arXiv:2406.09229
-
OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning 13 Jun 2024 · 0 repositories · arXiv:2406.08858
-
Optimizing Large Model Training through Overlapped Activation Recomputation 13 Jun 2024 · 0 repositories · arXiv:2406.08756
-
PC-LoRA: Low-Rank Adaptation for Progressive Model Compression with Knowledge Distillation 13 Jun 2024 · 0 repositories · arXiv:2406.09117
-
QMamba: On First Exploration of Vision Mamba for Image Quality Assessment 13 Jun 2024 · 1 repository · arXiv:2406.09546
-
ReadCtrl: Personalizing text generation with readability-controlled instruction learning 13 Jun 2024 · 0 repositories · arXiv:2406.09205
-
Research on Early Warning Model of Cardiovascular Disease Based on Computer Deep Learning 13 Jun 2024 · 0 repositories · arXiv:2406.08864
-
Scoreformer: A Surrogate Model For Large-Scale Prediction of Docking Scores 13 Jun 2024 · 0 repositories · arXiv:2406.09346
-
Separations in the Representational Capabilities of Transformers and Recurrent Architectures 13 Jun 2024 · 0 repositories · arXiv:2406.09347
-
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models 13 Jun 2024 · 0 repositories · arXiv:2406.09519
-
Transformers meet Neural Algorithmic Reasoners 13 Jun 2024 · 0 repositories · arXiv:2406.09308
-
Vertical LoRA: Dense Expectation-Maximization Interpretation of Transformers 13 Jun 2024 · 1 repository · arXiv:2406.09315
-
Why Warmup the Learning Rate? Underlying Mechanisms and Improvements 13 Jun 2024 · 0 repositories · arXiv:2406.09405Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
A Robust Pipeline for Classification and Detection of Bleeding Frames in Wireless Capsule Endoscopy using Swin Transformer and RT-DETR 12 Jun 2024 · 0 repositories · arXiv:2406.08046
-
A Sociotechnical Lens for Evaluating Computer Vision Models: A Case Study on Detecting and Reasoning about Gender and Emotion 12 Jun 2024 · 0 repositories · arXiv:2406.08222
-
Ad Auctions for LLMs via Retrieval Augmented Generation 12 Jun 2024 · 0 repositories · arXiv:2406.09459
-
AdaNCA: Neural Cellular Automata As Adaptors For More Robust Vision Transformer 12 Jun 2024 · 0 repositories · arXiv:2406.08298