Methods › General › Stochastic Optimization › Adam › Papers, page 31
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 31 of 244: papers 3,001 to 3,100 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning 26 Nov 2024 · 1 repository · arXiv:2411.17426
-
Distributed Sign Momentum with Local Steps for Training Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17866
-
ER2Score: LLM-based Explainable and Customizable Metric for Assessing Radiology Reports with Reward-Control Loss 26 Nov 2024 · 0 repositories · arXiv:2411.17301
-
Fairness And Performance In Harmony: Data Debiasing Is All You Need 26 Nov 2024 · 0 repositories · arXiv:2411.17374
-
Geometric Point Attention Transformer for 3D Shape Reassembly 26 Nov 2024 · 0 repositories · arXiv:2411.17788
-
"Give me the code" -- Log Analysis of First-Year CS Students' Interactions With GPT 26 Nov 2024 · 0 repositories · arXiv:2411.17855
-
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning 26 Nov 2024 · 1 repository · arXiv:2411.17100
-
Leveraging Large Language Models and Topic Modeling for Toxicity Classification 26 Nov 2024 · 1 repository · arXiv:2411.17876
-
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation 26 Nov 2024 · 2 repositories · arXiv:2411.17945Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
MAT: Multi-Range Attention Transformer for Efficient Image Super-Resolution 26 Nov 2024 · 1 repository · arXiv:2411.17214Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
MWFormer: Multi-Weather Image Restoration Using Degradation-Aware Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17226
-
On Limitations of LLM as Annotator for Low Resource Languages 26 Nov 2024 · 0 repositories · arXiv:2411.17637
-
ΩSFormer: Dual-Modal Ω-like Super-Resolution Transformer Network for Cross-scale and High-accuracy Terraced Field Vectorization Extraction 26 Nov 2024 · 0 repositories · arXiv:2411.17088
-
Pretrained LLM Adapted with LoRA as a Decision Transformer for Offline RL in Quantitative Trading 26 Nov 2024 · 1 repository · arXiv:2411.17900
-
Push the Limit of Multi-modal Emotion Recognition by Prompting LLMs with Receptive-Field-Aware Attention Weighting 26 Nov 2024 · 0 repositories · arXiv:2411.17674
-
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation 26 Nov 2024 · 0 repositories · arXiv:2411.17061
-
ShowUI: One Vision-Language-Action Model for GUI Visual Agent 26 Nov 2024 · 1 repository · arXiv:2411.17465Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
TED-VITON: Transformer-Empowered Diffusion Models for Virtual Try-On 26 Nov 2024 · 1 repository · arXiv:2411.17017
-
TinyViM: Frequency Decoupling for Tiny Hybrid Vision Mamba 26 Nov 2024 · 1 repository · arXiv:2411.17473
-
What Differentiates Educational Literature? A Multimodal Fusion Approach of Transformers and Computational Linguistics 26 Nov 2024 · 0 repositories · arXiv:2411.17593
-
Adaptive Circuit Behavior and Generalization in Mechanistic Interpretability 25 Nov 2024 · 0 repositories · arXiv:2411.16105
-
Are Transformers Truly Foundational for Robotics? 25 Nov 2024 · 0 repositories · arXiv:2411.16917
-
AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning 25 Nov 2024 · 1 repository · arXiv:2411.16495Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
CATP-LLM: Empowering Large Language Models for Cost-Aware Tool Planning 25 Nov 2024 · 0 repositories · arXiv:2411.16313
-
CMAViT: Integrating Climate, Managment, and Remote Sensing Data for Crop Yield Estimation with Multimodel Vision Transformers 25 Nov 2024 · 0 repositories · arXiv:2411.16989
-
Curvature in the Looking-Glass: Optimal Methods to Exploit Curvature of Expectation in the Loss Landscape 25 Nov 2024 · 0 repositories · arXiv:2411.16914
-
DF-GNN: Dynamic Fusion Framework for Attention Graph Neural Networks on GPUs 25 Nov 2024 · 1 repository · arXiv:2411.16127
-
Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16991
-
Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16797
-
Enhancing Fluorescence Lifetime Parameter Estimation Accuracy with Differential Transformer Based Deep Learning Model Incorporating Pixelwise Instrument Response Function 25 Nov 2024 · 0 repositories · arXiv:2411.16896
-
Fine-Tuning LLMs with Noisy Data for Political Argument Generation and Post Guidance 25 Nov 2024 · 0 repositories · arXiv:2411.16813
-
Human-Calibrated Automated Testing and Validation of Generative Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16391
-
LaB-RAG: Label Boosted Retrieval Augmented Generation for Radiology Report Generation 25 Nov 2024 · 1 repository · arXiv:2411.16523
-
Lion Cub: Minimizing Communication Overhead in Distributed Lion 25 Nov 2024 · 0 repositories · arXiv:2411.16462
-
MarketGPT: Developing a Pre-trained transformer (GPT) for Modeling Financial Time Series 25 Nov 2024 · 1 repository · arXiv:2411.16585Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
NormXLogit: The Head-on-Top Never Lies 25 Nov 2024 · 0 repositories · arXiv:2411.16252
-
Predictive Power of LLMs in Financial Markets 25 Nov 2024 · 0 repositories · arXiv:2411.16569
-
Scaling Spike-driven Transformer with Efficient Spike Firing Approximation Training 25 Nov 2024 · 1 repository · arXiv:2411.16061Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Soft-TransFormers for Continual Learning 25 Nov 2024 · 1 repository · arXiv:2411.16073
-
Solaris: A Foundation Model of the Sun 25 Nov 2024 · 0 repositories · arXiv:2411.16339
-
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training 25 Nov 2024 · 0 repositories · arXiv:2411.16618
-
Swin fMRI Transformer Predicts Early Neurodevelopmental Outcomes from Neonatal fMRI 25 Nov 2024 · 0 repositories · arXiv:2412.07783
-
Tree Transformers are an Ineffective Model of Syntactic Constituency 25 Nov 2024 · 0 repositories · arXiv:2411.16993
-
UltraSam: A Foundation Model for Ultrasound using Large Open-Access Segmentation Datasets 25 Nov 2024 · 1 repository · arXiv:2411.16222
-
VQ-SGen: A Vector Quantized Stroke Representation for Creative Sketch Generation 25 Nov 2024 · 0 repositories · arXiv:2411.16446
-
A Method for Building Large Language Models with Predefined KV Cache Capacity 24 Nov 2024 · 0 repositories · arXiv:2411.15785
-
Adaptive Methods through the Lens of SDEs: Theoretical Insights on the Role of Noise 24 Nov 2024 · 0 repositories · arXiv:2411.15958
-
Broad Critic Deep Actor Reinforcement Learning for Continuous Control 24 Nov 2024 · 0 repositories · arXiv:2411.15806
-
Development of Pre-Trained Transformer-based Models for the Nepali Language 24 Nov 2024 · 0 repositories · arXiv:2411.15734
-
Fixing the Perspective: A Critical Examination of Zero-1-to-3 24 Nov 2024 · 0 repositories · arXiv:2411.15706
-
Investigating Factuality in Long-Form Text Generation: The Roles of Self-Known and Self-Unknown 24 Nov 2024 · 0 repositories · arXiv:2411.15993
-
LTCF-Net: A Transformer-Enhanced Dual-Channel Fourier Framework for Low-Light Image Restoration 24 Nov 2024 · 0 repositories · arXiv:2411.15740
-
Medical Slice Transformer: Improved Diagnosis and Explainability on 3D Medical Images with DINOv2 24 Nov 2024 · 1 repository · arXiv:2411.15802
-
Nimbus: Secure and Efficient Two-Party Inference for Transformers 24 Nov 2024 · 1 repository · arXiv:2411.15707Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
RAMIE: Retrieval-Augmented Multi-task Information Extraction with Large Language Models on Dietary Supplements 24 Nov 2024 · 0 repositories · arXiv:2411.15700
-
A Comparative Analysis of Transformer and LSTM Models for Detecting Suicidal Ideation on Reddit 23 Nov 2024 · 1 repository · arXiv:2411.15404
-
"All that Glitters": Approaches to Evaluations with Unreliable Model and Human Annotations 23 Nov 2024 · 1 repository · arXiv:2411.15634
-
Best of Both Worlds: Advantages of Hybrid Graph Sequence Models 23 Nov 2024 · 0 repositories · arXiv:2411.15671
-
ChatBCI: A P300 Speller BCI Leveraging Large Language Models for Improved Sentence Composition in Realistic Scenarios 23 Nov 2024 · 0 repositories · arXiv:2411.15395
-
Improving Next Tokens via Second-Last Predictions with Generate and Refine 23 Nov 2024 · 0 repositories · arXiv:2411.15661
-
Inducing Human-like Biases in Moral Reasoning Language Models 23 Nov 2024 · 0 repositories · arXiv:2411.15386
-
TANGNN: a Concise, Scalable and Effective Graph Neural Networks with Top-m Attention Mechanism for Graph Representation Learning 23 Nov 2024 · 1 repository · arXiv:2411.15458
-
Traditional Chinese Medicine Case Analysis System for High-Level Semantic Abstraction: Optimized with Prompt and RAG 23 Nov 2024 · 0 repositories · arXiv:2411.15491
-
A New Way: Kronecker-Factored Approximate Curvature Deep Hedging and its Benefits 22 Nov 2024 · 1 repository · arXiv:2411.15002
-
A Real-Time DETR Approach to Bangladesh Road Object Detection for Autonomous Vehicles 22 Nov 2024 · 0 repositories · arXiv:2411.15110
-
AdamZ: An Enhanced Optimisation Method for Neural Network Training 22 Nov 2024 · 1 repository · arXiv:2411.15375
-
Astro-HEP-BERT: A bidirectional language model for studying the meanings of concepts in astrophysics and high energy physics 22 Nov 2024 · 0 repositories · arXiv:2411.14877
-
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective 22 Nov 2024 · 0 repositories · arXiv:2411.14654
-
Detecting Visual Triggers in Cannabis Imagery: A CLIP-Based Multi-Labeling Framework with Local-Global Aggregation 22 Nov 2024 · 0 repositories · arXiv:2412.08648
-
Don't Mesh with Me: Generating Constructive Solid Geometry Instead of Meshes by Fine-Tuning a Code-Generation LLM 22 Nov 2024 · 0 repositories · arXiv:2411.15279
-
ElastiFormer: Learned Redundancy Reduction in Transformer via Self-Distillation 22 Nov 2024 · 0 repositories · arXiv:2411.15281
-
AI Foundation Models for Wearable Movement Data in Mental Health Research 22 Nov 2024 · 1 repository · arXiv:2411.15240
-
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases 22 Nov 2024 · 1 repository · arXiv:2411.14790
-
Multiset Transformer: Advancing Representation Learning in Persistence Diagrams 22 Nov 2024 · 1 repository · arXiv:2411.14662
-
OminiControl: Minimal and Universal Control for Diffusion Transformer 22 Nov 2024 · 2 repositories · arXiv:2411.15098Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Point Cloud Understanding via Attention-Driven Contrastive Learning 22 Nov 2024 · 0 repositories · arXiv:2411.14744
-
Purrfessor: A Fine-tuned Multimodal LLaVA Diet Health Chatbot 22 Nov 2024 · 0 repositories · arXiv:2411.14925
-
RED: Effective Trajectory Representation Learning with Comprehensive Information 22 Nov 2024 · 0 repositories · arXiv:2411.15096
-
Resolution-Agnostic Transformer-based Climate Downscaling 22 Nov 2024 · 0 repositories · arXiv:2411.14774
-
ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data 22 Nov 2024 · 1 repository · arXiv:2411.15004Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Transforming NLU with Babylon: A Case Study in Development of Real-time, Edge-Efficient, Multi-Intent Translation System for Automated Drive-Thru Ordering 22 Nov 2024 · 0 repositories · arXiv:2411.15372
-
When Spatial meets Temporal in Action Recognition 22 Nov 2024 · 0 repositories · arXiv:2411.15284
-
An Experimental Study on Data Augmentation Techniques for Named Entity Recognition on Low-Resource Domains 21 Nov 2024 · 0 repositories · arXiv:2411.14551
-
Assessment of LLM Responses to End-user Security Questions 21 Nov 2024 · 0 repositories · arXiv:2411.14571
-
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels 21 Nov 2024 · 1 repository · arXiv:2411.13775Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
BERT-Based Approach for Automating Course Articulation Matrix Construction with Explainable AI 21 Nov 2024 · 1 repository · arXiv:2411.14254
-
Evaluating the Robustness of Analogical Reasoning in Large Language Models 21 Nov 2024 · 1 repository · arXiv:2411.14215Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Explaining GPT-4's Schema of Depression Using Machine Behavior Analysis 21 Nov 2024 · 0 repositories · arXiv:2411.13800
-
FastRAG: Retrieval Augmented Generation for Semi-structured Data 21 Nov 2024 · 0 repositories · arXiv:2411.13773
-
G-RAG: Knowledge Expansion in Material Science 21 Nov 2024 · 1 repository · arXiv:2411.14592
-
Generative Fuzzy System for Sequence Generation 21 Nov 2024 · 0 repositories · arXiv:2411.13867
-
Global and Local Attention-Based Transformer for Hyperspectral Image Change Detection 21 Nov 2024 · 1 repository · arXiv:2411.14109
-
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI 21 Nov 2024 · 1 repository · arXiv:2411.14522Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Learning from "Silly" Questions Improves Large Language Models, But Only Slightly 21 Nov 2024 · 0 repositories · arXiv:2411.14121
-
Parameter Efficient Mamba Tuning via Projector-targeted Diagonal-centric Linear Transformation 21 Nov 2024 · 0 repositories · arXiv:2411.15224
-
POS-tagging to highlight the skeletal structure of sentences 21 Nov 2024 · 2 repositories · arXiv:2411.14393
-
Stable Flow: Vital Layers for Training-Free Image Editing 21 Nov 2024 · 1 repository · arXiv:2411.14430Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Towards Knowledge Checking in Retrieval-augmented Generation: A Representation Perspective 21 Nov 2024 · 0 repositories · arXiv:2411.14572
-
Understanding World or Predicting Future? A Comprehensive Survey of World Models 21 Nov 2024 · 0 repositories · arXiv:2411.14499