Methods › General › Stochastic Optimization › Adam › Papers, page 16
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 16 of 244: papers 1,501 to 1,600 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Spiking Transformer:Introducing Accurate Addition-Only Spiking Self-Attention for Transformer 28 Feb 2025 · 0 repositories · arXiv:2503.00226
-
SuperRAG: Beyond RAG with Layout-Aware Graph Modeling 28 Feb 2025 · 0 repositories · arXiv:2503.04790
-
TeleRAG: Efficient Retrieval-Augmented Generation Inference with Lookahead Retrieval 28 Feb 2025 · 0 repositories · arXiv:2502.20969
-
The RAG Paradox: A Black-Box Attack Exploiting Unintentional Vulnerabilities in Retrieval-Augmented Generation Systems 28 Feb 2025 · 0 repositories · arXiv:2502.20995
-
Advanced Deep Learning Techniques for Analyzing Earnings Call Transcripts: Methodologies and Applications 27 Feb 2025 · 0 repositories · arXiv:2503.01886
-
An exploration of features to improve the generalisability of fake news detection models 27 Feb 2025 · 0 repositories · arXiv:2502.20299
-
An Integrated Deep Learning Framework Leveraging NASNet and Vision Transformer with MixProcessing for Accurate and Precise Diagnosis of Lung Diseases 27 Feb 2025 · 0 repositories · arXiv:2502.20570
-
Bridging Legal Knowledge and AI: Retrieval-Augmented Generation with Vector Stores, Knowledge Graphs, and Hierarchical Non-negative Matrix Factorization 27 Feb 2025 · 1 repository · arXiv:2502.20364
-
CirT: Global Subseasonal-to-Seasonal Forecasting with Geometry-inspired Transformer 27 Feb 2025 · 1 repository · arXiv:2502.19750Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
CNsum:Automatic Summarization for Chinese News Text 27 Feb 2025 · 0 repositories · arXiv:2502.19723
-
Enhancing 3D Gaze Estimation in the Wild using Weak Supervision with Gaze Following Labels 27 Feb 2025 · 0 repositories · arXiv:2502.20249
-
IL-SOAR : Imitation Learning with Soft Optimistic Actor cRitic 27 Feb 2025 · 0 repositories · arXiv:2502.19859
-
Long-Context Inference with Retrieval-Augmented Speculative Decoding 27 Feb 2025 · 1 repository · arXiv:2502.20330
-
Minds on the Move: Decoding Trajectory Prediction in Autonomous Driving with Cognitive Insights 27 Feb 2025 · 0 repositories · arXiv:2502.20084
-
Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think 27 Feb 2025 · 1 repository · arXiv:2502.20172
-
QORT-Former: Query-optimized Real-time Transformer for Understanding Two Hands Manipulating Objects 27 Feb 2025 · 0 repositories · arXiv:2502.19769
-
Regional climate projections using a deep-learning-based model-ranking and downscaling framework: Application to European climate zones 27 Feb 2025 · 0 repositories · arXiv:2502.20132
-
Revisit the Stability of Vanilla Federated Learning Under Diverse Conditions 27 Feb 2025 · 0 repositories · arXiv:2502.19849
-
SeisMoLLM: Advancing Seismic Monitoring via Cross-modal Transfer with Pre-trained Large Language Model 27 Feb 2025 · 1 repository · arXiv:2502.19960Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Thinking Slow, Fast: Scaling Inference Compute with Distilled Reasoners 27 Feb 2025 · 0 repositories · arXiv:2502.20339
-
WalnutData: A UAV Remote Sensing Dataset of Green Walnuts and Model Evaluation 27 Feb 2025 · 1 repository · arXiv:2502.20092
-
A Sliding Layer Merging Method for Efficient Depth-Wise Pruning in LLMs 26 Feb 2025 · 1 repository · arXiv:2502.19159
-
AKDT: Adaptive Kernel Dilation Transformer for Effective Image Denoising 26 Feb 2025 · 1 repository
-
Brain-inspired analogical mixture prototypes for few-shot class-incremental learning 26 Feb 2025 · 0 repositories · arXiv:2502.18923
-
Clip-TTS: Contrastive Text-content and Mel-spectrogram, A High-Quality Text-to-Speech Method based on Contextual Semantic Understanding 26 Feb 2025 · 0 repositories · arXiv:2502.18889
-
Cognitive networks highlight differences and similarities in the STEM mindsets of human and LLM-simulated trainees, experts and academics 26 Feb 2025 · 0 repositories · arXiv:2502.19529
-
CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition 26 Feb 2025 · 0 repositories · arXiv:2502.18913
-
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation 26 Feb 2025 · 0 repositories · arXiv:2502.18726
-
Efficient Federated Search for Retrieval-Augmented Generation 26 Feb 2025 · 0 repositories · arXiv:2502.19280
-
Enhanced Transformer-Based Tracking for Skiing Events: Overcoming Multi-Camera Challenges, Scale Variations and Rapid Motion -- SkiTB Visual Tracking Challenge 2025 26 Feb 2025 · 0 repositories · arXiv:2502.18867
-
Evaluating LLMs and Pre-trained Models for Text Summarization Across Diverse Datasets 26 Feb 2025 · 0 repositories · arXiv:2502.19339
-
Genotype-to-Phenotype Prediction in Rice with High-Dimensional Nonlinear Features 26 Feb 2025 · 0 repositories · arXiv:2502.18758
-
Learning to Align Multi-Faceted Evaluation: A Unified and Robust Framework 26 Feb 2025 · 0 repositories · arXiv:2502.18874
-
LORENZA: Enhancing Generalization in Low-Rank Gradient LLM Training via Efficient Zeroth-Order Adaptive SAM 26 Feb 2025 · 0 repositories · arXiv:2502.19571
-
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering 26 Feb 2025 · 0 repositories · arXiv:2502.18993
-
NeoBERT: A Next-Generation BERT 26 Feb 2025 · 1 repository · arXiv:2502.19587
-
Reimagining Personal Data: Unlocking the Potential of AI-Generated Images in Personal Data Meaning-Making 26 Feb 2025 · 0 repositories · arXiv:2502.18853
-
The Sharpness Disparity Principle in Transformers for Accelerating Language Model Pre-Training 26 Feb 2025 · 0 repositories · arXiv:2502.19002
-
Weaker LLMs' Opinions Also Matter: Mixture of Opinions Enhances LLM's Mathematical Reasoning 26 Feb 2025 · 0 repositories · arXiv:2502.19622
-
A Fusion Model for Art Style and Author Recognition Based on Convolutional Neural Networks and Transformers 25 Feb 2025 · 0 repositories · arXiv:2502.18083
-
ART: Anonymous Region Transformer for Variable Multi-Layer Transparent Image Generation 25 Feb 2025 · 1 repository · arXiv:2502.18364Syntology 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples)
-
Assessing Large Language Models in Agentic Multilingual National Bias 25 Feb 2025 · 0 repositories · arXiv:2502.17945
-
Automatic Vehicle Detection using DETR: A Transformer-Based Approach for Navigating Treacherous Roads 25 Feb 2025 · 0 repositories · arXiv:2502.17843
-
Bayesian Optimization for Controlled Image Editing via LLMs 25 Feb 2025 · 0 repositories · arXiv:2502.18116
-
Broadening Discovery through Structural Models: Multimodal Combination of Local and Structural Properties for Predicting Chemical Features 25 Feb 2025 · 0 repositories · arXiv:2502.17986
-
Detecting Knowledge Boundary of Vision Large Language Models by Sampling-Based Inference 25 Feb 2025 · 0 repositories · arXiv:2502.18023
-
Enhancing Speech Quality through the Integration of BGRU and Transformer Architectures 25 Feb 2025 · 0 repositories · arXiv:2502.17911
-
Enhancing Text Classification with a Novel Multi-Agent Collaboration Framework Leveraging BERT 25 Feb 2025 · 0 repositories · arXiv:2502.18653
-
Examining the Threat Landscape: Foundation Models and Model Stealing 25 Feb 2025 · 0 repositories · arXiv:2502.18077
-
Faster, Cheaper, Better: Multi-Objective Hyperparameter Optimization for LLM and RAG Systems 25 Feb 2025 · 0 repositories · arXiv:2502.18635
-
H-FLTN: A Privacy-Preserving Hierarchical Framework for Electric Vehicle Spatio-Temporal Charge Prediction 25 Feb 2025 · 0 repositories · arXiv:2502.18697
-
Independent Mobility GPT (IDM-GPT): A Self-Supervised Multi-Agent Large Language Model Framework for Customized Traffic Mobility Analysis Using Machine Learning Models 25 Feb 2025 · 0 repositories · arXiv:2502.18652
-
LAM: Large Avatar Model for One-shot Animatable Gaussian Head 25 Feb 2025 · 0 repositories · arXiv:2502.17796
-
Learning Structure-Supporting Dependencies via Keypoint Interactive Transformer for General Mammal Pose Estimation 25 Feb 2025 · 1 repository · arXiv:2502.18214
-
LevelRAG: Enhancing Retrieval-Augmented Generation with Multi-hop Logic Planning over Rewriting Augmented Searchers 25 Feb 2025 · 1 repository · arXiv:2502.18139
-
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks 25 Feb 2025 · 1 repository · arXiv:2502.17832
-
MuCoS: Efficient Drug-Target Prediction through Multi-Context-Aware Sampling 25 Feb 2025 · 0 repositories · arXiv:2502.17784
-
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation 25 Feb 2025 · 0 repositories · arXiv:2502.17839
-
Scaling LLM Pre-training with Vocabulary Curriculum 25 Feb 2025 · 0 repositories · arXiv:2502.17910
-
Self-Adjust Softmax 25 Feb 2025 · 0 repositories · arXiv:2502.18277
-
Stackelberg Game Preference Optimization for Data-Efficient Alignment of Language Models 25 Feb 2025 · 0 repositories · arXiv:2502.18099
-
ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents 25 Feb 2025 · 1 repository · arXiv:2502.18017Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
"Actionable Help" in Crises: A Novel Dataset and Resource-Efficient Models for Identifying Request and Offer Social Media Posts 24 Feb 2025 · 0 repositories · arXiv:2502.16839
-
Adversarial Training for Defense Against Label Poisoning Attacks 24 Feb 2025 · 1 repository · arXiv:2502.17121Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Applying LLMs to Active Learning: Towards Cost-Efficient Cross-Task Text Classification without Manually Labeled Data 24 Feb 2025 · 0 repositories · arXiv:2502.16892
-
Are Large Language Models Good Data Preprocessors? 24 Feb 2025 · 0 repositories · arXiv:2502.16790
-
Atten-Transformer: A Deep Learning Framework for User App Usage Prediction 24 Feb 2025 · 0 repositories · arXiv:2502.16957
-
Benchmarking Retrieval-Augmented Generation in Multi-Modal Contexts 24 Feb 2025 · 2 repositories · arXiv:2502.17297
-
CalibRefine: Deep Learning-Based Online Automatic Targetless LiDAR-Camera Calibration with Iterative and Attention-Driven Post-Refinement 24 Feb 2025 · 1 repository · arXiv:2502.17648
-
CipherPrune: Efficient and Scalable Private Transformer Inference 24 Feb 2025 · 1 repository · arXiv:2502.16782Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation 24 Feb 2025 · 0 repositories · arXiv:2502.17198
-
Disentangling Visual Transformers: Patch-level Interpretability for Image Classification 24 Feb 2025 · 0 repositories · arXiv:2502.17196
-
ENACT-Heart -- ENsemble-based Assessment Using CNN and Transformer on Heart Sounds 24 Feb 2025 · 0 repositories · arXiv:2502.16914
-
Enhancing Image Matting in Real-World Scenes with Mask-Guided Iterative Refinement 24 Feb 2025 · 0 repositories · arXiv:2502.17093
-
Evaluating the Effect of Retrieval Augmentation on Social Biases 24 Feb 2025 · 0 repositories · arXiv:2502.17611
-
Functional BART with Shape Priors: A Bayesian Tree Approach to Constrained Functional Regression 24 Feb 2025 · 0 repositories · arXiv:2502.16888
-
GaussianFlowOcc: Sparse and Weakly Supervised Occupancy Estimation using Gaussian Splatting and Temporal Flow 24 Feb 2025 · 0 repositories · arXiv:2502.17288
-
LettuceDetect: A Hallucination Detection Framework for RAG Applications 24 Feb 2025 · 2 repositories · arXiv:2502.17125
-
LLM Inference Acceleration via Efficient Operation Fusion 24 Feb 2025 · 0 repositories · arXiv:2502.17728
-
Logic Haystacks: Probing LLMs Long-Context Logical Reasoning (Without Easily Identifiable Unrelated Padding) 24 Feb 2025 · 0 repositories · arXiv:2502.17169
-
MaxGlaViT: A novel lightweight vision transformer-based approach for early diagnosis of glaucoma stages from fundus images 24 Feb 2025 · 1 repository · arXiv:2502.17154
-
MDN: Mamba-Driven Dualstream Network For Medical Hyperspectral Image Segmentation 24 Feb 2025 · 0 repositories · arXiv:2502.17255
-
MEMERAG: A Multilingual End-to-End Meta-Evaluation Benchmark for Retrieval Augmented Generation 24 Feb 2025 · 1 repository · arXiv:2502.17163
-
Mitigating Bias in RAG: Controlling the Embedder 24 Feb 2025 · 1 repository · arXiv:2502.17390
-
Mutual Reinforcement of LLM Dialogue Synthesis and Summarization Capabilities for Few-Shot Dialogue Summarization 24 Feb 2025 · 0 repositories · arXiv:2502.17328
-
Stable-SPAM: How to Train in 4-Bit More Stably than 16-Bit Adam 24 Feb 2025 · 1 repository · arXiv:2502.17055
-
Towards Typologically Aware Rescoring to Mitigate Unfaithfulness in Lower-Resource Languages 24 Feb 2025 · 0 repositories · arXiv:2502.17664
-
Unraveling the geometry of visual relational reasoning 24 Feb 2025 · 1 repository · arXiv:2502.17382
-
A Split-Window Transformer for Multi-Model Sequence Spammer Detection using Multi-Model Variational Autoencoder 23 Feb 2025 · 0 repositories · arXiv:2502.16483
-
D2S-FLOW: Automated Parameter Extraction from Datasheets for SPICE Model Generation Using Large Language Models 23 Feb 2025 · 0 repositories · arXiv:2502.16540
-
AeroReformer: Aerial Referring Transformer for UAV-based Referring Image Segmentation 23 Feb 2025 · 1 repository · arXiv:2502.16680
-
Co-MTP: A Cooperative Trajectory Prediction Framework with Multi-Temporal Fusion for Autonomous Driving 23 Feb 2025 · 1 repository · arXiv:2502.16589
-
Code Summarization Beyond Function Level 23 Feb 2025 · 1 repository · arXiv:2502.16704
-
Dynamic LLM Routing and Selection based on User Preferences: Balancing Performance, Cost, and Ethics 23 Feb 2025 · 0 repositories · arXiv:2502.16696
-
GS-TransUNet: Integrated 2D Gaussian Splatting and Transformer UNet for Accurate Skin Lesion Analysis 23 Feb 2025 · 1 repository · arXiv:2502.16748
-
Layer-Wise Evolution of Representations in Fine-Tuned Transformers: Insights from Sparse AutoEncoders 23 Feb 2025 · 0 repositories · arXiv:2502.16722
-
Optimizing Retrieval-Augmented Generation of Medical Content for Spaced Repetition Learning 23 Feb 2025 · 0 repositories · arXiv:2503.01859
-
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning 23 Feb 2025 · 1 repository · arXiv:2502.16496
-
Reasoning about Affordances: Causal and Compositional Reasoning in LLMs 23 Feb 2025 · 0 repositories · arXiv:2502.16606
-
Retrieval-Augmented Visual Question Answering via Built-in Autoregressive Search Engines 23 Feb 2025 · 0 repositories · arXiv:2502.16641