Methods › General › Stochastic Optimization › Adam › Papers, page 29
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 29 of 244: papers 2,801 to 2,900 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Inverting Transformer-based Vision Models 9 Dec 2024 · 2 repositories · arXiv:2412.06534
-
LLM as HPC Expert: Extending RAG Architecture for HPC Data 9 Dec 2024 · 0 repositories · arXiv:2501.14733
-
Normalizing Flows are Capable Generative Models 9 Dec 2024 · 3 repositories · arXiv:2412.06329Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06249
-
Parkinson's Disease Diagnosis Through Deep Learning: A Novel LSTM-Based Approach for Freezing of Gait Detection 9 Dec 2024 · 0 repositories · arXiv:2412.06709
-
S²FT: Efficient, Scalable and Generalizable LLM Fine-tuning by Structured Sparsity 9 Dec 2024 · 0 repositories · arXiv:2412.06289
-
SiReRAG: Indexing Similar and Related Information for Multihop Reasoning 9 Dec 2024 · 0 repositories · arXiv:2412.06206Syntology 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
The Computational Limits of State-Space Models and Mamba via the Lens of Circuit Complexity 9 Dec 2024 · 0 repositories · arXiv:2412.06148
-
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.17821
-
Toward Non-Invasive Diagnosis of Bankart Lesions with Deep Learning 9 Dec 2024 · 1 repository · arXiv:2412.06717
-
Unseen Attack Detection in Software-Defined Networking Using a BERT-Based Large Language Model 9 Dec 2024 · 0 repositories · arXiv:2412.06239
-
A Collaborative Multi-Agent Approach to Retrieval-Augmented Generation Across Diverse Data 8 Dec 2024 · 0 repositories · arXiv:2412.05838
-
Enhanced Computationally Efficient Long LoRA Inspired Perceiver Architectures for Auto-Regressive Language Modeling 8 Dec 2024 · 0 repositories · arXiv:2412.06106
-
Evaluating Robustness of LLMs on Crisis-Related Microblogs across Events, Information Types, and Linguistic Features 8 Dec 2024 · 0 repositories · arXiv:2412.10413
-
Fully Open Source Moxin-7B Technical Report 8 Dec 2024 · 1 repository · arXiv:2412.06845
-
KITE-DDI: A Knowledge graph Integrated Transformer Model for accurately predicting Drug-Drug Interaction Events from Drug SMILES and Biomedical Knowledge Graph 8 Dec 2024 · 0 repositories · arXiv:2412.05770
-
Language-Guided Image Tokenization for Generation 8 Dec 2024 · 0 repositories · arXiv:2412.05796
-
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor 8 Dec 2024 · 1 repository · arXiv:2412.07801
-
M³-20M: A Large-Scale Multi-Modal Molecule Dataset for AI-driven Drug Design and Discovery 8 Dec 2024 · 1 repository · arXiv:2412.06847
-
Mixture-of-PageRanks: Replacing Long-Context with Real-Time, Sparse GraphRAG 8 Dec 2024 · 0 repositories · arXiv:2412.06078
-
Paddy Disease Detection and Classification Using Computer Vision Techniques: A Mobile Application to Detect Paddy Disease 8 Dec 2024 · 0 repositories · arXiv:2412.05996
-
A Comparative Study on Code Generation with Transformers 7 Dec 2024 · 0 repositories · arXiv:2412.05749
-
BERTCaps: BERT Capsule for Persian Multi-Domain Sentiment Analysis 7 Dec 2024 · 0 repositories · arXiv:2412.05591
-
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds 7 Dec 2024 · 1 repository · arXiv:2412.05631Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
Innovative Sentiment Analysis and Prediction of Stock Price Using FinBERT, GPT-4 and Logistic Regression: A Data-Driven Approach 7 Dec 2024 · 0 repositories · arXiv:2412.06837
-
KG-Retriever: Efficient Knowledge Indexing for Retrieval-Augmented Large Language Models 7 Dec 2024 · 1 repository · arXiv:2412.05547Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
M³PC: Test-time Model Predictive Control for Pretrained Masked Trajectory Model 7 Dec 2024 · 1 repository · arXiv:2412.05675Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
PrivAgent: Agentic-based Red-teaming for LLM Privacy Leakage 7 Dec 2024 · 1 repository · arXiv:2412.05734
-
RefSAM3D: Adapting SAM with Cross-modal Reference for 3D Medical Image Segmentation 7 Dec 2024 · 0 repositories · arXiv:2412.05605
-
Shifting NER into High Gear: The Auto-AdvER Approach 7 Dec 2024 · 0 repositories · arXiv:2412.05655
-
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering 7 Dec 2024 · 0 repositories · arXiv:2412.06832
-
Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning 7 Dec 2024 · 2 repositories · arXiv:2412.05586Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo 6 Dec 2024 · 0 repositories · arXiv:2412.05223
-
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG 6 Dec 2024 · 0 repositories · arXiv:2412.05447
-
Are Frontier Large Language Models Suitable for Q&A in Science Centres? 6 Dec 2024 · 0 repositories · arXiv:2412.05200
-
BEExformer: A Fast Inferencing Transformer Architecture via Binarization with Multiple Early Exits 6 Dec 2024 · 0 repositories · arXiv:2412.05225
-
DHIL-GT: Scalable Graph Transformer with Decoupled Hierarchy Labeling 6 Dec 2024 · 0 repositories · arXiv:2412.04738
-
Enhancing Cross-Language Code Translation via Task-Specific Embedding Alignment in Retrieval-Augmented Generation 6 Dec 2024 · 0 repositories · arXiv:2412.05159
-
Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System 6 Dec 2024 · 0 repositories · arXiv:2412.06828
-
Feature Group Tabular Transformer: A Novel Approach to Traffic Crash Modeling and Causality Analysis 6 Dec 2024 · 0 repositories · arXiv:2412.06825
-
NLP-ADBench: NLP Anomaly Detection Benchmark 6 Dec 2024 · 1 repository · arXiv:2412.04784
-
Privacy-Preserving Retrieval-Augmented Generation with Differential Privacy 6 Dec 2024 · 0 repositories · arXiv:2412.04697
-
QueEn: A Large Language Model for Quechua-English Translation 6 Dec 2024 · 0 repositories · arXiv:2412.05184
-
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens 6 Dec 2024 · 1 repository · arXiv:2412.04680Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Towards Understanding the Role of Sharpness-Aware Minimization Algorithms for Out-of-Distribution Generalization 6 Dec 2024 · 0 repositories · arXiv:2412.05169
-
Addressing Hallucinations with RAG and NMISS in Italian Healthcare LLM Chatbots 5 Dec 2024 · 0 repositories · arXiv:2412.04235
-
ARTeFACT: Benchmarking Segmentation Models on Diverse Analogue Media Damage 5 Dec 2024 · 0 repositories · arXiv:2412.04580
-
Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding 5 Dec 2024 · 0 repositories · arXiv:2412.03980
-
Cubify Anything: Scaling Indoor 3D Object Detection 5 Dec 2024 · 1 repository · arXiv:2412.04458
-
DEIM: DETR with Improved Matching for Fast Convergence 5 Dec 2024 · 1 repository · arXiv:2412.04234Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Dynamic Graph Representation with Contrastive Learning for Financial Market Prediction: Integrating Temporal Evolution and Static Relations 5 Dec 2024 · 1 repository · arXiv:2412.04034
-
Exploring AI Text Generation, Retrieval-Augmented Generation, and Detection Technologies: a Comprehensive Overview 5 Dec 2024 · 0 repositories · arXiv:2412.03933
-
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning 5 Dec 2024 · 1 repository · arXiv:2412.04661
-
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments? 5 Dec 2024 · 0 repositories · arXiv:2412.03856
-
TransAdapter: Vision Transformer for Feature-Centric Unsupervised Domain Adaptation 5 Dec 2024 · 1 repository · arXiv:2412.04073
-
Uniform Discretized Integrated Gradients: An effective attribution based method for explaining large language models 5 Dec 2024 · 0 repositories · arXiv:2412.03886
-
A Water Efficiency Dataset for African Data Centers 4 Dec 2024 · 0 repositories · arXiv:2412.03716
-
Advanced Risk Prediction and Stability Assessment of Banks Using Time Series Transformer Models 4 Dec 2024 · 0 repositories · arXiv:2412.03606
-
Advancing Conversational Psychotherapy: Integrating Privacy, Dual-Memory, and Domain Expertise with Large Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.02987
-
AntLM: Bridging Causal and Masked Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.03275
-
Controlling the Mutation in Large Language Models for the Efficient Evolution of Algorithms 4 Dec 2024 · 0 repositories · arXiv:2412.03250
-
Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts? 4 Dec 2024 · 0 repositories · arXiv:2412.03235
-
EMPATH: MediaPipe-Aided Ensemble Learning with Attention-Based Transformers for Accurate Recognition of Bangla Word-Level Sign Language 4 Dec 2024 · 1 repository
-
FANAL -- Financial Activity News Alerting Language Modeling Framework 4 Dec 2024 · 0 repositories · arXiv:2412.03527
-
Interpreting Transformers for Jet Tagging 4 Dec 2024 · 1 repository · arXiv:2412.03673Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MaterialPicker: Multi-Modal Material Generation with Diffusion Transformers 4 Dec 2024 · 0 repositories · arXiv:2412.03225
-
Multi-Branch Mutual-Distillation Transformer for EEG-Based Seizure Subtype Classification 4 Dec 2024 · 0 repositories · arXiv:2412.15224
-
Multimodal Sentiment Analysis Based on BERT and ResNet 4 Dec 2024 · 0 repositories · arXiv:2412.03625
-
Navigation World Models 4 Dec 2024 · 1 repository · arXiv:2412.03572Syntology 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention 4 Dec 2024 · 0 repositories · arXiv:2412.03520
-
Theoretical limitations of multi-layer Transformer 4 Dec 2024 · 1 repository · arXiv:2412.02975
-
Achieving Semantic Consistency: Contextualized Word Representations for Political Text Analysis 3 Dec 2024 · 0 repositories · arXiv:2412.04505
-
CAISSON: Concept-Augmented Inference Suite of Self-Organizing Neural Networks 3 Dec 2024 · 0 repositories · arXiv:2412.02835
-
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity 3 Dec 2024 · 0 repositories · arXiv:2412.02252
-
Conformal Symplectic Optimization for Stable Reinforcement Learning 3 Dec 2024 · 1 repository · arXiv:2412.02291
-
CPTQuant -- A Novel Mixed Precision Post-Training Quantization Techniques for Large Language Models 3 Dec 2024 · 0 repositories · arXiv:2412.03599
-
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators 3 Dec 2024 · 1 repository · arXiv:2412.02467
-
FCL-ViT: Task-Aware Attention Tuning for Continual Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02509
-
Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model 3 Dec 2024 · 0 repositories · arXiv:2412.02802
-
GQWformer: A Quantum-based Transformer for Graph Representation Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02285
-
Gracefully Filtering Backdoor Samples for Generative Large Language Models without Retraining 3 Dec 2024 · 1 repository · arXiv:2412.02454
-
Impact of Data Snooping on Deep Learning Models for Locating Vulnerabilities in Lifted Code 3 Dec 2024 · 0 repositories · arXiv:2412.02048
-
Leveraging Large Language Models for Comparative Literature Summarization with Reflective Incremental Mechanisms 3 Dec 2024 · 0 repositories · arXiv:2412.02149
-
MAGMA: Manifold Regularization for MAEs 3 Dec 2024 · 1 repository · arXiv:2412.02871
-
OCR Hinders RAG: Evaluating the Cascading Impact of OCR on Retrieval-Augmented Generation 3 Dec 2024 · 1 repository · arXiv:2412.02592
-
Optimization of Transformer heart disease prediction model based on particle swarm optimization algorithm 3 Dec 2024 · 0 repositories · arXiv:2412.02801
-
Patent-CR: A Dataset for Patent Claim Revision 3 Dec 2024 · 0 repositories · arXiv:2412.02549
-
RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models 3 Dec 2024 · 1 repository · arXiv:2412.02830
-
Revisiting the Initial Steps in Adaptive Gradient Descent Optimization 3 Dec 2024 · 0 repositories · arXiv:2412.02153
-
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction 3 Dec 2024 · 0 repositories · arXiv:2412.02698
-
Semantic Tokens in Retrieval Augmented Generation 3 Dec 2024 · 0 repositories · arXiv:2412.02563
-
The Asymptotic Behavior of Attention in Transformers 3 Dec 2024 · 0 repositories · arXiv:2412.02682
-
Transformer-Based Auxiliary Loss for Face Recognition Across Age Variations 3 Dec 2024 · 0 repositories · arXiv:2412.02198
-
Automated Extraction of Acronym-Expansion Pairs from Scientific Papers 2 Dec 2024 · 0 repositories · arXiv:2412.01093
-
Convolutional Transformer Neural Collaborative Filtering 2 Dec 2024 · 0 repositories · arXiv:2412.01376
-
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation 2 Dec 2024 · 0 repositories · arXiv:2412.01429
-
FGATT: A Robust Framework for Wireless Data Imputation Using Fuzzy Graph Attention Networks and Transformer Encoders 2 Dec 2024 · 0 repositories · arXiv:2412.01979
-
GETAE: Graph information Enhanced deep neural NeTwork ensemble ArchitecturE for fake news detection 2 Dec 2024 · 1 repository · arXiv:2412.01825
-
Global Average Feature Augmentation for Robust Semantic Segmentation with Transformers 2 Dec 2024 · 0 repositories · arXiv:2412.01941