Methods › General › Stochastic Optimization › Adam › Papers, page 5
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 5 of 244: papers 401 to 500 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Hierarchical Document Refinement for Long-context Retrieval-augmented Generation 15 May 2025 · 1 repository · arXiv:2505.10413
-
Leveraging Graph Retrieval-Augmented Generation to Support Learners' Understanding of Knowledge Concepts in MOOCs 15 May 2025 · 0 repositories · arXiv:2505.10074
-
On Technique Identification and Threat-Actor Attribution using LLMs and Embedding Models 15 May 2025 · 1 repository · arXiv:2505.11547
-
One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems 15 May 2025 · 0 repositories · arXiv:2505.11548
-
Pre-Act: Multi-Step Planning and Reasoning Improves Acting in LLM Agents 15 May 2025 · 0 repositories · arXiv:2505.09970
-
Private Transformer Inference in MLaaS: A Survey 15 May 2025 · 0 repositories · arXiv:2505.10315
-
Rethinking Prompt Optimizers: From Prompt Merits to Optimization 15 May 2025 · 1 repository · arXiv:2505.09930
-
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and 𝒪(T) Complexity 15 May 2025 · 1 repository · arXiv:2505.10352Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
VRU-CIPI: Crossing Intention Prediction at Intersections for Improving Vulnerable Road Users Safety 15 May 2025 · 0 repositories · arXiv:2505.09935
-
A Comprehensive Analysis of Large Language Model Outputs: Similarity, Diversity, and Bias 14 May 2025 · 0 repositories · arXiv:2505.09056
-
AdaFortiTran: An Adaptive Transformer Model for Robust OFDM Channel Estimation 14 May 2025 · 1 repository · arXiv:2505.09076
-
Atomic Consistency Preference Optimization for Long-Form Question Answering 14 May 2025 · 1 repository · arXiv:2505.09039
-
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer 14 May 2025 · 1 repository · arXiv:2505.09114
-
BrainNetMLP: An Efficient and Effective Baseline for Functional Brain Network Classification 14 May 2025 · 1 repository · arXiv:2505.11538
-
CXMArena: Unified Dataset to benchmark performance in realistic CXM Scenarios 14 May 2025 · 1 repository · arXiv:2505.09436
-
FAS-LLM: Large Language Model-Based Channel Prediction for OTFS-Enabled Satellite-FAS Links 14 May 2025 · 0 repositories · arXiv:2505.09751
-
How Hungry is AI? Benchmarking Energy, Water, and Carbon Footprint of LLM Inference 14 May 2025 · 0 repositories · arXiv:2505.09598
-
LAS: Loss-less ANN-SNN Conversion for Fully Spike-Driven Large Language Models 14 May 2025 · 1 repository · arXiv:2505.09659
-
Multilingual Machine Translation with Quantum Encoder Decoder Attention-based Convolutional Variational Circuits 14 May 2025 · 0 repositories · arXiv:2505.09407
-
Out-of-distribution generalisation is hard: evidence from ARC-like tasks 14 May 2025 · 0 repositories · arXiv:2505.09716
-
Quotient Complex Transformer (QCformer) for Perovskite Data Analysis 14 May 2025 · 0 repositories · arXiv:2505.09174
-
TopoDiT-3D: Topology-Aware Diffusion Transformer with Bottleneck Structure for 3D Point Cloud Generation 14 May 2025 · 1 repository · arXiv:2505.09140
-
Zero-Shot Multi-modal Large Language Model v.s. Supervised Deep Learning: A Comparative Study on CT-Based Intracranial Hemorrhage Subtyping 14 May 2025 · 1 repository · arXiv:2505.09252
-
A Deep Learning-Driven Inhalation Injury Grading Assistant Using Bronchoscopy Images 13 May 2025 · 0 repositories · arXiv:2505.08517
-
A Head to Predict and a Head to Question: Pre-trained Uncertainty Quantification Heads for Hallucination Detection in LLM Outputs 13 May 2025 · 1 repository · arXiv:2505.08200Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
A suite of LMs comprehend puzzle statements as well as humans 13 May 2025 · 0 repositories · arXiv:2505.08996
-
AC-Reason: Towards Theory-Guided Actual Causality Reasoning with Large Language Models 13 May 2025 · 1 repository · arXiv:2505.08750
-
Achieving Scalable Robot Autonomy via neurosymbolic planning using lightweight local LLM 13 May 2025 · 1 repository · arXiv:2505.08492
-
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets 13 May 2025 · 0 repositories · arXiv:2505.08259
-
Constrained Edge AI Deployment: Fine-Tuning vs Distillation for LLM Compression 13 May 2025 · 0 repositories · arXiv:2505.18166
-
Deep reinforcement learning-based longitudinal control strategy for automated vehicles at signalised intersections 13 May 2025 · 0 repositories · arXiv:2505.08896
-
Enhancing Thyroid Cytology Diagnosis with RAG-Optimized LLMs and Pa-thology Foundation Models 13 May 2025 · 0 repositories · arXiv:2505.08590
-
Evaluating LLM Metrics Through Real-World Capabilities 13 May 2025 · 0 repositories · arXiv:2505.08253
-
Evaluating the Effectiveness of Black-Box Prompt Optimization as the Scale of LLMs Continues to Grow 13 May 2025 · 0 repositories · arXiv:2505.08303
-
For GPT-4 as with Humans: Information Structure Predicts Acceptability of Long-Distance Dependencies 13 May 2025 · 0 repositories · arXiv:2505.09005
-
Hakim: Farsi Text Embedding Model 13 May 2025 · 0 repositories · arXiv:2505.08435
-
HealthBench: Evaluating Large Language Models Towards Improved Human Health 13 May 2025 · 1 repository · arXiv:2505.08775Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
IterKey: Iterative Keyword Generation with LLMs for Enhanced Retrieval Augmented Generation 13 May 2025 · 0 repositories · arXiv:2505.08450
-
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? 13 May 2025 · 1 repository · arXiv:2505.08468
-
Lost in Transmission: When and Why LLMs Fail to Reason Globally 13 May 2025 · 0 repositories · arXiv:2505.08140Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Memorization-Compression Cycles Improve Generalization 13 May 2025 · 0 repositories · arXiv:2505.08727
-
Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control 13 May 2025 · 0 repositories · arXiv:2505.09029
-
OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning 13 May 2025 · 1 repository · arXiv:2505.08617
-
Optimizing Retrieval-Augmented Generation: Analysis of Hyperparameter Impact on Performance and Efficiency 13 May 2025 · 0 repositories · arXiv:2505.08445
-
Probability Consistency in Large Language Models: Theoretical Foundations Meet Empirical Discrepancies 13 May 2025 · 1 repository · arXiv:2505.08739
-
SAR-GTR: Attributed Scattering Information Guided SAR Graph Transformer Recognition Algorithm 13 May 2025 · 0 repositories · arXiv:2505.08547
-
Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing 13 May 2025 · 0 repositories · arXiv:2505.08651
-
Securing RAG: A Risk Assessment and Mitigation Framework 13 May 2025 · 0 repositories · arXiv:2505.08728
-
Small but Significant: On the Promise of Small Language Models for Accessible AIED 13 May 2025 · 0 repositories · arXiv:2505.08588
-
Structural-Temporal Coupling Anomaly Detection with Dynamic Graph Transformer 13 May 2025 · 1 repository · arXiv:2505.08330
-
WixQA: A Multi-Dataset Benchmark for Enterprise Retrieval-Augmented Generation 13 May 2025 · 0 repositories · arXiv:2505.08643
-
A Generative Re-ranking Model for List-level Multi-objective Optimization at Taobao 12 May 2025 · 0 repositories · arXiv:2505.07197
-
AIS Data-Driven Maritime Monitoring Based on Transformer: A Comprehensive Review 12 May 2025 · 1 repository · arXiv:2505.07374
-
An Extra RMSNorm is All You Need for Fine Tuning to 1.58 Bits 12 May 2025 · 0 repositories · arXiv:2505.08823
-
Benchmarking Retrieval-Augmented Generation for Chemistry 12 May 2025 · 0 repositories · arXiv:2505.07671
-
Comparative sentiment analysis of public perception: Monkeypox vs. COVID-19 behavioral insights 12 May 2025 · 0 repositories · arXiv:2505.07430
-
DynamicRAG: Leveraging Outputs of Large Language Model as Feedback for Dynamic Reranking in Retrieval-Augmented Generation 12 May 2025 · 1 repository · arXiv:2505.07233
-
Efficient and Reproducible Biomedical Question Answering using Retrieval Augmented Generation 12 May 2025 · 1 repository · arXiv:2505.07917
-
Fused3S: Fast Sparse Attention on Tensor Cores 12 May 2025 · 1 repository · arXiv:2505.08098
-
Generative Pre-trained Autoregressive Diffusion Transformer 12 May 2025 · 0 repositories · arXiv:2505.07344
-
GRADA: Graph-based Reranker against Adversarial Documents Attack 12 May 2025 · 1 repository · arXiv:2505.07546
-
HAMLET: Healthcare-focused Adaptive Multilingual Learning Embedding-based Topic Modeling 12 May 2025 · 0 repositories · arXiv:2505.07157
-
Hybrid Spiking Vision Transformer for Object Detection with Event Cameras 12 May 2025 · 0 repositories · arXiv:2505.07715
-
KAQG: A Knowledge-Graph-Enhanced RAG for Difficulty-Controlled Question Generation 12 May 2025 · 0 repositories · arXiv:2505.07618
-
LAMM-ViT: AI Face Detection via Layer-Aware Modulation of Region-Guided Attention 12 May 2025 · 0 repositories · arXiv:2505.07734
-
MAIS: Memory-Attention for Interactive Segmentation 12 May 2025 · 0 repositories · arXiv:2505.07511
-
Multi-Objective Reinforcement Learning for Energy-Efficient Industrial Control 12 May 2025 · 0 repositories · arXiv:2505.07607
-
No Query, No Access 12 May 2025 · 0 repositories · arXiv:2505.07258
-
Pre-training vs. Fine-tuning: A Reproducibility Study on Dense Retrieval Knowledge Acquisition 12 May 2025 · 1 repository · arXiv:2505.07166
-
SEReDeEP: Hallucination Detection in Retrieval-Augmented Models via Semantic Entropy and Context-Parameter Fusion 12 May 2025 · 0 repositories · arXiv:2505.07528
-
Sleep Position Classification using Transfer Learning for Bed-based Pressure Sensors 12 May 2025 · 0 repositories · arXiv:2505.08111
-
Statistical CSI-Based Distributed Precoding Design for OFDM-Cooperative Multi-Satellite Systems 12 May 2025 · 0 repositories · arXiv:2505.08038
-
The Geography of Transportation Cybersecurity: Visitor Flows, Industry Clusters, and Spatial Dynamics 12 May 2025 · 0 repositories · arXiv:2505.08822
-
Topology-Guided Knowledge Distillation for Efficient Point Cloud Processing 12 May 2025 · 1 repository · arXiv:2505.08101
-
Towards Requirements Engineering for RAG Systems 12 May 2025 · 0 repositories · arXiv:2505.07553
-
Trial and Trust: Addressing Byzantine Attacks with Comprehensive Defense Strategy 12 May 2025 · 0 repositories · arXiv:2505.07614
-
UMoE: Unifying Attention and FFN with Shared Experts 12 May 2025 · 0 repositories · arXiv:2505.07260
-
Why Uncertainty Estimation Methods Fall Short in RAG: An Axiomatic Analysis 12 May 2025 · 0 repositories · arXiv:2505.07459
-
Evaluating Reasoning LLMs for Suicide Screening with the Columbia-Suicide Severity Rating Scale 11 May 2025 · 1 repository · arXiv:2505.13480
-
IM-BERT: Enhancing Robustness of BERT through the Implicit Euler Method 11 May 2025 · 0 repositories · arXiv:2505.06889
-
Image Classification Using a Diffusion Model as a Pre-Training Model 11 May 2025 · 0 repositories · arXiv:2505.06890
-
NeuRN: Neuro-inspired Domain Generalization for Image Classification 11 May 2025 · 0 repositories · arXiv:2505.06881
-
Streaming Krylov-Accelerated Stochastic Gradient Descent 11 May 2025 · 0 repositories · arXiv:2505.07046
-
Technical Report for ICRA 2025 GOOSE 2D Semantic Segmentation Challenge: Leveraging Color Shift Correction, RoPE-Swin Backbone, and Quantile-based Label Denoising Strategy for Robust Outdoor Scene Understanding 11 May 2025 · 0 repositories · arXiv:2505.06991
-
The Distracting Effect: Understanding Irrelevant Passages in RAG 11 May 2025 · 0 repositories · arXiv:2505.06914
-
Boosting Neural Language Inference via Cascaded Interactive Reasoning 10 May 2025 · 0 repositories · arXiv:2505.06607
-
MacRAG: Compress, Slice, and Scale-up for Multi-Scale Adaptive Context RAG 10 May 2025 · 1 repository · arXiv:2505.06569
-
OMGM: Orchestrate Multiple Granularities and Modalities for Efficient Multimodal Retrieval 10 May 2025 · 0 repositories · arXiv:2505.07879Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Probing In-Context Learning: Impact of Task Complexity and Model Architecture on Generalization and Efficiency 10 May 2025 · 1 repository · arXiv:2505.06475
-
QoS-Efficient Serving of Multiple Mixture-of-Expert LLMs Using Partial Runtime Reconfiguration 10 May 2025 · 0 repositories · arXiv:2505.06481
-
REFINE-AF: A Task-Agnostic Framework to Align Language Models via Self-Generated Instructions using Reinforcement Learning from Automated Feedback 10 May 2025 · 0 repositories · arXiv:2505.06548
-
The Sound of Populism: Distinct Linguistic Features Across Populist Variants 10 May 2025 · 0 repositories · arXiv:2505.07874
-
Underwater object detection in sonar imagery with detection transformer and Zero-shot neural architecture search 10 May 2025 · 0 repositories · arXiv:2505.06694
-
Utilizing LLMs to Investigate the Disputed Role of Evidence in Electronic Cigarette Health Policy Formation in Australia and the UK 10 May 2025 · 0 repositories · arXiv:2505.06782
-
xGen-small Technical Report 10 May 2025 · 0 repositories · arXiv:2505.06496
-
Accurate and Efficient Multivariate Time Series Forecasting via Offline Clustering 9 May 2025 · 0 repositories · arXiv:2505.05738
-
An empathic GPT-based chatbot to talk about mental disorders with Spanish teenagers 9 May 2025 · 0 repositories · arXiv:2505.05828
-
Attention on Multiword Expressions: A Multilingual Study of BERT-based Models with Regard to Idiomaticity and Microsyntax 9 May 2025 · 1 repository · arXiv:2505.06062
-
Camera Control at the Edge with Language Models for Scene Understanding 9 May 2025 · 0 repositories · arXiv:2505.06402
-
CellVerse: Do Large Language Models Really Understand Cell Biology? 9 May 2025 · 0 repositories · arXiv:2505.07865