Methods › General › Stochastic Optimization › Adam › Papers, page 38
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 38 of 244: papers 3,701 to 3,800 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Supervised Chain of Thought 18 Oct 2024 · 0 repositories · arXiv:2410.14198
-
TimeSeriesExam: A time series understanding exam 18 Oct 2024 · 1 repository · arXiv:2410.14752Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Transfer Learning on Transformers for Building Energy Consumption Forecasting -- A Comparative Study 18 Oct 2024 · 0 repositories · arXiv:2410.14107
-
xPerT: Extended Persistence Transformer 18 Oct 2024 · 2 repositories · arXiv:2410.14193
-
How Does Knowledge Selection Help Retrieval Augmented Generation? 17 Oct 2024 · 0 repositories · arXiv:2410.13258
-
Active-Dormant Attention Heads: Mechanistically Demystifying Extreme-Token Phenomena in LLMs 17 Oct 2024 · 1 repository · arXiv:2410.13835Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Adversarial Testing as a Tool for Interpretability: Length-based Overfitting of Elementary Functions in Transformers 17 Oct 2024 · 0 repositories · arXiv:2410.13802
-
Better to Ask in English: Evaluation of Large Language Models on English, Low-resource and Cross-Lingual Settings 17 Oct 2024 · 0 repositories · arXiv:2410.13153
-
Co-Segmentation without any Pixel-level Supervision with Application to Large-Scale Sketch Classification 17 Oct 2024 · 0 repositories · arXiv:2410.13582
-
D-FINE: Redefine Regression Task in DETRs as Fine-grained Distribution Refinement 17 Oct 2024 · 5 repositories · arXiv:2410.13842Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
Detecting AI-Generated Texts in Cross-Domains 17 Oct 2024 · 1 repository · arXiv:2410.13966
-
DurIAN-E 2: Duration Informed Attention Network with Adaptive Variational Autoencoder and Adversarial Learning for Expressive Text-to-Speech Synthesis 17 Oct 2024 · 0 repositories · arXiv:2410.13288
-
Enhancing Generalization in Sparse Mixture of Experts Models: The Case for Increased Expert Activation in Compositional Tasks 17 Oct 2024 · 0 repositories · arXiv:2410.13964
-
Enhancing Text Generation in Joint NLG/NLU Learning Through Curriculum Learning, Semi-Supervised Training, and Advanced Optimization Techniques 17 Oct 2024 · 0 repositories · arXiv:2410.13498
-
Evaluating Self-Generated Documents for Enhancing Retrieval-Augmented Generation with Large Language Models 17 Oct 2024 · 0 repositories · arXiv:2410.13192
-
FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs 17 Oct 2024 · 2 repositories · arXiv:2410.13210
-
Help Me Identify: Is an LLM+VQA System All We Need to Identify Visual Concepts? 17 Oct 2024 · 1 repository · arXiv:2410.13651
-
Hiformer: Hybrid Frequency Feature Enhancement Inverted Transformer for Long-Term Wind Power Prediction 17 Oct 2024 · 1 repository · arXiv:2410.13303
-
Integrating Temporal Representations for Dynamic Memory Retrieval and Management in Large Language Models 17 Oct 2024 · 0 repositories · arXiv:2410.13553
-
IterSelectTune: An Iterative Training Framework for Efficient Instruction-Tuning Data Selection 17 Oct 2024 · 0 repositories · arXiv:2410.13464
-
Jailbreaking LLM-Controlled Robots 17 Oct 2024 · 0 repositories · arXiv:2410.13691
-
Learning Graph Quantized Tokenizers 17 Oct 2024 · 1 repository · arXiv:2410.13798Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Linguistically Grounded Analysis of Language Models using Shapley Head Values 17 Oct 2024 · 0 repositories · arXiv:2410.13396
-
LoLDU: Low-Rank Adaptation via Lower-Diag-Upper Decomposition for Parameter-Efficient Fine-Tuning 17 Oct 2024 · 1 repository · arXiv:2410.13618Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Looking Inward: Language Models Can Learn About Themselves by Introspection 17 Oct 2024 · 1 repository · arXiv:2410.13787Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MarineFormer: A Spatio-Temporal Attention Model for USV Navigation in Dynamic Marine Environments 17 Oct 2024 · 0 repositories · arXiv:2410.13973
-
MCQG-SRefine: Multiple Choice Question Generation and Evaluation with Iterative Self-Critique, Correction, and Comparison Feedback 17 Oct 2024 · 1 repository · arXiv:2410.13191
-
Measuring and Modifying the Readability of English Texts with GPT-4 17 Oct 2024 · 1 repository · arXiv:2410.14028
-
Judgment of Learning: A Human Ability Beyond Generative Artificial Intelligence 17 Oct 2024 · 0 repositories · arXiv:2410.13392
-
MIRAGE-Bench: Automatic Multilingual Benchmark Arena for Retrieval-Augmented Generation Systems 17 Oct 2024 · 1 repository · arXiv:2410.13716
-
On the Learn-to-Optimize Capabilities of Transformers in In-Context Sparse Recovery 17 Oct 2024 · 0 repositories · arXiv:2410.13981
-
Personalized Adaptation via In-Context Preference Learning 17 Oct 2024 · 0 repositories · arXiv:2410.14001
-
Precipitation Nowcasting Using Diffusion Transformer with Causal Attention 17 Oct 2024 · 0 repositories · arXiv:2410.13314
-
RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rewards 17 Oct 2024 · 1 repository · arXiv:2410.13509Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
RGB to Hyperspectral: Spectral Reconstruction for Enhanced Surgical Imaging 17 Oct 2024 · 0 repositories · arXiv:2410.13570
-
SBI-RAG: Enhancing Math Word Problem Solving for Students through Schema-Based Instruction and Retrieval-Augmented Generation 17 Oct 2024 · 1 repository · arXiv:2410.13293
-
SouLLMate: An Application Enhancing Diverse Mental Health Support with Adaptive LLMs, Prompt Engineering, and RAG Techniques 17 Oct 2024 · 0 repositories · arXiv:2410.16322
-
Temporal-Enhanced Multimodal Transformer for Referring Multi-Object Tracking and Segmentation 17 Oct 2024 · 0 repositories · arXiv:2410.13437
-
Towards Cross-Cultural Machine Translation with Retrieval-Augmented Generation from Multilingual Knowledge Graphs 17 Oct 2024 · 0 repositories · arXiv:2410.14057Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Training Compute-Optimal Vision Transformers for Brain Encoding 17 Oct 2024 · 0 repositories · arXiv:2410.19810
-
UniGS: Modeling Unitary 3D Gaussians for Novel View Synthesis from Sparse-view Images 17 Oct 2024 · 2 repositories · arXiv:2410.13195
-
Advancing Fairness in Natural Language Processing: From Traditional Methods to Explainability 16 Oct 2024 · 0 repositories · arXiv:2410.12511
-
Agent Skill Acquisition for Large Language Models via CycleQD 16 Oct 2024 · 1 repository · arXiv:2410.14735Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
AT-RAG: An Adaptive RAG Model Enhancing Query Efficiency with Topic Filtering and Iterative Reasoning 16 Oct 2024 · 1 repository · arXiv:2410.12886
-
CCSBench: Evaluating Compositional Controllability in LLMs for Scientific Document Summarization 16 Oct 2024 · 0 repositories · arXiv:2410.12601
-
CoFE-RAG: A Comprehensive Full-chain Evaluation Framework for Retrieval-Augmented Generation with Enhanced Data Diversity 16 Oct 2024 · 1 repository · arXiv:2410.12248
-
Communication-Efficient and Tensorized Federated Fine-Tuning of Large Language Models 16 Oct 2024 · 0 repositories · arXiv:2410.13097
-
Context-Scaling versus Task-Scaling in In-Context Learning 16 Oct 2024 · 0 repositories · arXiv:2410.12783
-
Evaluating Morphological Compositional Generalization in Large Language Models 16 Oct 2024 · 1 repository · arXiv:2410.12656Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Evaluation of Attribution Bias in Retrieval-Augmented Large Language Models 16 Oct 2024 · 0 repositories · arXiv:2410.12380
-
Exploring Large Language Models for Hate Speech Detection in Rioplatense Spanish 16 Oct 2024 · 0 repositories · arXiv:2410.12174
-
FusionLLM: A Decentralized LLM Training System on Geo-distributed GPUs with Adaptive Compression 16 Oct 2024 · 0 repositories · arXiv:2410.12707
-
Hypothesis Testing the Circuit Hypothesis in LLMs 16 Oct 2024 · 1 repository · arXiv:2410.13032Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Identifying Task Groupings for Multi-Task Learning Using Pointwise V-Usable Information 16 Oct 2024 · 0 repositories · arXiv:2410.12774
-
Is Semantic Chunking Worth the Computational Cost? 16 Oct 2024 · 0 repositories · arXiv:2410.13070
-
Kallini et al. (2024) do not compare impossible languages with constituency-based ones 16 Oct 2024 · 0 repositories · arXiv:2410.12271
-
MambaBEV: An efficient 3D detection model with Mamba2 16 Oct 2024 · 0 repositories · arXiv:2410.12673
-
Meta-Chunking: Learning Text Segmentation and Semantic Completion via Logical Perception 16 Oct 2024 · 1 repository · arXiv:2410.12788Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
MIRROR: A Novel Approach for the Automated Evaluation of Open-Ended Question Generation 16 Oct 2024 · 0 repositories · arXiv:2410.12893
-
MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models 16 Oct 2024 · 1 repository · arXiv:2410.13085Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
MSc-SQL: Multi-Sample Critiquing Small Language Models For Text-To-SQL Translation 16 Oct 2024 · 1 repository · arXiv:2410.12916Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
On A Scale From 1 to 5: Quantifying Hallucination in Faithfulness Evaluation 16 Oct 2024 · 0 repositories · arXiv:2410.12222
-
ShapefileGPT: A Multi-Agent Large Language Model Framework for Automated Shapefile Processing 16 Oct 2024 · 0 repositories · arXiv:2410.12376
-
Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective 16 Oct 2024 · 1 repository · arXiv:2410.12490Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
AttentiveMOS: A Lightweight Attention-Only Model for Speech Quality Prediction 16 Oct 2024 · 0 repositories · arXiv:2410.12675
-
Table-LLM-Specialist: Language Model Specialists for Tables using Iterative Generator-Validator Fine-tuning 16 Oct 2024 · 0 repositories · arXiv:2410.12164
-
Tracking Universal Features Through Fine-Tuning and Model Merging 16 Oct 2024 · 0 repositories · arXiv:2410.12391
-
Unifying Economic and Language Models for Enhanced Sentiment Analysis of the Oil Market 16 Oct 2024 · 0 repositories · arXiv:2410.12473
-
Unitary Multi-Margin BERT for Robust Natural Language Processing 16 Oct 2024 · 1 repository · arXiv:2410.12759
-
When Not to Answer: Evaluating Prompts on GPT Models for Effective Abstention in Unanswerable Math Word Problems 16 Oct 2024 · 0 repositories · arXiv:2410.13029
-
Athena: Retrieval-augmented Legal Judgment Prediction with Large Language Models 15 Oct 2024 · 0 repositories · arXiv:2410.11195
-
Bypassing the Exponential Dependency: Looped Transformers Efficiently Learn In-context by Multi-step Gradient Descent 15 Oct 2024 · 0 repositories · arXiv:2410.11268
-
Cognitive Overload Attack:Prompt Injection for Long Context 15 Oct 2024 · 1 repository · arXiv:2410.11272
-
De-jargonizing Science for Journalists with GPT-4: A Pilot Study 15 Oct 2024 · 1 repository · arXiv:2410.12069
-
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation 15 Oct 2024 · 1 repository · arXiv:2410.11317Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
DODT: Enhanced Online Decision Transformer Learning through Dreamer's Actor-Critic Trajectory Forecasting 15 Oct 2024 · 0 repositories · arXiv:2410.11359
-
DynamicER: Resolving Emerging Mentions to Dynamic Entities for RAG 15 Oct 2024 · 1 repository · arXiv:2410.11494
-
Efficient Partitioning Vision Transformer on Edge Devices for Distributed Inference 15 Oct 2024 · 0 repositories · arXiv:2410.11650
-
EFILN: The Electric Field Inversion-Localization Network for High-Precision Underwater Positioning 15 Oct 2024 · 0 repositories · arXiv:2410.11223
-
Evidence of Cognitive Deficits andDevelopmental Advances in Generative AI: A Clock Drawing Test Analysis 15 Oct 2024 · 0 repositories · arXiv:2410.11756
-
From promise to practice: realizing high-performance decentralized training 15 Oct 2024 · 2 repositories · arXiv:2410.11998
-
Holistic Reasoning with Long-Context LMs: A Benchmark for Database Operations on Massive Textual Data 15 Oct 2024 · 0 repositories · arXiv:2410.11996
-
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection 15 Oct 2024 · 0 repositories · arXiv:2410.11230
-
Jigsaw Puzzles: Splitting Harmful Questions to Jailbreak Large Language Models 15 Oct 2024 · 1 repository · arXiv:2410.11459
-
Meta-DT: Offline Meta-RL as Conditional Sequence Modeling with World Model Disentanglement 15 Oct 2024 · 1 repository · arXiv:2410.11448Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions 15 Oct 2024 · 0 repositories · arXiv:2410.11833
-
MoH: Multi-Head Attention as Mixture-of-Head Attention 15 Oct 2024 · 3 repositories · arXiv:2410.11842Syntology official (archive's flag): 3 ran · 14 ran (of which 5 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 8 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples) · 3 pointer-only (licence)
-
MTU-Bench: A Multi-granularity Tool-Use Benchmark for Large Language Models 15 Oct 2024 · 1 repository · arXiv:2410.11710
-
Multiview Scene Graph 15 Oct 2024 · 1 repository · arXiv:2410.11187Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 20 unverified (of 37 harvested samples) · 37 pointer-only (licence)
-
Nonlinear Gaussian process tomography with imposed non-negativity constraints on physical quantities for plasma diagnostics 15 Oct 2024 · 0 repositories · arXiv:2410.11454
-
On the Capacity of Citation Generation by Large Language Models 15 Oct 2024 · 0 repositories · arXiv:2410.11217
-
Pixology: Probing the Linguistic and Visual Capabilities of Pixel-based Language Models 15 Oct 2024 · 1 repository · arXiv:2410.12011Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
ReDeEP: Detecting Hallucination in Retrieval-Augmented Generation via Mechanistic Interpretability 15 Oct 2024 · 0 repositories · arXiv:2410.11414
-
Rethinking Graph Transformer Architecture Design for Node Classification 15 Oct 2024 · 0 repositories · arXiv:2410.11189
-
Retrieval Augmented Spelling Correction for E-Commerce Applications 15 Oct 2024 · 0 repositories · arXiv:2410.11655
-
RuleRAG: Rule-guided retrieval-augmented generation with language models for question answering 15 Oct 2024 · 1 repository · arXiv:2410.22353
-
SeaDATE: Remedy Dual-Attention Transformer with Semantic Alignment via Contrast Learning for Multimodal Object Detection 15 Oct 2024 · 0 repositories · arXiv:2410.11358
-
SEER: Self-Aligned Evidence Extraction for Retrieval-Augmented Generation 15 Oct 2024 · 0 repositories · arXiv:2410.11315
-
Selection-p: Self-Supervised Task-Agnostic Prompt Compression for Faithfulness and Transferability 15 Oct 2024 · 0 repositories · arXiv:2410.11786
-
Self-adaptive Multimodal Retrieval-Augmented Generation 15 Oct 2024 · 1 repository · arXiv:2410.11321