Methods › General › Stochastic Optimization › Adam › Papers, page 76
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 76 of 244: papers 7,501 to 7,600 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Explainable Deep Learning: A Visual Analytics Approach with Transition Matrices 29 Mar 2024 · 1 repository
-
Latxa: An Open Language Model and Evaluation Suite for Basque 29 Mar 2024 · 1 repository · arXiv:2403.20266
-
LayerNorm: A key component in parameter-efficient fine-tuning 29 Mar 2024 · 0 repositories · arXiv:2403.20284
-
Localising the Seizure Onset Zone from Single-Pulse Electrical Stimulation Responses with a CNN Transformer 29 Mar 2024 · 1 repository · arXiv:2403.20324
-
MANGO: A Benchmark for Evaluating Mapping and Navigation Abilities of Large Language Models 29 Mar 2024 · 1 repository · arXiv:2403.19913Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
On-the-fly Definition Augmentation of LLMs for Biomedical NER 29 Mar 2024 · 1 repository · arXiv:2404.00152Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
ReALM: Reference Resolution As Language Modeling 29 Mar 2024 · 0 repositories · arXiv:2403.20329
-
SceneTracker: Long-term Scene Flow Estimation Network 29 Mar 2024 · 1 repository · arXiv:2403.19924Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Shallow Cross-Encoders for Low-Latency Retrieval 29 Mar 2024 · 1 repository · arXiv:2403.20222
-
A Novel Stochastic Transformer-based Approach for Post-Traumatic Stress Disorder Detection using Audio Recording of Clinical Interviews 28 Mar 2024 · 0 repositories · arXiv:2403.19441
-
A Review of Multi-Modal Large Language and Vision Models 28 Mar 2024 · 0 repositories · arXiv:2404.01322
-
AAPMT: AGI Assessment Through Prompt and Metric Transformer 28 Mar 2024 · 1 repository · arXiv:2403.19101
-
AlloyBERT: Alloy Property Prediction with Large Language Models 28 Mar 2024 · 0 repositories · arXiv:2403.19783
-
Are Large Language Models Good at Utility Judgments? 28 Mar 2024 · 1 repository · arXiv:2403.19216Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Checkpoint Merging via Bayesian Optimization in LLM Pretraining 28 Mar 2024 · 0 repositories · arXiv:2403.19390
-
Code Comparison Tuning for Code Large Language Models 28 Mar 2024 · 0 repositories · arXiv:2403.19121
-
DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs 28 Mar 2024 · 3 repositories · arXiv:2403.19588Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Enhancing Efficiency in Vision Transformer Networks: Design Techniques and Insights 28 Mar 2024 · 0 repositories · arXiv:2403.19882
-
FACTOID: FACtual enTailment fOr hallucInation Detection 28 Mar 2024 · 0 repositories · arXiv:2403.19113
-
Generating Multi-Aspect Queries for Conversational Search 28 Mar 2024 · 0 repositories · arXiv:2403.19302
-
Genetic Quantization-Aware Approximation for Non-Linear Operations in Transformers 28 Mar 2024 · 1 repository · arXiv:2403.19591
-
Intelligent Classification and Personalized Recommendation of E-commerce Products Based on Machine Learning 28 Mar 2024 · 0 repositories · arXiv:2403.19345
-
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models 28 Mar 2024 · 1 repository · arXiv:2403.19521Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Jamba: A Hybrid Transformer-Mamba Language Model 28 Mar 2024 · 3 repositories · arXiv:2403.19887
-
Just-DNA-Seq, open-source personal genomics platform: longevity science for everyone 28 Mar 2024 · 0 repositories · arXiv:2403.19087
-
Keypoint Action Tokens Enable In-Context Imitation Learning in Robotics 28 Mar 2024 · 0 repositories · arXiv:2403.19578
-
MATEval: A Multi-Agent Discussion Framework for Advancing Open-Ended Text Evaluation 28 Mar 2024 · 1 repository · arXiv:2403.19305
-
Risk prediction of pathological gambling on social media 28 Mar 2024 · 0 repositories · arXiv:2403.19358
-
Single-Shared Network with Prior-Inspired Loss for Parameter-Efficient Multi-Modal Imaging Skin Lesion Classification 28 Mar 2024 · 0 repositories · arXiv:2403.19203
-
A Novel Corpus of Annotated Medical Imaging Reports and Information Extraction Results Using BERT-based Language Models 27 Mar 2024 · 1 repository · arXiv:2403.18975
-
A Survey on Large Language Models from Concept to Implementation 27 Mar 2024 · 0 repositories · arXiv:2403.18969
-
AcTED: Automatic Acquisition of Typical Event Duration for Semi-supervised Temporal Commonsense QA 27 Mar 2024 · 0 repositories · arXiv:2403.18504
-
Attention-aware semantic relevance predicting Chinese sentence reading 27 Mar 2024 · 0 repositories · arXiv:2403.18542
-
BioMedLM: A 2.7B Parameter Language Model Trained On Biomedical Text 27 Mar 2024 · 1 repository · arXiv:2403.18421Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
BLADE: Enhancing Black-box Large Language Models with Small Domain-Specific Models 27 Mar 2024 · 0 repositories · arXiv:2403.18365
-
Boosting Conversational Question Answering with Fine-Grained Retrieval-Augmentation and Self-Check 27 Mar 2024 · 0 repositories · arXiv:2403.18243
-
CPR: Retrieval Augmented Generation for Copyright Protection 27 Mar 2024 · 0 repositories · arXiv:2403.18920
-
Cross-domain Fiber Cluster Shape Analysis for Language Performance Cognitive Score Prediction 27 Mar 2024 · 0 repositories · arXiv:2403.19001
-
Evaluating Large Language Models for Health-Related Text Classification Tasks with Public Social Media Data 27 Mar 2024 · 0 repositories · arXiv:2403.19031
-
Faster Convergence for Transformer Fine-tuning with Line Search Methods 27 Mar 2024 · 1 repository · arXiv:2403.18506
-
Cross-System Categorization of Abnormal Traces in Microservice-Based Systems via Meta-Learning 27 Mar 2024 · 0 repositories · arXiv:2403.18998
-
Fourier or Wavelet bases as counterpart self-attention in spikformer for efficient visual classification 27 Mar 2024 · 0 repositories · arXiv:2403.18228
-
Fusion approaches for emotion recognition from speech using acoustic and text-based features 27 Mar 2024 · 0 repositories · arXiv:2403.18635
-
Illicit object detection in X-ray images using Vision Transformers 27 Mar 2024 · 0 repositories · arXiv:2403.19043
-
Improving Line Search Methods for Large Scale Neural Network Training 27 Mar 2024 · 1 repository · arXiv:2403.18519
-
Learning in PINNs: Phase transition, total diffusion, and generalization 27 Mar 2024 · 0 repositories · arXiv:2403.18494
-
LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research Practices 27 Mar 2024 · 0 repositories · arXiv:2403.18173
-
Long-form factuality in large language models 27 Mar 2024 · 3 repositories · arXiv:2403.18802Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
mALBERT: Is a Compact Multilingual BERT Model Still Worth It? 27 Mar 2024 · 0 repositories · arXiv:2403.18338
-
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models 27 Mar 2024 · 2 repositories · arXiv:2403.18814Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
ParCo: Part-Coordinating Text-to-Motion Synthesis 27 Mar 2024 · 1 repository · arXiv:2403.18512
-
RankMamba: Benchmarking Mamba's Document Ranking Performance in the Era of Transformers 27 Mar 2024 · 1 repository · arXiv:2403.18276
-
Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers 27 Mar 2024 · 1 repository · arXiv:2403.18938
-
SemRoDe: Macro Adversarial Training to Learn Representations That are Robust to Word-Level Attacks 27 Mar 2024 · 1 repository · arXiv:2403.18423Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
ViTAR: Vision Transformer with Any Resolution 27 Mar 2024 · 0 repositories · arXiv:2403.18361
-
Vulnerability Detection with Code Language Models: How Far Are We? 27 Mar 2024 · 1 repository · arXiv:2403.18624Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching 26 Mar 2024 · 0 repositories · arXiv:2403.17312
-
Are Compressed Language Models Less Subgroup Robust? 26 Mar 2024 · 1 repository · arXiv:2403.17811
-
Automated Report Generation for Lung Cytological Images Using a CNN Vision Classifier and Multiple-Transformer Text Decoders: Preliminary Study 26 Mar 2024 · 0 repositories · arXiv:2403.18151
-
CCDSReFormer: Traffic Flow Prediction with a Criss-Crossed Dual-Stream Enhanced Rectified Transformer Model 26 Mar 2024 · 0 repositories · arXiv:2403.17753
-
Constructions Are So Difficult That Even Large Language Models Get Them Right for the Wrong Reasons 26 Mar 2024 · 1 repository · arXiv:2403.17760
-
Decoding Probing: Revealing Internal Linguistic Structures in Neural Language Models using Minimal Pairs 26 Mar 2024 · 0 repositories · arXiv:2403.17299
-
Disambiguate Entity Matching using Large Language Models through Relation Discovery 26 Mar 2024 · 0 repositories · arXiv:2403.17344
-
Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization 26 Mar 2024 · 1 repository · arXiv:2403.18120Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation 26 Mar 2024 · 1 repository · arXiv:2403.18080
-
ELLEN: Extremely Lightly Supervised Learning For Efficient Named Entity Recognition 26 Mar 2024 · 1 repository · arXiv:2403.17385
-
Enhancing Legal Document Retrieval: A Multi-Phase Approach with Large Language Models 26 Mar 2024 · 0 repositories · arXiv:2403.18093
-
Evaluating the Efficacy of Prompt-Engineered Large Multimodal Models Versus Fine-Tuned Vision Transformers in Image-Based Security Applications 26 Mar 2024 · 0 repositories · arXiv:2403.17787
-
Fingerprinting web servers through Transformer-encoded HTTP response headers 26 Mar 2024 · 1 repository · arXiv:2404.00056
-
Hierarchical Light Transformer Ensembles for Multimodal Trajectory Forecasting 26 Mar 2024 · 1 repository · arXiv:2403.17678
-
Hierarchical Multi-label Classification for Fine-level Event Extraction from Aviation Accident Reports 26 Mar 2024 · 0 repositories · arXiv:2403.17914
-
Integrative Graph-Transformer Framework for Histopathology Whole Slide Image Representation and Classification 26 Mar 2024 · 0 repositories · arXiv:2403.18134
-
InternLM2 Technical Report 26 Mar 2024 · 3 repositories · arXiv:2403.17297Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Large Language Models Are State-of-the-Art Evaluator for Grammatical Error Correction 26 Mar 2024 · 0 repositories · arXiv:2403.17540
-
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution 26 Mar 2024 · 0 repositories · arXiv:2403.17927
-
Mechanistic Design and Scaling of Hybrid Architectures 26 Mar 2024 · 1 repository · arXiv:2403.17844Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Not All Similarities Are Created Equal: Leveraging Data-Driven Biases to Inform GenAI Copyright Disputes 26 Mar 2024 · 0 repositories · arXiv:2403.17691
-
OmniVid: A Generative Framework for Universal Video Understanding 26 Mar 2024 · 1 repository · arXiv:2403.17935Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Rotate to Scan: UNet-like Mamba with Triplet SSM Module for Medical Image Segmentation 26 Mar 2024 · 0 repositories · arXiv:2403.17701
-
SGHormer: An Energy-Saving Graph Transformer Driven by Spikes 26 Mar 2024 · 1 repository · arXiv:2403.17656
-
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic 26 Mar 2024 · 1 repository · arXiv:2403.17933
-
Supervisory Prompt Training 26 Mar 2024 · 0 repositories · arXiv:2403.18051
-
Targeted Visualization of the Backbone of Encoder LLMs 26 Mar 2024 · 1 repository · arXiv:2403.18872
-
Verbing Weirds Language (Models): Evaluation of English Zero-Derivation in Five LLMs 26 Mar 2024 · 0 repositories · arXiv:2403.17856
-
3D-EffiViTCaps: 3D Efficient Vision Transformer with Capsule for Medical Image Segmentation 25 Mar 2024 · 1 repository · arXiv:2403.16350
-
A comparative analysis of embedding models for patent similarity 25 Mar 2024 · 0 repositories · arXiv:2403.16630
-
A comparison of Human, GPT-3.5, and GPT-4 Performance in a University-Level Coding Course 25 Mar 2024 · 1 repository · arXiv:2403.16977
-
A Study on How Attention Scores in the BERT Model are Aware of Lexical Categories in Syntactic and Semantic Tasks on the GLUE Benchmark 25 Mar 2024 · 0 repositories · arXiv:2403.16447
-
An End-to-End Structure with Novel Position Mechanism and Improved EMD for Stock Forecasting 25 Mar 2024 · 1 repository · arXiv:2404.07969
-
ChebMixer: Efficient Graph Representation Learning with MLP Mixer 25 Mar 2024 · 0 repositories · arXiv:2403.16358
-
CT-Bound: Robust Boundary Detection From Noisy Images Via Hybrid Convolution and Transformer Neural Networks 25 Mar 2024 · 1 repository · arXiv:2403.16494
-
CVT-xRF: Contrastive In-Voxel Transformer for 3D Consistent Radiance Fields from Sparse Inputs 25 Mar 2024 · 0 repositories · arXiv:2403.16885
-
Do LLM Agents Have Regret? A Case Study in Online Learning and Games 25 Mar 2024 · 0 repositories · arXiv:2403.16843
-
DOCTR: Disentangled Object-Centric Transformer for Point Scene Understanding 25 Mar 2024 · 1 repository · arXiv:2403.16431
-
Text Understanding in GPT-4 vs Humans 25 Mar 2024 · 0 repositories · arXiv:2403.17196
-
Grammatical vs Spelling Error Correction: An Investigation into the Responsiveness of Transformer-based Language Models using BART and MarianMT 25 Mar 2024 · 0 repositories · arXiv:2403.16655
-
Iterative Refinement of Project-Level Code Context for Precise Code Generation with Compiler Feedback 25 Mar 2024 · 1 repository · arXiv:2403.16792
-
Linear Cross-document Event Coreference Resolution with X-AMR 25 Mar 2024 · 1 repository · arXiv:2404.08656
-
LSTTN: A Long-Short Term Transformer-based Spatio-temporal Neural Network for Traffic Flow Forecasting 25 Mar 2024 · 1 repository · arXiv:2403.16495
-
ModeTv2: GPU-accelerated Motion Decomposition Transformer for Pairwise Optimization in Medical Image Registration 25 Mar 2024 · 2 repositories · arXiv:2403.16526