Methods › General › Stochastic Optimization › Adam › Papers, page 101
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 101 of 244: papers 10,001 to 10,100 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Accumulating Word Representations in Multi-level Context Integration for ERC Task 6 Nov 2023 · 1 repository
-
Adapting Pre-trained Generative Models for Extractive Question Answering 6 Nov 2023 · 0 repositories · arXiv:2311.02961
-
Can LLMs Follow Simple Rules? 6 Nov 2023 · 1 repository · arXiv:2311.04235Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
DeepInception: Hypnotize Large Language Model to Be Jailbreaker 6 Nov 2023 · 1 repository · arXiv:2311.03191Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
In-Context Learning for Knowledge Base Question Answering for Unmanned Systems based on Large Language Models 6 Nov 2023 · 0 repositories · arXiv:2311.02956
-
Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch 6 Nov 2023 · 3 repositories · arXiv:2311.03099Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Leveraging Transformers to Improve Breast Cancer Classification and Risk Assessment with Multi-modal and Longitudinal Data 6 Nov 2023 · 0 repositories · arXiv:2311.03217
-
Machine Learning-Based Tea Leaf Disease Detection: A Comprehensive Review 6 Nov 2023 · 0 repositories · arXiv:2311.03240
-
Nexus at ArAIEval Shared Task: Fine-Tuning Arabic Language Models for Propaganda and Disinformation Detection 6 Nov 2023 · 0 repositories · arXiv:2311.03184
-
p-Laplacian Transformer 6 Nov 2023 · 0 repositories · arXiv:2311.03235
-
Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation 6 Nov 2023 · 0 repositories · arXiv:2311.03348
-
SugarViT -- Multi-objective Regression of UAV Images with Vision Transformers and Deep Label Distribution Learning Demonstrated on Disease Severity Prediction in Sugar Beet 6 Nov 2023 · 0 repositories · arXiv:2311.03076
-
TSP-Transformer: Task-Specific Prompts Boosted Transformer for Holistic Scene Understanding 6 Nov 2023 · 1 repository · arXiv:2311.03427
-
Unraveling Downstream Gender Bias from Large Language Models: A Study on AI Educational Writing Assistance 6 Nov 2023 · 1 repository · arXiv:2311.03311
-
A Critical Perceptual Pre-trained Model for Complex Trajectory Recovery 5 Nov 2023 · 0 repositories · arXiv:2311.02631
-
Attention Modules Improve Image-Level Anomaly Detection for Industrial Inspection: A DifferNet Case Study 5 Nov 2023 · 1 repository · arXiv:2311.02747
-
AI-TA: Towards an Intelligent Question-Answer Teaching Assistant using Open-Source LLMs 5 Nov 2023 · 1 repository · arXiv:2311.02775
-
Evaluating the Potential of Leading Large Language Models in Reasoning Biology Questions 5 Nov 2023 · 0 repositories · arXiv:2311.07582
-
GPT-4V-AD: Exploring Grounding Potential of VQA-oriented GPT-4V for Zero-shot Anomaly Detection 5 Nov 2023 · 1 repository · arXiv:2311.02612
-
Extraction of Atypical Aspects from Customer Reviews: Datasets and Experiments with Language Models 5 Nov 2023 · 2 repositories · arXiv:2311.02702
-
FloodBrain: Flood Disaster Reporting by Web-based Retrieval Augmented Generation with an LLM 5 Nov 2023 · 0 repositories · arXiv:2311.02597
-
Large language models implicitly learn to straighten neural sentence trajectories to construct a predictive representation of natural language 5 Nov 2023 · 0 repositories · arXiv:2311.04930
-
UID as a Guiding Metric for Automated Authorship Obfuscation 5 Nov 2023 · 0 repositories · arXiv:2312.03709
-
MFTCoder: Boosting Code LLMs with Multitask Fine-Tuning 4 Nov 2023 · 1 repository · arXiv:2311.02303
-
Ultra-Long Sequence Distributed Transformer 4 Nov 2023 · 0 repositories · arXiv:2311.02382
-
Understanding the Natural Language of DNA using Encoder-Decoder Foundation Models with Byte-level Precision 4 Nov 2023 · 0 repositories · arXiv:2311.02333
-
You Only Forward Once: Prediction and Rationalization in A Single Forward Pass 4 Nov 2023 · 0 repositories · arXiv:2311.02344
-
An Empirical Study of Benchmarking Chinese Aspect Sentiment Quad Prediction 3 Nov 2023 · 0 repositories · arXiv:2311.01713
-
Automating Governing Knowledge Commons and Contextual Integrity (GKC-CI) Privacy Policy Annotations with Large Language Models 3 Nov 2023 · 1 repository · arXiv:2311.02192
-
Capturing Local and Global Features in Medical Images by Using Ensemble CNN-Transformer 3 Nov 2023 · 0 repositories · arXiv:2311.01731
-
COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning 3 Nov 2023 · 0 repositories · arXiv:2311.02248
-
Data-Free Distillation of Language Model by Text-to-Text Transfer 3 Nov 2023 · 0 repositories · arXiv:2311.01689
-
Depth-guided Free-space Segmentation for a Mobile Robot 3 Nov 2023 · 0 repositories · arXiv:2311.01966
-
DialogBench: Evaluating LLMs as Human-like Dialogue Systems 3 Nov 2023 · 1 repository · arXiv:2311.01677
-
Efficient Black-Box Adversarial Attacks on Neural Text Detectors 3 Nov 2023 · 1 repository · arXiv:2311.01873
-
Emergence of Abstract State Representations in Embodied Sequence Modeling 3 Nov 2023 · 0 repositories · arXiv:2311.02171
-
Epidemic Decision-making System Based Federated Reinforcement Learning 3 Nov 2023 · 0 repositories · arXiv:2311.01749
-
Exploring the Numerical Reasoning Capabilities of Language Models: A Comprehensive Analysis on Tabular Data 3 Nov 2023 · 0 repositories · arXiv:2311.02216
-
FaMeSumm: Investigating and Improving Faithfulness of Medical Summarization 3 Nov 2023 · 1 repository · arXiv:2311.02271Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
GateLoop: Fully Data-Controlled Linear Recurrence for Sequence Modeling 3 Nov 2023 · 3 repositories · arXiv:2311.01927Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
High Probability Convergence of Adam Under Unbounded Gradients and Affine Variance Noise 3 Nov 2023 · 0 repositories · arXiv:2311.02000
-
Multi-scale Time-stepping of Partial Differential Equations with Transformers 3 Nov 2023 · 1 repository · arXiv:2311.02225
-
PPTC Benchmark: Evaluating Large Language Models for PowerPoint Task Completion 3 Nov 2023 · 1 repository · arXiv:2311.01767Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples)
-
Simplifying Transformer Blocks 3 Nov 2023 · 1 repository · arXiv:2311.01906
-
The Potential of Wearable Sensors for Assessing Patient Acuity in Intensive Care Unit (ICU) 3 Nov 2023 · 0 repositories · arXiv:2311.02251
-
The risks of risk-based AI regulation: taking liability seriously 3 Nov 2023 · 0 repositories · arXiv:2311.14684
-
Towards a Unified Transformer-based Framework for Scene Graph Generation and Human-object Interaction Detection 3 Nov 2023 · 0 repositories · arXiv:2311.01755
-
ATGNN: Audio Tagging Graph Neural Network 2 Nov 2023 · 0 repositories · arXiv:2311.01526
-
Deep Double Descent for Time Series Forecasting: Avoiding Undertrained Models 2 Nov 2023 · 0 repositories · arXiv:2311.01442
-
Distilling Knowledge from CNN-Transformer Models for Enhanced Human Action Recognition 2 Nov 2023 · 0 repositories · arXiv:2311.01283
-
Efficient Vision Transformer for Accurate Traffic Sign Detection 2 Nov 2023 · 0 repositories · arXiv:2311.01429
-
Enriching Phrases with Coupled Pixel and Object Contexts for Panoptic Narrative Grounding 2 Nov 2023 · 0 repositories · arXiv:2311.01091
-
Generative Input: Towards Next-Generation Input Methods Paradigm 2 Nov 2023 · 0 repositories · arXiv:2311.01166
-
Hybrid-Fusion Transformer for Multisequence MRI 2 Nov 2023 · 1 repository · arXiv:2311.01308
-
Copilot4D: Learning Unsupervised World Models for Autonomous Driving via Discrete Diffusion 2 Nov 2023 · 0 repositories · arXiv:2311.01017
-
Long Story Short: a Summarize-then-Search Method for Long Video Question Answering 2 Nov 2023 · 1 repository · arXiv:2311.01233
-
M&M3D: Multi-Dataset Training and Efficient Network for Multi-view 3D Object Detection 2 Nov 2023 · 1 repository · arXiv:2311.00986
-
Measuring Five Accountable Talk Moves to Improve Instruction at Scale 2 Nov 2023 · 0 repositories · arXiv:2311.10749
-
On the Convergence of Encoder-only Shallow Transformers 2 Nov 2023 · 0 repositories · arXiv:2311.01575
-
Scattering Vision Transformer: Spectral Mixing Matters 2 Nov 2023 · 0 repositories · arXiv:2311.01310
-
Server-side Rescoring of Spoken Entity-centric Knowledge Queries for Virtual Assistants 2 Nov 2023 · 0 repositories · arXiv:2311.01398
-
Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model 2 Nov 2023 · 1 repository · arXiv:2311.00968Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples)
-
1DFormer: a Transformer Architecture Learning 1D Landmark Representations for Facial Landmark Tracking 1 Nov 2023 · 0 repositories · arXiv:2311.00241
-
A Spatial-Temporal Transformer based Framework For Human Pose Assessment And Correction in Education Scenarios 1 Nov 2023 · 0 repositories · arXiv:2311.00401
-
An Improved Transformer-based Model for Detecting Phishing, Spam, and Ham: A Large Language Model Approach 1 Nov 2023 · 0 repositories · arXiv:2311.04913
-
Are Large Language Models Reliable Judges? A Study on the Factuality Evaluation Capabilities of LLMs 1 Nov 2023 · 0 repositories · arXiv:2311.00681
-
Attention Alignment and Flexible Positional Embeddings Improve Transformer Length Extrapolation 1 Nov 2023 · 0 repositories · arXiv:2311.00684
-
Can Large Language Models Capture Public Opinion about Global Warming? An Empirical Assessment of Algorithmic Fidelity and Bias 1 Nov 2023 · 0 repositories · arXiv:2311.00217
-
Continuous Training and Fine-tuning for Domain-Specific Language Models in Medical Question Answering 1 Nov 2023 · 0 repositories · arXiv:2311.00204
-
COSTAR: Improved Temporal Counterfactual Estimation with Self-Supervised Learning 1 Nov 2023 · 1 repository · arXiv:2311.00886
-
Detecting Visual Cues in the Intensive Care Unit and Association with Patient Clinical Status 1 Nov 2023 · 0 repositories · arXiv:2311.00565
-
Entity Alignment Method of Science and Technology Patent based on Graph Convolution Network and Information Fusion 1 Nov 2023 · 0 repositories · arXiv:2311.00300
-
From Text to Structure: Using Large Language Models to Support the Development of Legal Expert Systems 1 Nov 2023 · 1 repository · arXiv:2311.04911
-
Improving Robustness for Vision Transformer with a Simple Dynamic Scanning Augmentation 1 Nov 2023 · 0 repositories · arXiv:2311.00441
-
Is GPT Powerful Enough to Analyze the Emotions of Memes? 1 Nov 2023 · 0 repositories · arXiv:2311.00223
-
Rethinking Decision Transformer via Hierarchical Reinforcement Learning 1 Nov 2023 · 0 repositories · arXiv:2311.00267
-
Syntactic Inductive Bias in Transformer Language Models: Especially Helpful for Low-Resource Languages? 1 Nov 2023 · 1 repository · arXiv:2311.00268
-
Advances in Embodied Navigation Using Large Language Models: A Survey 1 Nov 2023 · 1 repository · arXiv:2311.00530
-
Unsupervised Lexical Simplification with Context Augmentation 1 Nov 2023 · 1 repository · arXiv:2311.00310
-
A Systematic Review for Transformer-based Long-term Series Forecasting 31 Oct 2023 · 0 repositories · arXiv:2310.20218
-
BERTwich: Extending BERT's Capabilities to Model Dialectal and Noisy Text 31 Oct 2023 · 0 repositories · arXiv:2311.00116
-
Breaking the Token Barrier: Chunking and Convolution for Efficient Long Text Classification with BERT 31 Oct 2023 · 0 repositories · arXiv:2310.20558
-
Breathing Life into Faces: Speech-driven 3D Facial Animation with Natural Head Pose and Detailed Shape 31 Oct 2023 · 0 repositories · arXiv:2310.20240
-
Causal Interpretation of Self-Attention in Pre-Trained Transformers 31 Oct 2023 · 1 repository · arXiv:2310.20307
-
ChipNeMo: Domain-Adapted LLMs for Chip Design 31 Oct 2023 · 0 repositories · arXiv:2311.00176
-
Diversified Node Sampling based Hierarchical Transformer Pooling for Graph Representation Learning 31 Oct 2023 · 0 repositories · arXiv:2310.20250
-
Do large language models solve verbal analogies like children do? 31 Oct 2023 · 0 repositories · arXiv:2310.20384
-
Does GPT-4 pass the Turing test? 31 Oct 2023 · 0 repositories · arXiv:2310.20216
-
EELBERT: Tiny Models through Dynamic Embeddings 31 Oct 2023 · 0 repositories · arXiv:2310.20144
-
Efficient Classification of Student Help Requests in Programming Courses Using Large Language Models 31 Oct 2023 · 0 repositories · arXiv:2310.20105
-
FA Team at the NTCIR-17 UFO Task 31 Oct 2023 · 0 repositories · arXiv:2310.20322
-
GAR-meets-RAG Paradigm for Zero-Shot Information Retrieval 31 Oct 2023 · 0 repositories · arXiv:2310.20158
-
Generate What You Prefer: Reshaping Sequential Recommendation via Guided Diffusion 31 Oct 2023 · 1 repository · arXiv:2310.20453
-
Global Transformer Architecture for Indoor Room Temperature Forecasting 31 Oct 2023 · 0 repositories · arXiv:2310.20476
-
GraphTransformers for Geospatial Forecasting of Hurricane Trajectories 31 Oct 2023 · 0 repositories · arXiv:2310.20174
-
In Search of Lost Online Test-time Adaptation: A Survey 31 Oct 2023 · 1 repository · arXiv:2310.20199Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 7 pointer-only (licence)
-
Increasing The Performance of Cognitively Inspired Data-Efficient Language Models via Implicit Structure Building 31 Oct 2023 · 1 repository · arXiv:2310.20589
-
Interactive Multi-fidelity Learning for Cost-effective Adaptation of Language Model with Sparse Human Supervision 31 Oct 2023 · 0 repositories · arXiv:2310.20153
-
Learning From Mistakes Makes LLM Better Reasoner 31 Oct 2023 · 1 repository · arXiv:2310.20689Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
PsyCoT: Psychological Questionnaire as Powerful Chain-of-Thought for Personality Detection 31 Oct 2023 · 1 repository · arXiv:2310.20256