Methods › General › Stochastic Optimization › Adam › Papers, page 127
Adam
Papers archive 2025-07-28
archive papers tagged: 24,390 · with a code link: 10,944 · where Syntology ran a sample: 3,424 (2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,424 of 24,390 tagged: 2,899 with a run with no instrument failure, 525 where every run was a failure of Syntology's instrument)
Page 127 of 244: papers 12,601 to 12,700 of 24,390, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fine-tuned LLMs Know More, Hallucinate Less with Few-Shot Sequence-to-Sequence Semantic Parsing over Wikidata 23 May 2023 · 1 repository · arXiv:2305.14202
-
Condensing Multilingual Knowledge with Lightweight Language-Specific Modules 23 May 2023 · 1 repository · arXiv:2305.13993Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Connecting the Dots: What Graph-Based Text Representations Work Best for Text Classification Using Graph Neural Networks? 23 May 2023 · 1 repository · arXiv:2305.14578
-
Cross-Attention is Not Enough: Incongruity-Aware Dynamic Hierarchical Fusion for Multimodal Affect Recognition 23 May 2023 · 1 repository · arXiv:2305.13583
-
Dancing Between Success and Failure: Edit-level Simplification Evaluation using SALSA 23 May 2023 · 0 repositories · arXiv:2305.14458
-
Deduction under Perturbed Evidence: Probing Student Simulation Capabilities of Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.14507
-
Detecting automatically the layout of clinical documents to enhance the performances of downstream natural language processing 23 May 2023 · 0 repositories · arXiv:2305.13817
-
Dynosaur: A Dynamic Growth Paradigm for Instruction-Tuning Data Curation 23 May 2023 · 1 repository · arXiv:2305.14327Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
Eeg2vec: Self-Supervised Electroencephalographic Representation Learning 23 May 2023 · 0 repositories · arXiv:2305.13957
-
Few-Shot Data Synthesis for Open Domain Multi-Hop Question Answering 23 May 2023 · 0 repositories · arXiv:2305.13691
-
Benchmarking Machine Translation with Cultural Awareness 23 May 2023 · 1 repository · arXiv:2305.14328
-
Enhancing Black-Box Few-Shot Text Classification with Prompt-Based Data Augmentation 23 May 2023 · 0 repositories · arXiv:2305.13785
-
Advancing Precise Outline-Conditioned Text Generation with Task Duality and Explicit Outline Control 23 May 2023 · 0 repositories · arXiv:2305.14459
-
Evaluating Factual Consistency of Summaries with Large Language Models 23 May 2023 · 2 repositories · arXiv:2305.14069
-
Exploring Large Language Models for Classical Philology 23 May 2023 · 1 repository · arXiv:2305.13698
-
FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation 23 May 2023 · 4 repositories · arXiv:2305.14251Syntology official (archive's flag): 10 ran · 11 ran (of which 1 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples) · 6 pointer-only (licence)
-
From Characters to Words: Hierarchical Pre-trained Language Model for Open-vocabulary Language Understanding 23 May 2023 · 0 repositories · arXiv:2305.14571
-
GenSpectrum Chat: Data Exploration in Public Health Using Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.13821
-
Goat: Fine-tuned LLaMA Outperforms GPT-4 on Arithmetic Tasks 23 May 2023 · 1 repository · arXiv:2305.14201
-
Handling Realistic Label Noise in BERT Text Classification 23 May 2023 · 0 repositories · arXiv:2305.16337
-
HumBEL: A Human-in-the-Loop Approach for Evaluating Demographic Factors of Language Models in Human-Machine Conversations 23 May 2023 · 1 repository · arXiv:2305.14195
-
IfQA: A Dataset for Open-domain Question Answering under Counterfactual Presuppositions 23 May 2023 · 0 repositories · arXiv:2305.14010
-
Images in Language Space: Exploring the Suitability of Large Language Models for Vision & Language Tasks 23 May 2023 · 1 repository · arXiv:2305.13782
-
INSTRUCTSCORE: Explainable Text Generation Evaluation with Finegrained Feedback 23 May 2023 · 2 repositories · arXiv:2305.14282
-
SciMON: Scientific Inspiration Machines Optimized for Novelty 23 May 2023 · 1 repository · arXiv:2305.14259
-
Let's Think Frame by Frame with VIP: A Video Infilling and Prediction Dataset for Evaluating Video Chain-of-Thought 23 May 2023 · 1 repository · arXiv:2305.13903
-
LLM-powered Data Augmentation for Enhanced Cross-lingual Performance 23 May 2023 · 1 repository · arXiv:2305.14288
-
LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond 23 May 2023 · 1 repository · arXiv:2305.14540
-
MathDial: A Dialogue Tutoring Dataset with Rich Pedagogical Properties Grounded in Math Reasoning Problems 23 May 2023 · 1 repository · arXiv:2305.14536Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
NAIL: Lexical Retrieval Indices with Efficient Non-Autoregressive Decoders 23 May 2023 · 0 repositories · arXiv:2305.14499
-
NarrativeXL: A Large-scale Dataset For Long-Term Memory Models 23 May 2023 · 1 repository · arXiv:2305.13877
-
On Robustness of Finetuned Transformer-based NLP Models 23 May 2023 · 1 repository · arXiv:2305.14453
-
On Structural Expressive Power of Graph Transformers 23 May 2023 · 0 repositories · arXiv:2305.13987
-
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification 23 May 2023 · 1 repository · arXiv:2305.14032Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Physics of Language Models: Part 1, Learning Hierarchical Language Structures 23 May 2023 · 0 repositories · arXiv:2305.13673
-
Pre-training Multi-task Contrastive Learning Models for Scientific Literature Understanding 23 May 2023 · 0 repositories · arXiv:2305.14232
-
Probing Brain Context-Sensitivity with Masked-Attention Generation 23 May 2023 · 0 repositories · arXiv:2305.13863
-
QLoRA: Efficient Finetuning of Quantized LLMs 23 May 2023 · 20 repositories · arXiv:2305.14314Syntology official (archive's flag): 3 ran · 18 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 2 violated, 2 with no contract checked; 12 where Syntology's instrument failed) · 8 unverified (of 26 harvested samples) · 17 pointer-only (licence)
-
Rethinking Speech Recognition with A Multimodal Perspective via Acoustic and Semantic Cooperative Decoding 23 May 2023 · 0 repositories · arXiv:2305.14049
-
Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training 23 May 2023 · 7 repositories · arXiv:2305.14342Syntology 13 ran (of which 4 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 19 harvested samples)
-
Source-Free Domain Adaptation for RGB-D Semantic Segmentation with Vision Transformers 23 May 2023 · 0 repositories · arXiv:2305.14269
-
Sources of Hallucination by Large Language Models on Inference Tasks 23 May 2023 · 1 repository · arXiv:2305.14552
-
Text Is All You Need: Learning Language Representations for Sequential Recommendation 23 May 2023 · 1 repository · arXiv:2305.13731Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards A Unified View of Sparse Feed-Forward Network in Pretraining Large Language Model 23 May 2023 · 0 repositories · arXiv:2305.13999
-
ReadMe++: Benchmarking Multilingual Language Models for Multi-Domain Readability Assessment 23 May 2023 · 1 repository · arXiv:2305.14463
-
Training Transitive and Commutative Multimodal Transformers with LoReTTa 23 May 2023 · 0 repositories · arXiv:2305.14243
-
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs 23 May 2023 · 0 repositories · arXiv:2305.14279
-
Understanding and Improving Optimization in Predictive Coding Networks 23 May 2023 · 1 repository · arXiv:2305.13562
-
VDD: Varied Drone Dataset for Semantic Segmentation 23 May 2023 · 1 repository · arXiv:2305.13608
-
WikiChat: Stopping the Hallucination of Large Language Model Chatbots by Few-Shot Grounding on Wikipedia 23 May 2023 · 1 repository · arXiv:2305.14292Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
ZeroSCROLLS: A Zero-Shot Benchmark for Long Text Understanding 23 May 2023 · 1 repository · arXiv:2305.14196Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Let GPT be a Math Tutor: Teaching Math Word Problem Solvers with Customized Exercise Generation 22 May 2023 · 0 repositories · arXiv:2305.14386
-
A Study of Generative Large Language Model for Medical Research and Healthcare 22 May 2023 · 1 repository · arXiv:2305.13523
-
Adaptive action supervision in reinforcement learning from real-world multi-agent demonstrations 22 May 2023 · 0 repositories · arXiv:2305.13030
-
Large Language Models are Not Yet Human-Level Evaluators for Abstractive Summarization 22 May 2023 · 1 repository · arXiv:2305.13091Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Beneath Surface Similarity: Large Language Models Make Reasonable Scientific Analogies after Structure Abduction 22 May 2023 · 1 repository · arXiv:2305.12660
-
Bidirectional Transformer Reranker for Grammatical Error Correction 22 May 2023 · 1 repository · arXiv:2305.13000
-
Can ChatGPT Defend its Belief in Truth? Evaluating LLM Reasoning via Debate 22 May 2023 · 0 repositories · arXiv:2305.13160
-
Can Large Language Models emulate an inductive Thematic Analysis of semi-structured interviews? An exploration and provocation on the limits of the approach and the model 22 May 2023 · 1 repository · arXiv:2305.13014
-
Can LLMs facilitate interpretation of pre-trained language models? 22 May 2023 · 0 repositories · arXiv:2305.13386
-
Cognitive network science reveals bias in GPT-3, ChatGPT, and GPT-4 mirroring math anxiety in high-school students 22 May 2023 · 0 repositories · arXiv:2305.18320
-
A 4D Hybrid Algorithm to Scale Parallel Training to Thousands of GPUs 22 May 2023 · 1 repository · arXiv:2305.13525
-
Table Meets LLM: Can Large Language Models Understand Structured Table Data? A Benchmark and Empirical Study 22 May 2023 · 1 repository · arXiv:2305.13062
-
ExplainCPE: A Free-text Explanation Benchmark of Chinese Pharmacist Examination 22 May 2023 · 1 repository · arXiv:2305.12945
-
Exploring Energy-based Language Models with Different Architectures and Training Methods for Speech Recognition 22 May 2023 · 2 repositories · arXiv:2305.12676
-
Flover: A Temporal Fusion Framework for Efficient Autoregressive Model Parallel Inference 22 May 2023 · 1 repository · arXiv:2305.13484
-
G3Detector: General GPT-Generated Text Detector 22 May 2023 · 0 repositories · arXiv:2305.12680
-
GATology for Linguistics: What Syntactic Dependencies It Knows 22 May 2023 · 0 repositories · arXiv:2305.13403
-
GNCformer Enhanced Self-attention for Automatic Speech Recognition 22 May 2023 · 0 repositories · arXiv:2305.12755
-
How Language Model Hallucinations Can Snowball 22 May 2023 · 1 repository · arXiv:2305.13534
-
SimCSE++: Improving Contrastive Learning for Sentence Embeddings from Two Perspectives 22 May 2023 · 0 repositories · arXiv:2305.13192
-
InheritSumm: A General, Versatile and Compact Summarizer by Distilling from GPT 22 May 2023 · 0 repositories · arXiv:2305.13083
-
Iterative Forward Tuning Boosts In-Context Learning in Language Models 22 May 2023 · 1 repository · arXiv:2305.13016
-
Language-Agnostic Bias Detection in Language Models with Bias Probing 22 May 2023 · 1 repository · arXiv:2305.13302
-
Learning Pedestrian Actions to Ensure Safe Autonomous Driving 22 May 2023 · 0 repositories · arXiv:2305.13051
-
Learning Subpocket Prototypes for Generalizable Structure-based Drug Design 22 May 2023 · 1 repository · arXiv:2305.13997Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
LLMs for Knowledge Graph Construction and Reasoning: Recent Capabilities and Future Opportunities 22 May 2023 · 1 repository · arXiv:2305.13168Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Atomic Inference for NLI with Generated Facts as Atoms 22 May 2023 · 1 repository · arXiv:2305.13214
-
MAILEX: Email Event and Argument Extraction 22 May 2023 · 1 repository · arXiv:2305.13469
-
Measuring Inductive Biases of In-Context Learning with Underspecified Demonstrations 22 May 2023 · 1 repository · arXiv:2305.13299
-
Multi-Task Instruction Tuning of LLaMa for Specific Scenarios: A Preliminary Study on Writing Assistance 22 May 2023 · 0 repositories · arXiv:2305.13225
-
nnDetection for Intracranial Aneurysms Detection and Localization 22 May 2023 · 1 repository · arXiv:2305.13398
-
Investigating the Role of Feed-Forward Networks in Transformers Using Parallel Attention and Feed-Forward Net Design 22 May 2023 · 0 repositories · arXiv:2305.13297
-
RecurrentGPT: Interactive Generation of (Arbitrarily) Long Text 22 May 2023 · 2 repositories · arXiv:2305.13304Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
RWKV: Reinventing RNNs for the Transformer Era 22 May 2023 · 14 repositories · arXiv:2305.13048Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
SCITAB: A Challenging Benchmark for Compositional Reasoning and Claim Verification on Scientific Tables 22 May 2023 · 1 repository · arXiv:2305.13186
-
Stock and market index prediction using Informer network 22 May 2023 · 0 repositories · arXiv:2305.14382
-
Syntactic Knowledge via Graph Attention with BERT in Machine Translation 22 May 2023 · 0 repositories · arXiv:2305.13413
-
The Emergence of Economic Rationality of GPT 22 May 2023 · 0 repositories · arXiv:2305.12763
-
Tokenized Graph Transformer with Neighborhood Augmentation for Node Classification in Large Graphs 22 May 2023 · 0 repositories · arXiv:2305.12677
-
Investigating Agency of LLMs in Human-AI Collaboration Tasks 22 May 2023 · 0 repositories · arXiv:2305.12815
-
VDT: General-purpose Video Diffusion Transformers via Mask Modeling 22 May 2023 · 1 repository · arXiv:2305.13311Syntology official (archive's flag): 4 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
VideoLLM: Modeling Video Sequence with Large Language Models 22 May 2023 · 1 repository · arXiv:2305.13292
-
Why current rain denoising models fail on CycleGAN created rain images in autonomous driving 22 May 2023 · 0 repositories · arXiv:2305.12983
-
A Deeper (Autoregressive) Approach to Non-Convergent Discourse Parsing 21 May 2023 · 0 repositories · arXiv:2305.12510
-
A Symbolic Framework for Evaluating Mathematical Reasoning and Generalisation with Transformers 21 May 2023 · 0 repositories · arXiv:2305.12563
-
BertRLFuzzer: A BERT and Reinforcement Learning Based Fuzzer 21 May 2023 · 1 repository · arXiv:2305.12534
-
BiasAsker: Measuring the Bias in Conversational AI System 21 May 2023 · 1 repository · arXiv:2305.12434
-
Abstract Meaning Representation-Based Logic-Driven Data Augmentation for Logical Reasoning 21 May 2023 · 1 repository · arXiv:2305.12599
-
Evaluating the Performance of Large Language Models on GAOKAO Benchmark 21 May 2023 · 1 repository · arXiv:2305.12474Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)