Methods › General › Output Functions › Softmax › Papers, page 214
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 214 of 375: papers 21,301 to 21,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SciMON: Scientific Inspiration Machines Optimized for Novelty 23 May 2023 · 1 repository · arXiv:2305.14259
-
Let's Think Frame by Frame with VIP: A Video Infilling and Prediction Dataset for Evaluating Video Chain-of-Thought 23 May 2023 · 1 repository · arXiv:2305.13903
-
LLM-powered Data Augmentation for Enhanced Cross-lingual Performance 23 May 2023 · 1 repository · arXiv:2305.14288
-
LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond 23 May 2023 · 1 repository · arXiv:2305.14540
-
MathDial: A Dialogue Tutoring Dataset with Rich Pedagogical Properties Grounded in Math Reasoning Problems 23 May 2023 · 1 repository · arXiv:2305.14536Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
mmT5: Modular Multilingual Pre-Training Solves Source Language Hallucinations 23 May 2023 · 0 repositories · arXiv:2305.14224
-
NAIL: Lexical Retrieval Indices with Efficient Non-Autoregressive Decoders 23 May 2023 · 0 repositories · arXiv:2305.14499
-
NarrativeXL: A Large-scale Dataset For Long-Term Memory Models 23 May 2023 · 1 repository · arXiv:2305.13877
-
On Robustness of Finetuned Transformer-based NLP Models 23 May 2023 · 1 repository · arXiv:2305.14453
-
On Structural Expressive Power of Graph Transformers 23 May 2023 · 0 repositories · arXiv:2305.13987
-
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification 23 May 2023 · 1 repository · arXiv:2305.14032Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Physics of Language Models: Part 1, Learning Hierarchical Language Structures 23 May 2023 · 0 repositories · arXiv:2305.13673
-
Pre-training Multi-task Contrastive Learning Models for Scientific Literature Understanding 23 May 2023 · 0 repositories · arXiv:2305.14232
-
Probing Brain Context-Sensitivity with Masked-Attention Generation 23 May 2023 · 0 repositories · arXiv:2305.13863
-
QLoRA: Efficient Finetuning of Quantized LLMs 23 May 2023 · 20 repositories · arXiv:2305.14314Syntology official (archive's flag): 3 ran · 18 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 2 violated, 2 with no contract checked; 12 where Syntology's instrument failed) · 8 unverified (of 26 harvested samples) · 17 pointer-only (licence)
-
Rethinking Speech Recognition with A Multimodal Perspective via Acoustic and Semantic Cooperative Decoding 23 May 2023 · 0 repositories · arXiv:2305.14049
-
SMAP: A Novel Heterogeneous Information Framework for Scenario-based Optimal Model Assignment 23 May 2023 · 0 repositories · arXiv:2305.13634
-
Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training 23 May 2023 · 7 repositories · arXiv:2305.14342Syntology 13 ran (of which 4 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 19 harvested samples)
-
Sorted Convolutional Network for Achieving Continuous Rotational Invariance 23 May 2023 · 0 repositories · arXiv:2305.14462
-
Source-Free Domain Adaptation for RGB-D Semantic Segmentation with Vision Transformers 23 May 2023 · 0 repositories · arXiv:2305.14269
-
Sources of Hallucination by Large Language Models on Inference Tasks 23 May 2023 · 1 repository · arXiv:2305.14552
-
Text Is All You Need: Learning Language Representations for Sequential Recommendation 23 May 2023 · 1 repository · arXiv:2305.13731Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards A Unified View of Sparse Feed-Forward Network in Pretraining Large Language Model 23 May 2023 · 0 repositories · arXiv:2305.13999
-
ReadMe++: Benchmarking Multilingual Language Models for Multi-Domain Readability Assessment 23 May 2023 · 1 repository · arXiv:2305.14463
-
Training Transitive and Commutative Multimodal Transformers with LoReTTa 23 May 2023 · 0 repositories · arXiv:2305.14243
-
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs 23 May 2023 · 0 repositories · arXiv:2305.14279
-
VDD: Varied Drone Dataset for Semantic Segmentation 23 May 2023 · 1 repository · arXiv:2305.13608
-
Weakly Supervised 3D Open-vocabulary Segmentation 23 May 2023 · 1 repository · arXiv:2305.14093Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 10 pointer-only (licence)
-
WikiChat: Stopping the Hallucination of Large Language Model Chatbots by Few-Shot Grounding on Wikipedia 23 May 2023 · 1 repository · arXiv:2305.14292Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
ZeroSCROLLS: A Zero-Shot Benchmark for Long Text Understanding 23 May 2023 · 1 repository · arXiv:2305.14196Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Let GPT be a Math Tutor: Teaching Math Word Problem Solvers with Customized Exercise Generation 22 May 2023 · 0 repositories · arXiv:2305.14386
-
A Study of Generative Large Language Model for Medical Research and Healthcare 22 May 2023 · 1 repository · arXiv:2305.13523
-
Adaptive action supervision in reinforcement learning from real-world multi-agent demonstrations 22 May 2023 · 0 repositories · arXiv:2305.13030
-
An Optimized Ensemble Deep Learning Model For Brain Tumor Classification 22 May 2023 · 0 repositories · arXiv:2305.12844
-
Large Language Models are Not Yet Human-Level Evaluators for Abstractive Summarization 22 May 2023 · 1 repository · arXiv:2305.13091Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Beneath Surface Similarity: Large Language Models Make Reasonable Scientific Analogies after Structure Abduction 22 May 2023 · 1 repository · arXiv:2305.12660
-
Bidirectional Transformer Reranker for Grammatical Error Correction 22 May 2023 · 1 repository · arXiv:2305.13000
-
Can ChatGPT Defend its Belief in Truth? Evaluating LLM Reasoning via Debate 22 May 2023 · 0 repositories · arXiv:2305.13160
-
Can Large Language Models emulate an inductive Thematic Analysis of semi-structured interviews? An exploration and provocation on the limits of the approach and the model 22 May 2023 · 1 repository · arXiv:2305.13014
-
Can LLMs facilitate interpretation of pre-trained language models? 22 May 2023 · 0 repositories · arXiv:2305.13386
-
Cognitive network science reveals bias in GPT-3, ChatGPT, and GPT-4 mirroring math anxiety in high-school students 22 May 2023 · 0 repositories · arXiv:2305.18320
-
A 4D Hybrid Algorithm to Scale Parallel Training to Thousands of GPUs 22 May 2023 · 1 repository · arXiv:2305.13525
-
DeepJSCC-l++: Robust and Bandwidth-Adaptive Wireless Image Transmission 22 May 2023 · 1 repository · arXiv:2305.13161
-
DiffAVA: Personalized Text-to-Audio Generation with Visual Alignment 22 May 2023 · 0 repositories · arXiv:2305.12903
-
Efficient Large-Scale Visual Representation Learning And Evaluation 22 May 2023 · 0 repositories · arXiv:2305.13399
-
Table Meets LLM: Can Large Language Models Understand Structured Table Data? A Benchmark and Empirical Study 22 May 2023 · 1 repository · arXiv:2305.13062
-
ExplainCPE: A Free-text Explanation Benchmark of Chinese Pharmacist Examination 22 May 2023 · 1 repository · arXiv:2305.12945
-
Exploring Energy-based Language Models with Different Architectures and Training Methods for Speech Recognition 22 May 2023 · 2 repositories · arXiv:2305.12676
-
Finding the Pillars of Strength for Multi-Head Attention 22 May 2023 · 2 repositories · arXiv:2305.14380
-
Flover: A Temporal Fusion Framework for Efficient Autoregressive Model Parallel Inference 22 May 2023 · 1 repository · arXiv:2305.13484
-
G3Detector: General GPT-Generated Text Detector 22 May 2023 · 0 repositories · arXiv:2305.12680
-
GATology for Linguistics: What Syntactic Dependencies It Knows 22 May 2023 · 0 repositories · arXiv:2305.13403
-
GNCformer Enhanced Self-attention for Automatic Speech Recognition 22 May 2023 · 0 repositories · arXiv:2305.12755
-
GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints 22 May 2023 · 4 repositories · arXiv:2305.13245Syntology 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
How Language Model Hallucinations Can Snowball 22 May 2023 · 1 repository · arXiv:2305.13534
-
SimCSE++: Improving Contrastive Learning for Sentence Embeddings from Two Perspectives 22 May 2023 · 0 repositories · arXiv:2305.13192
-
InheritSumm: A General, Versatile and Compact Summarizer by Distilling from GPT 22 May 2023 · 0 repositories · arXiv:2305.13083
-
Iterative Forward Tuning Boosts In-Context Learning in Language Models 22 May 2023 · 1 repository · arXiv:2305.13016
-
Language-Agnostic Bias Detection in Language Models with Bias Probing 22 May 2023 · 1 repository · arXiv:2305.13302
-
Learning Pedestrian Actions to Ensure Safe Autonomous Driving 22 May 2023 · 0 repositories · arXiv:2305.13051
-
Learning Subpocket Prototypes for Generalizable Structure-based Drug Design 22 May 2023 · 1 repository · arXiv:2305.13997Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
LLMs for Knowledge Graph Construction and Reasoning: Recent Capabilities and Future Opportunities 22 May 2023 · 1 repository · arXiv:2305.13168Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Atomic Inference for NLI with Generated Facts as Atoms 22 May 2023 · 1 repository · arXiv:2305.13214
-
MAILEX: Email Event and Argument Extraction 22 May 2023 · 1 repository · arXiv:2305.13469
-
Materialistic: Selecting Similar Materials in Images 22 May 2023 · 0 repositories · arXiv:2305.13291
-
Measuring Inductive Biases of In-Context Learning with Underspecified Demonstrations 22 May 2023 · 1 repository · arXiv:2305.13299
-
Multi-Task Instruction Tuning of LLaMa for Specific Scenarios: A Preliminary Study on Writing Assistance 22 May 2023 · 0 repositories · arXiv:2305.13225
-
Investigating the Role of Feed-Forward Networks in Transformers Using Parallel Attention and Feed-Forward Net Design 22 May 2023 · 0 repositories · arXiv:2305.13297
-
RecurrentGPT: Interactive Generation of (Arbitrarily) Long Text 22 May 2023 · 2 repositories · arXiv:2305.13304Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
RWKV: Reinventing RNNs for the Transformer Era 22 May 2023 · 14 repositories · arXiv:2305.13048Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
SCITAB: A Challenging Benchmark for Compositional Reasoning and Claim Verification on Scientific Tables 22 May 2023 · 1 repository · arXiv:2305.13186
-
SPARSEFIT: Few-shot Prompting with Sparse Fine-tuning for Jointly Generating Predictions and Natural Language Explanations 22 May 2023 · 1 repository · arXiv:2305.13235
-
Spatiotemporal Attention-based Semantic Compression for Real-time Video Recognition 22 May 2023 · 0 repositories · arXiv:2305.12796
-
Stock and market index prediction using Informer network 22 May 2023 · 0 repositories · arXiv:2305.14382
-
Syntactic Knowledge via Graph Attention with BERT in Machine Translation 22 May 2023 · 0 repositories · arXiv:2305.13413
-
The Emergence of Economic Rationality of GPT 22 May 2023 · 0 repositories · arXiv:2305.12763
-
Tokenized Graph Transformer with Neighborhood Augmentation for Node Classification in Large Graphs 22 May 2023 · 0 repositories · arXiv:2305.12677
-
Investigating Agency of LLMs in Human-AI Collaboration Tasks 22 May 2023 · 0 repositories · arXiv:2305.12815
-
TSPTQ-ViT: Two-scaled post-training quantization for vision transformer 22 May 2023 · 0 repositories · arXiv:2305.12901
-
VDT: General-purpose Video Diffusion Transformers via Mask Modeling 22 May 2023 · 1 repository · arXiv:2305.13311Syntology official (archive's flag): 4 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
VideoLLM: Modeling Video Sequence with Large Language Models 22 May 2023 · 1 repository · arXiv:2305.13292
-
Why current rain denoising models fail on CycleGAN created rain images in autonomous driving 22 May 2023 · 0 repositories · arXiv:2305.12983
-
A Deeper (Autoregressive) Approach to Non-Convergent Discourse Parsing 21 May 2023 · 0 repositories · arXiv:2305.12510
-
A Symbolic Framework for Evaluating Mathematical Reasoning and Generalisation with Transformers 21 May 2023 · 0 repositories · arXiv:2305.12563
-
BertRLFuzzer: A BERT and Reinforcement Learning Based Fuzzer 21 May 2023 · 1 repository · arXiv:2305.12534
-
Bi-ViT: Pushing the Limit of Vision Transformer Quantization 21 May 2023 · 0 repositories · arXiv:2305.12354
-
BiasAsker: Measuring the Bias in Conversational AI System 21 May 2023 · 1 repository · arXiv:2305.12434
-
Abstract Meaning Representation-Based Logic-Driven Data Augmentation for Logical Reasoning 21 May 2023 · 1 repository · arXiv:2305.12599
-
Evaluating the Performance of Large Language Models on GAOKAO Benchmark 21 May 2023 · 1 repository · arXiv:2305.12474Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Explaining How Transformers Use Context to Build Predictions 21 May 2023 · 1 repository · arXiv:2305.12535Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
F-PABEE: Flexible-patience-based Early Exiting for Single-label and Multi-label text Classification Tasks 21 May 2023 · 0 repositories · arXiv:2305.11916
-
Gene Set Summarization using Large Language Models 21 May 2023 · 1 repository · arXiv:2305.13338Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
GPT-3.5, GPT-4, or BARD? Evaluating LLMs Reasoning Ability in Zero-Shot Setting and Performance Boosting Through Prompts 21 May 2023 · 0 repositories · arXiv:2305.12477
-
DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection 21 May 2023 · 0 repositories · arXiv:2305.12519
-
HIINT: Historical, Intra- and Inter- personal Dynamics Modeling with Cross-person Memory Transformer 21 May 2023 · 0 repositories · arXiv:2305.12369
-
Infor-Coef: Information Bottleneck-based Dynamic Token Downsampling for Compact and Efficient language model 21 May 2023 · 0 repositories · arXiv:2305.12458
-
IR Models and the COVID-19 Pandemic: A Comparative Study of Performance and Challenges 21 May 2023 · 0 repositories · arXiv:2305.12528
-
Learning Joint 2D & 3D Diffusion Models for Complete Molecule Generation 21 May 2023 · 2 repositories · arXiv:2305.12347Syntology official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Model-Generated Pretraining Signals Improves Zero-Shot Generalization of Text-to-Text Transformers 21 May 2023 · 1 repository · arXiv:2305.12567
-
Multi-Head State Space Model for Speech Recognition 21 May 2023 · 0 repositories · arXiv:2305.12498