Methods › General › Output Functions › Softmax › Papers, page 11
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 11 of 375: papers 1,001 to 1,100 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Only Large Weights (And Not Skip Connections) Can Prevent the Perils of Rank Collapse 22 May 2025 · 0 repositories · arXiv:2505.16284
-
OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning 22 May 2025 · 1 repository · arXiv:2505.16974
-
Partitioning and Observability in Linear Systems via Submodular Optimization 22 May 2025 · 0 repositories · arXiv:2505.16169
-
PaTH Attention: Position Encoding via Accumulating Householder Transformations 22 May 2025 · 2 repositories · arXiv:2505.16381Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Personalizing Student-Agent Interactions Using Log-Contextualized Retrieval Augmented Generation (RAG) 22 May 2025 · 0 repositories · arXiv:2505.17238
-
Pursuing Temporal-Consistent Video Virtual Try-On via Dynamic Pose Interaction 22 May 2025 · 0 repositories · arXiv:2505.16980
-
R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning 22 May 2025 · 3 repositories · arXiv:2505.17005Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
RAP: Runtime-Adaptive Pruning for LLM Inference 22 May 2025 · 0 repositories · arXiv:2505.17138
-
Realistic Evaluation of TabPFN v2 in Open Environments 22 May 2025 · 0 repositories · arXiv:2505.16226
-
Reasoning in Neurosymbolic AI 22 May 2025 · 0 repositories · arXiv:2505.20313
-
REPA Works Until It Doesn't: Early-Stopped, Holistic Alignment Supercharges Diffusion Training 22 May 2025 · 1 repository · arXiv:2505.16792Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Representation Discrepancy Bridging Method for Remote Sensing Image-Text Retrieval 22 May 2025 · 0 repositories · arXiv:2505.16756
-
Reward-Aware Proto-Representations in Reinforcement Learning 22 May 2025 · 0 repositories · arXiv:2505.16217
-
SafeKey: Amplifying Aha-Moment Insights for Safety Reasoning 22 May 2025 · 0 repositories · arXiv:2505.16186
-
SAMba-UNet: Synergizing SAM2 and Mamba in UNet with Heterogeneous Aggregation for Cardiac MRI Segmentation 22 May 2025 · 0 repositories · arXiv:2505.16304
-
Scalable Graph Generative Modeling via Substructure Sequences 22 May 2025 · 1 repository · arXiv:2505.16130Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Search Wisely: Mitigating Sub-optimal Agentic Searches By Reducing Uncertainty 22 May 2025 · 0 repositories · arXiv:2505.17281
-
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding 22 May 2025 · 0 repositories · arXiv:2505.16652
-
Self-Classification Enhancement and Correction for Weakly Supervised Object Detection 22 May 2025 · 0 repositories · arXiv:2505.16294
-
SELF: Self-Extend the Context Length With Logistic Growth Function 22 May 2025 · 1 repository · arXiv:2505.17296
-
SHaDe: Compact and Consistent Dynamic 3D Reconstruction via Tri-Plane Deformation and Latent Diffusion 22 May 2025 · 0 repositories · arXiv:2505.16535
-
Shadows in the Attention: Contextual Perturbation and Representation Drift in the Dynamics of Hallucination in LLMs 22 May 2025 · 0 repositories · arXiv:2505.16894
-
Sketchy Bounding-box Supervision for 3D Instance Segmentation 22 May 2025 · 1 repository · arXiv:2505.16399
-
SpecMaskFoley: Steering Pretrained Spectral Masked Generative Transformer Toward Synchronized Video-to-audio Synthesis via ControlNet 22 May 2025 · 0 repositories · arXiv:2505.16195
-
Style Transfer with Diffusion Models for Synthetic-to-Real Domain Adaptation 22 May 2025 · 1 repository · arXiv:2505.16360
-
Swin Transformer for Robust CGI Images Detection: Intra- and Inter-Dataset Analysis across Multiple Color Spaces 22 May 2025 · 0 repositories · arXiv:2505.16253
-
Temporal and Spatial Feature Fusion Framework for Dynamic Micro Expression Recognition 22 May 2025 · 0 repositories · arXiv:2505.16372
-
Temporal Object Captioning for Street Scene Videos from LiDAR Tracks 22 May 2025 · 0 repositories · arXiv:2505.16594
-
The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm 22 May 2025 · 0 repositories · arXiv:2505.16932
-
Three Minds, One Legend: Jailbreak Large Reasoning Model with Adaptive Stacked Ciphers 22 May 2025 · 0 repositories · arXiv:2505.16241
-
Training-Free Efficient Video Generation via Dynamic Token Carving 22 May 2025 · 1 repository · arXiv:2505.16864
-
Training-Free Reasoning and Reflection in MLLMs 22 May 2025 · 0 repositories · arXiv:2505.16151
-
Transformer brain encoders explain human high-level visual responses 22 May 2025 · 1 repository · arXiv:2505.17329
-
Transformer Copilot: Learning from The Mistake Log in LLM Fine-tuning 22 May 2025 · 1 repository · arXiv:2505.16270Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Tropical Attention: Neural Algorithmic Reasoning for Combinatorial Algorithms 22 May 2025 · 0 repositories · arXiv:2505.17190Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
Understanding Differential Transformer Unchains Pretrained Self-Attentions 22 May 2025 · 0 repositories · arXiv:2505.16333
-
VoxRAG: A Step Toward Transcription-Free RAG Systems in Spoken Question Answering 22 May 2025 · 0 repositories · arXiv:2505.17326
-
Walk&Retrieve: Simple Yet Effective Zero-shot Retrieval-Augmented Generation via Knowledge Graph Walks 22 May 2025 · 1 repository · arXiv:2505.16849
-
When can isotropy help adapt LLMs' next word prediction to numerical domains? 22 May 2025 · 0 repositories · arXiv:2505.17135
-
When Do LLMs Admit Their Mistakes? Understanding the Role of Model Belief in Retraction 22 May 2025 · 1 repository · arXiv:2505.16170
-
Why Can Accurate Models Be Learned from Inaccurate Annotations? 22 May 2025 · 0 repositories · arXiv:2505.16159
-
Zebra-Llama: Towards Extremely Efficient Hybrid Models 22 May 2025 · 0 repositories · arXiv:2505.17272
-
Zero-Shot Hyperspectral Pansharpening Using Hysteresis-Based Tuning for Spectral Quality Control 22 May 2025 · 1 repository · arXiv:2505.16658
-
A Taxonomy of Structure from Motion Methods 21 May 2025 · 0 repositories · arXiv:2505.15814
-
AdUE: Improving uncertainty estimation head for LoRA adapters in LLMs 21 May 2025 · 0 repositories · arXiv:2505.15443
-
An Approach Towards Identifying Bangladeshi Leaf Diseases through Transfer Learning and XAI 21 May 2025 · 0 repositories · arXiv:2505.16033
-
An Efficient Private GPT Never Autoregressively Decodes 21 May 2025 · 0 repositories · arXiv:2505.15252
-
An Exploratory Approach Towards Investigating and Explaining Vision Transformer and Transfer Learning for Brain Disease Detection 21 May 2025 · 0 repositories · arXiv:2505.16039
-
Beyond Node Attention: Multi-Scale Harmonic Encoding for Feature-Wise Graph Message Passing 21 May 2025 · 0 repositories · arXiv:2505.15015
-
BountyBench: Dollar Impact of AI Agent Attackers and Defenders on Real-World Cybersecurity Systems 21 May 2025 · 0 repositories · arXiv:2505.15216
-
BR-TaxQA-R: A Dataset for Question Answering with References for Brazilian Personal Income Tax Law, including case law 21 May 2025 · 0 repositories · arXiv:2505.15916
-
CEBSNet: Change-Excited and Background-Suppressed Network with Temporal Dependency Modeling for Bitemporal Change Detection 21 May 2025 · 0 repositories · arXiv:2505.15322
-
Collaborative Problem-Solving in an Optimization Game 21 May 2025 · 1 repository · arXiv:2505.15490
-
Comprehensive Lung Disease Detection Using Deep Learning Models and Hybrid Chest X-ray Data with Explainable AI 21 May 2025 · 0 repositories · arXiv:2505.16028
-
Convolutional Long Short-Term Memory Neural Networks Based Numerical Simulation of Flow Field 21 May 2025 · 0 repositories · arXiv:2505.15533
-
Decouple and Orthogonalize: A Data-Free Framework for LoRA Merging 21 May 2025 · 0 repositories · arXiv:2505.15875
-
Diffusion vs. Autoregressive Language Models: A Text Embedding Perspective 21 May 2025 · 0 repositories · arXiv:2505.15045
-
DISCO Balances the Scales: Adaptive Domain- and Difficulty-Aware Reinforcement Learning on Imbalanced Data 21 May 2025 · 0 repositories · arXiv:2505.15074
-
dKV-Cache: The Cache for Diffusion Language Models 21 May 2025 · 2 repositories · arXiv:2505.15781
-
Domain Adaptive Skin Lesion Classification via Conformal Ensemble of Vision Transformers 21 May 2025 · 0 repositories · arXiv:2505.15997
-
Exploring the Innovation Opportunities for Pre-trained Models 21 May 2025 · 0 repositories · arXiv:2505.15790
-
Filtering Learning Histories Enhances In-Context Reinforcement Learning 21 May 2025 · 0 repositories · arXiv:2505.15143
-
Fourier-Invertible Neural Encoder (FINE) for Homogeneous Flows 21 May 2025 · 0 repositories · arXiv:2505.15329
-
Gated Integration of Low-Rank Adaptation for Continual Learning of Language Models 21 May 2025 · 1 repository · arXiv:2505.15424Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Guidelines for the Quality Assessment of Energy-Aware NAS Benchmarks 21 May 2025 · 0 repositories · arXiv:2505.15631
-
Hallucinate at the Last in Long Response Generation: A Case Study on Long Document Summarization 21 May 2025 · 0 repositories · arXiv:2505.15291
-
HDLxGraph: Bridging Large Language Models and HDL Repositories via HDL Graph Databases 21 May 2025 · 1 repository · arXiv:2505.15701
-
Higher-order Structure Boosts Link Prediction on Temporal Graphs 21 May 2025 · 0 repositories · arXiv:2505.15746
-
Hunyuan-TurboS: Advancing Large Language Models through Mamba-Transformer Synergy and Adaptive Chain-of-Thought 21 May 2025 · 0 repositories · arXiv:2505.15431
-
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation 21 May 2025 · 0 repositories · arXiv:2505.15872
-
Internal and External Impacts of Natural Language Processing Papers 21 May 2025 · 0 repositories · arXiv:2505.16061
-
Interspatial Attention for Efficient 4D Human Video Generation 21 May 2025 · 0 repositories · arXiv:2505.15800
-
Leveraging Foundation Models for Multimodal Graph-Based Action Recognition 21 May 2025 · 0 repositories · arXiv:2505.15192
-
Leveraging Large Language Models for Command Injection Vulnerability Analysis in Python: An Empirical Study on Popular Open-Source Projects 21 May 2025 · 0 repositories · arXiv:2505.15088
-
Leveraging the Powerful Attention of a Pre-trained Diffusion Model for Exemplar-based Image Colorization 21 May 2025 · 1 repository · arXiv:2505.15812
-
LFTF: Locating First and Then Fine-Tuning for Mitigating Gender Bias in Large Language Models 21 May 2025 · 0 repositories · arXiv:2505.15475
-
LogiCase: Effective Test Case Generation from Logical Description in Competitive Programming 21 May 2025 · 0 repositories · arXiv:2505.15039Syntology 11 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 8 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Lossless Token Merging Even Without Fine-Tuning in Vision Transformers 21 May 2025 · 0 repositories · arXiv:2505.15160
-
MaxPoolBERT: Enhancing BERT Classification via Layer- and Token-Wise Aggregation 21 May 2025 · 0 repositories · arXiv:2505.15696
-
Mechanistic Insights into Grokking from the Embedding Layer 21 May 2025 · 0 repositories · arXiv:2505.15624
-
Mitigating Spurious Correlations with Causal Logit Perturbation 21 May 2025 · 0 repositories · arXiv:2505.15246
-
MonoSplat: Generalizable 3D Gaussian Splatting from Monocular Depth Foundation Models 21 May 2025 · 1 repository · arXiv:2505.15185Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples)
-
Moonbeam: A MIDI Foundation Model Using Both Absolute and Relative Music Attributes 21 May 2025 · 1 repository · arXiv:2505.15559Syntology 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
NL-Debugging: Exploiting Natural Language as an Intermediate Representation for Code Debugging 21 May 2025 · 0 repositories · arXiv:2505.15356
-
Oversmoothing, "Oversquashing", Heterophily, Long-Range, and more: Demystifying Common Beliefs in Graph Machine Learning 21 May 2025 · 0 repositories · arXiv:2505.15547
-
Ranking Free RAG: Replacing Re-ranking with Selection in RAG for Sensitive Domains 21 May 2025 · 0 repositories · arXiv:2505.16014
-
RePPL: Recalibrating Perplexity by Uncertainty in Semantic Propagation and Language Generation for Explainable QA Hallucination Detection 21 May 2025 · 0 repositories · arXiv:2505.15386
-
Reranking with Compressed Document Representation 21 May 2025 · 0 repositories · arXiv:2505.15394
-
RLBenchNet: The Right Network for the Right Reinforcement Learning Task 21 May 2025 · 1 repository · arXiv:2505.15040
-
Robo-DM: Data Management For Large Robot Datasets 21 May 2025 · 0 repositories · arXiv:2505.15558
-
RoT: Enhancing Table Reasoning with Iterative Row-Wise Traversals 21 May 2025 · 0 repositories · arXiv:2505.15110
-
SAMA-UNet: Enhancing Medical Image Segmentation with Self-Adaptive Mamba-Like Attention and Causal-Resonance Learning 21 May 2025 · 1 repository · arXiv:2505.15234
-
Scaling Diffusion Transformers Efficiently via μP 21 May 2025 · 1 repository · arXiv:2505.15270
-
Seeing the Trees for the Forest: Rethinking Weakly-Supervised Medical Visual Grounding 21 May 2025 · 0 repositories · arXiv:2505.15123
-
Set-LLM: A Permutation-Invariant LLM 21 May 2025 · 0 repositories · arXiv:2505.15433
-
Short-Range Dependency Effects on Transformer Instability and a Decomposed Attention Solution 21 May 2025 · 0 repositories · arXiv:2505.15548
-
Silent Leaks: Implicit Knowledge Extraction Attack on RAG Systems through Benign Queries 21 May 2025 · 1 repository · arXiv:2505.15420
-
Single LLM, Multiple Roles: A Unified Retrieval-Augmented Generation Framework Using Role-Specific Token Optimization 21 May 2025 · 0 repositories · arXiv:2505.15444
-
Small Language Models in the Real World: Insights from Industrial Text Classification 21 May 2025 · 0 repositories · arXiv:2505.16078
-
SUS backprop: linear backpropagation algorithm for long inputs in transformers 21 May 2025 · 0 repositories · arXiv:2505.15080