Methods › General › Output Functions › Softmax › Papers, page 38
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 38 of 375: papers 3,701 to 3,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
X2I: Seamless Integration of Multimodal Understanding into Diffusion Transformer via Attention Distillation 8 Mar 2025 · 1 repository · arXiv:2503.06134Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding 8 Mar 2025 · 0 repositories · arXiv:2503.06287
-
A Hybrid Model/Data-Driven Solution to Channel, Position and Orientation Tracking in mmWave Vehicular Systems 7 Mar 2025 · 0 repositories · arXiv:2503.05091
-
A Real-time Multimodal Transformer Neural Network-powered Wildfire Forecasting System 7 Mar 2025 · 0 repositories · arXiv:2503.05971
-
A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models 7 Mar 2025 · 0 repositories · arXiv:2503.05613
-
BARK: A Fully Bayesian Tree Kernel for Black-box Optimization 7 Mar 2025 · 0 repositories · arXiv:2503.05574
-
CASP: Compression of Large Multimodal Models Based on Attention Sparsity 7 Mar 2025 · 1 repository · arXiv:2503.05936
-
ColFigPhotoAttnNet: Reliable Finger Photo Presentation Attack Detection Leveraging Window-Attention on Color Spaces 7 Mar 2025 · 1 repository · arXiv:2503.05247
-
CoMoGaussian: Continuous Motion-Aware Gaussian Splatting from Motion-Blurred Images 7 Mar 2025 · 1 repository · arXiv:2503.05332
-
Deep Frequency Attention Networks for Single Snapshot Sparse Array Interpolation 7 Mar 2025 · 0 repositories · arXiv:2503.05486
-
Energy-Free Sensing and Context Recognition Using Photovoltaic Cells 7 Mar 2025 · 1 repository · arXiv:2503.05406
-
Evaluating Large Language Models in Code Generation: INFINITE Methodology for Defining the Inference Index 7 Mar 2025 · 0 repositories · arXiv:2503.05852
-
Simulating and Analysing Human Survey Responses with Large Language Models: A Case Study in Energy Stated Preference 7 Mar 2025 · 0 repositories · arXiv:2503.10652
-
Evaluating open-source Large Language Models for automated fact-checking 7 Mar 2025 · 0 repositories · arXiv:2503.05565
-
Explaining the Unexplainable: A Systematic Review of Explainable AI in Finance 7 Mar 2025 · 0 repositories · arXiv:2503.05966
-
Exploring FMCW Radars and Feature Maps for Activity Recognition: A Benchmark Study 7 Mar 2025 · 0 repositories · arXiv:2503.05629
-
FastMap: Fast Queries Initialization Based Vectorized HD Map Reconstruction Framework 7 Mar 2025 · 1 repository · arXiv:2503.05492
-
FMCHS: Advancing Traditional Chinese Medicine Herb Recommendation with Fusion of Multiscale Correlations of Herbs and Symptoms 7 Mar 2025 · 0 repositories · arXiv:2503.05167
-
FMT:A Multimodal Pneumonia Detection Model Based on Stacking MOE Framework 7 Mar 2025 · 0 repositories · arXiv:2503.05626
-
GaussianCAD: Robust Self-Supervised CAD Reconstruction from Three Orthographic Views Using 3D Gaussian Splatting 7 Mar 2025 · 0 repositories · arXiv:2503.05161
-
Language modelling techniques for analysing the impact of human genetic variation 7 Mar 2025 · 0 repositories · arXiv:2503.10655
-
Leveraging Approximate Caching for Faster Retrieval-Augmented Generation 7 Mar 2025 · 0 repositories · arXiv:2503.05530
-
Leveraging Semantic Type Dependencies for Clinical Named Entity Recognition 7 Mar 2025 · 0 repositories · arXiv:2503.05373
-
Lightweight Hypercomplex MRI Reconstruction: A Generalized Kronecker-Parameterized Approach 7 Mar 2025 · 0 repositories · arXiv:2503.05063
-
Look Before You Leap: Using Serialized State Machine for Language Conditioned Robotic Manipulation 7 Mar 2025 · 0 repositories · arXiv:2503.05114
-
MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice 7 Mar 2025 · 0 repositories · arXiv:2503.05978
-
MastermindEval: A Simple But Scalable Reasoning Benchmark 7 Mar 2025 · 1 repository · arXiv:2503.05891Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MedCAM-OsteoCls: Medical Context Aware Multimodal Classification of Knee Osteoarthritis 7 Mar 2025 · 1 repository
-
Mol-CADiff: Causality-Aware Autoregressive Diffusion for Molecule Generation 7 Mar 2025 · 0 repositories · arXiv:2503.05499
-
MPTSNet: Integrating Multiscale Periodic Local Patterns and Global Dependencies for Multivariate Time Series Classification 7 Mar 2025 · 1 repository · arXiv:2503.05582
-
Personalized Federated Learning via Learning Dynamic Graphs 7 Mar 2025 · 0 repositories · arXiv:2503.05474
-
Pi-GPS: Enhancing Geometry Problem Solving by Unleashing the Power of Diagrammatic Information 7 Mar 2025 · 0 repositories · arXiv:2503.05543Syntology 7 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 3 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Quantifying the Robustness of Retrieval-Augmented Language Models Against Spurious Features in Grounding Data 7 Mar 2025 · 0 repositories · arXiv:2503.05587
-
R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning 7 Mar 2025 · 5 repositories · arXiv:2503.05592Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
S2S-Arena, Evaluating Speech2Speech Protocols on Instruction Following with Paralinguistic Information 7 Mar 2025 · 0 repositories · arXiv:2503.05085
-
Slim attention: cut your context memory in half without loss of accuracy -- K-cache is all you need for MHA 7 Mar 2025 · 1 repository · arXiv:2503.05840
-
SplatPose: Geometry-Aware 6-DoF Pose Estimation from Single RGB Image via 3D Gaussian Splatting 7 Mar 2025 · 0 repositories · arXiv:2503.05174
-
Task-oriented Uncertainty Collaborative Learning for Label-Efficient Brain Tumor Segmentation 7 Mar 2025 · 1 repository · arXiv:2503.05682
-
Tractable Representations for Convergent Approximation of Distributional HJB Equations 7 Mar 2025 · 0 repositories · arXiv:2503.05563
-
Zero-shot Medical Event Prediction Using a Generative Pre-trained Transformer on Electronic Health Records 7 Mar 2025 · 0 repositories · arXiv:2503.05893
-
A Generalist Cross-Domain Molecular Learning Framework for Structure-Based Drug Discovery 6 Mar 2025 · 0 repositories · arXiv:2503.04362
-
A Unified Framework with Novel Metrics for Evaluating the Effectiveness of XAI Techniques in LLMs 6 Mar 2025 · 0 repositories · arXiv:2503.05050
-
Beyond RAG: Task-Aware KV Cache Compression for Comprehensive Knowledge Reasoning 6 Mar 2025 · 0 repositories · arXiv:2503.04973
-
BicliqueEncoder: An Efficient Method for Link Prediction in Bipartite Networks using Formal Concept Analysis and Transformer Encoder 6 Mar 2025 · 0 repositories · arXiv:2503.07645
-
Can We Optimize Deep RL Policy Weights as Trajectory Modeling? 6 Mar 2025 · 0 repositories · arXiv:2503.04074
-
Chart-HQA: A Benchmark for Hypothetical Question Answering in Charts 6 Mar 2025 · 0 repositories · arXiv:2503.04095
-
Collapse of Dense Retrievers: Short, Early, and Literal Biases Outranking Factual Evidence 6 Mar 2025 · 0 repositories · arXiv:2503.05037
-
Compositional Causal Reasoning Evaluation in Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04556
-
Conformal forecasting for surgical instrument trajectory 6 Mar 2025 · 0 repositories · arXiv:2503.04191
-
DB-Explore: Automated Database Exploration and Instruction Synthesis for Text-to-SQL 6 Mar 2025 · 0 repositories · arXiv:2503.04959
-
Early Detection of Mental Health Issues Using Social Media Posts 6 Mar 2025 · 0 repositories · arXiv:2503.07653
-
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection 6 Mar 2025 · 1 repository · arXiv:2503.04156
-
Gate-Shift-Pose: Enhancing Action Recognition in Sports with Skeleton Information 6 Mar 2025 · 1 repository · arXiv:2503.04470
-
GBT-SAM: Adapting a Foundational Deep Learning Model for Generalizable Brain Tumor Segmentation via Efficient Integration of Multi-Parametric MRI Data 6 Mar 2025 · 1 repository · arXiv:2503.04325
-
Hedging with Sparse Reward Reinforcement Learning 6 Mar 2025 · 0 repositories · arXiv:2503.04218
-
High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects 6 Mar 2025 · 0 repositories · arXiv:2503.04862
-
HILGEN: Hierarchically-Informed Data Generation for Biomedical NER Using Knowledgebases and Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04930
-
HybridNorm: Towards Stable and Efficient Transformer Training via Hybrid Normalization 6 Mar 2025 · 1 repository · arXiv:2503.04598
-
In-depth Analysis of Graph-based RAG in a Unified Framework 6 Mar 2025 · 0 repositories · arXiv:2503.04338
-
Incentivizing Multi-Tenant Split Federated Learning for Foundation Models at the Network Edge 6 Mar 2025 · 0 repositories · arXiv:2503.04971
-
Interpretable Transformation and Analysis of Timelines through Learning via Surprisability 6 Mar 2025 · 0 repositories · arXiv:2503.04502
-
Joint Masked Reconstruction and Contrastive Learning for Mining Interactions Between Proteins 6 Mar 2025 · 1 repository · arXiv:2503.04650
-
Layer-Specific Scaling of Positional Encodings for Superior Long-Context Modeling 6 Mar 2025 · 0 repositories · arXiv:2503.04355
-
Learning Transformer-based World Models with Contrastive Predictive Coding 6 Mar 2025 · 0 repositories · arXiv:2503.04416
-
Learning Wideband User Scheduling and Hybrid Precoding with Graph Neural Networks 6 Mar 2025 · 0 repositories · arXiv:2503.04233
-
LEDiT: Your Length-Extrapolatable Diffusion Transformer without Positional Encoding 6 Mar 2025 · 0 repositories · arXiv:2503.04344
-
Leveraging Large Language Models to Address Data Scarcity in Machine Learning: Applications in Graphene Synthesis 6 Mar 2025 · 1 repository · arXiv:2503.04870
-
Multi-modal Summarization in Model-Based Engineering: Automotive Software Development Case Study 6 Mar 2025 · 0 repositories · arXiv:2503.04506
-
Scale-Invariant Adversarial Attack against Arbitrary-scale Super-resolution 6 Mar 2025 · 0 repositories · arXiv:2503.04385
-
Toward Lightweight and Fast Decoders for Diffusion Models in Image and Video Generation 6 Mar 2025 · 1 repository · arXiv:2503.04871
-
Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation with Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04280
-
A Multimodal Framework for Topic Propagation Classification in Social Networks 5 Mar 2025 · 0 repositories · arXiv:2503.03112
-
Addressing Overprescribing Challenges: Fine-Tuning Large Language Models for Medication Recommendation Tasks 5 Mar 2025 · 1 repository · arXiv:2503.03687Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation 5 Mar 2025 · 0 repositories · arXiv:2503.03556
-
AHCPTQ: Accurate and Hardware-Compatible Post-Training Quantization for Segment Anything Model 5 Mar 2025 · 0 repositories · arXiv:2503.03088
-
All-atom Diffusion Transformers: Unified generative modelling of molecules and materials 5 Mar 2025 · 1 repository · arXiv:2503.03965Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
An Aspect Extraction Framework using Different Embedding Types, Learning Models, and Dependency Structure 5 Mar 2025 · 1 repository · arXiv:2503.03512
-
Analogical Reasoning Inside Large Language Models: Concept Vectors and the Limits of Abstraction 5 Mar 2025 · 1 repository · arXiv:2503.03666
-
BANet: Bilateral Aggregation Network for Mobile Stereo Matching 5 Mar 2025 · 1 repository · arXiv:2503.03259Syntology official (archive's flag): 8 ran · 8 ran (of which 4 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples)
-
Can Frontier LLMs Replace Annotators in Biomedical Text Mining? Analyzing Challenges and Exploring Solutions 5 Mar 2025 · 1 repository · arXiv:2503.03261
-
Conformal Transformations for Symmetric Power Transformers 5 Mar 2025 · 0 repositories · arXiv:2503.03269
-
Convergence Rates for Softmax Gating Mixture of Experts 5 Mar 2025 · 0 repositories · arXiv:2503.03213
-
DA-STGCN: 4D Trajectory Prediction Based on Spatiotemporal Feature Extraction 5 Mar 2025 · 0 repositories · arXiv:2503.04823
-
Deictic Codes, Demonstratives, and Reference: A Step Toward Solving the Grounding Problem 5 Mar 2025 · 0 repositories · arXiv:2503.03495
-
DTU-Net: A Multi-Scale Dilated Transformer Network for Nonlinear Hyperspectral Unmixing 5 Mar 2025 · 0 repositories · arXiv:2503.03465
-
DualDiff+: Dual-Branch Diffusion for High-Fidelity Video Generation with Reward Guidance 5 Mar 2025 · 1 repository · arXiv:2503.03689
-
Golden Cudgel Network for Real-Time Semantic Segmentation 5 Mar 2025 · 1 repository · arXiv:2503.03325
-
Intermediate-Task Transfer Learning: Leveraging Sarcasm Detection for Stance Detection 5 Mar 2025 · 0 repositories · arXiv:2503.03172
-
Introduction to Artificial Consciousness: History, Current Trends and Ethical Challenges 5 Mar 2025 · 0 repositories · arXiv:2503.05823
-
Knowledge Augmentation in Federation: Rethinking What Collaborative Learning Can Bring Back to Decentralized Data 5 Mar 2025 · 0 repositories · arXiv:2503.03140
-
Learning to Reduce Search Space for Generalizable Neural Routing Solver 5 Mar 2025 · 0 repositories · arXiv:2503.03137
-
Large language models in finance : what is financial sentiment? 5 Mar 2025 · 0 repositories · arXiv:2503.03612
-
MA-LoT: Multi-Agent Lean-based Long Chain-of-Thought Reasoning enhances Formal Theorem Proving 5 Mar 2025 · 1 repository · arXiv:2503.03205
-
Multi-View Depth Consistent Image Generation Using Generative AI Models: Application on Architectural Design of University Buildings 5 Mar 2025 · 0 repositories · arXiv:2503.03068
-
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs 5 Mar 2025 · 0 repositories · arXiv:2503.03304
-
Partial Convolution Meets Visual Attention 5 Mar 2025 · 0 repositories · arXiv:2503.03148
-
PathRWKV: Enabling Whole Slide Prediction with Recurrent-Transformer 5 Mar 2025 · 0 repositories · arXiv:2503.03199
-
Personalized Federated Fine-tuning for Heterogeneous Data: An Automatic Rank Learning Approach via Two-Level LoRA 5 Mar 2025 · 0 repositories · arXiv:2503.03920
-
Petri Timo 5 Mar 2025 · 0 repositories
-
PowerAttention: Exponentially Scaling of Receptive Fields for Effective Sparse Attention 5 Mar 2025 · 0 repositories · arXiv:2503.03588