Methods › General › Output Functions › Softmax › Papers, page 4
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 4 of 375: papers 301 to 400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MedMoE: Modality-Specialized Mixture of Experts for Medical Vision-Language Understanding 10 Jun 2025 · 0 repositories · arXiv:2506.08356
-
Mitigating Posterior Salience Attenuation in Long-Context LLMs with Positional Contrastive Decoding 10 Jun 2025 · 0 repositories · arXiv:2506.08371
-
MLVTG: Mamba-Based Feature Alignment and LLM-Driven Purification for Multi-Modal Video Temporal Grounding 10 Jun 2025 · 0 repositories · arXiv:2506.08512
-
NAM: A Normalization Attention Model for Personalized Product Search In Fliggy 10 Jun 2025 · 0 repositories · arXiv:2506.08382
-
Olica: Efficient Structured Pruning of Large Language Models without Retraining 10 Jun 2025 · 1 repository · arXiv:2506.08436Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
PatchGuard: Adversarially Robust Anomaly Detection and Localization through Vision Transformers and Pseudo Anomalies 10 Jun 2025 · 2 repositories · arXiv:2506.09237
-
PlantBert: An Open Source Language Model for Plant Science 10 Jun 2025 · 0 repositories · arXiv:2506.08897
-
Plug-and-Play Linear Attention for Pre-trained Image and Video Restoration Models 10 Jun 2025 · 1 repository · arXiv:2506.08520
-
Robust Visual Localization via Semantic-Guided Multi-Scale Transformer 10 Jun 2025 · 0 repositories · arXiv:2506.08526
-
ScalableHD: Scalable and High-Throughput Hyperdimensional Computing Inference on Multi-Core CPUs 10 Jun 2025 · 0 repositories · arXiv:2506.09282
-
SDMPrune: Self-Distillation MLP Pruning for Efficient Large Language Models 10 Jun 2025 · 1 repository · arXiv:2506.11120
-
SeerAttention-R: Sparse Attention Adaptation for Long Reasoning 10 Jun 2025 · 1 repository · arXiv:2506.08889
-
Self-Anchored Attention Model for Sample-Efficient Classification of Prosocial Text Chat 10 Jun 2025 · 0 repositories · arXiv:2506.09259
-
SEMA: a Scalable and Efficient Mamba like Attention via Token Localization and Averaging 10 Jun 2025 · 0 repositories · arXiv:2506.08297Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 2 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
SurfR: Surface Reconstruction with Multi-scale Attention 10 Jun 2025 · 0 repositories · arXiv:2506.08635
-
TACTIC: Translation Agents with Cognitive-Theoretic Interactive Collaboration 10 Jun 2025 · 1 repository · arXiv:2506.08403
-
The impact of fine tuning in LLaMA on hallucinations for named entity extraction in legal documentation 10 Jun 2025 · 0 repositories · arXiv:2506.08827
-
The Predictive Brain: Neural Correlates of Word Expectancy Align with Large Language Model Prediction Probabilities 10 Jun 2025 · 0 repositories · arXiv:2506.08511
-
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers 10 Jun 2025 · 0 repositories · arXiv:2506.08641
-
Transformers Meet Hyperspectral Imaging: A Comprehensive Study of Models, Challenges and Open Problems 10 Jun 2025 · 0 repositories · arXiv:2506.08596
-
XGraphRAG: Interactive Visual Analysis for Graph-based Retrieval-Augmented Generation 10 Jun 2025 · 1 repository · arXiv:2506.13782
-
4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular Videos 9 Jun 2025 · 0 repositories · arXiv:2506.08015
-
A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation 9 Jun 2025 · 0 repositories · arXiv:2506.08210
-
A Unified Anti-Jamming Design in Complex Environments Based on Cross-Modal Fusion and Intelligent Decision-Making 9 Jun 2025 · 0 repositories · arXiv:2506.07532
-
Adapter Naturally Serves as Decoupler for Cross-Domain Few-Shot Semantic Segmentation 9 Jun 2025 · 0 repositories · arXiv:2506.07376
-
Adaptive Blind Super-Resolution Network for Spatial-Specific and Spatial-Agnostic Degradations 9 Jun 2025 · 0 repositories · arXiv:2506.07705
-
Benchmarking Foundation Speech and Language Models for Alzheimer's Disease and Related Dementia Detection from Spontaneous Speech 9 Jun 2025 · 0 repositories · arXiv:2506.11119
-
Bingo: Boosting Efficient Reasoning of LLMs via Dynamic and Significance-based Reinforcement Learning 9 Jun 2025 · 0 repositories · arXiv:2506.08125
-
Can AI Validate Science? Benchmarking LLMs for Accurate Scientific Claim → Evidence Reasoning 9 Jun 2025 · 1 repository · arXiv:2506.08235
-
Can Hessian-Based Insights Support Fault Diagnosis in Attention-based Models? 9 Jun 2025 · 0 repositories · arXiv:2506.07871
-
CrosswalkNet: An Optimized Deep Learning Framework for Pedestrian Crosswalk Detection in Aerial Images with High-Performance Computing 9 Jun 2025 · 0 repositories · arXiv:2506.07885
-
CyberV: Cybernetics for Test-time Scaling in Video Understanding 9 Jun 2025 · 1 repository · arXiv:2506.07971
-
DLNet: Direction-Aware Feature Integration for Robust Lane Detection in Complex Environments 9 Jun 2025 · 1 repository
-
Evidential Spectrum-Aware Contrastive Learning for OOD Detection in Dynamic Graphs 9 Jun 2025 · 1 repository · arXiv:2506.07417
-
Generative Voice Bursts during Phone Call 9 Jun 2025 · 0 repositories · arXiv:2506.07526
-
HAELT: A Hybrid Attentive Ensemble Learning Transformer Framework for High-Frequency Stock Price Forecasting 9 Jun 2025 · 0 repositories · arXiv:2506.13981
-
Hierarchical Lexical Graph for Enhanced Multi-Hop Retrieval 9 Jun 2025 · 1 repository · arXiv:2506.08074
-
Knowledge Compression via Question Generation: Enhancing Multihop Document Retrieval without Fine-tuning 9 Jun 2025 · 0 repositories · arXiv:2506.13778
-
Lightweight Sequential Transformers for Blood Glucose Level Prediction in Type-1 Diabetes 9 Jun 2025 · 0 repositories · arXiv:2506.07864
-
LlamaRec-LKG-RAG: A Single-Pass, Learnable Knowledge Graph-RAG Framework for LLM-Based Ranking 9 Jun 2025 · 1 repository · arXiv:2506.07449
-
LLM-BT-Terms: Back-Translation as a Framework for Terminology Standardization and Dynamic Semantic Embedding 9 Jun 2025 · 0 repositories · arXiv:2506.08174
-
LLM-driven Indoor Scene Layout Generation via Scaled Human-aligned Data Synthesis and Multi-Stage Preference Optimization 9 Jun 2025 · 0 repositories · arXiv:2506.07570
-
LUCIFER: Language Understanding and Context-Infused Framework for Exploration and Behavior Refinement 9 Jun 2025 · 0 repositories · arXiv:2506.07915
-
M2Restore: Mixture-of-Experts-based Mamba-CNN Fusion Framework for All-in-One Image Restoration 9 Jun 2025 · 0 repositories · arXiv:2506.07814
-
MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation 9 Jun 2025 · 0 repositories · arXiv:2506.07999
-
MiniCPM4: Ultra-Efficient LLMs on End Devices 9 Jun 2025 · 1 repository · arXiv:2506.07900Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Mondrian: Transformer Operators via Domain Decomposition 9 Jun 2025 · 0 repositories · arXiv:2506.08226
-
Multilingual Hate Speech Detection in Social Media Using Translation-Based Approaches with Large Language Models 9 Jun 2025 · 0 repositories · arXiv:2506.08147
-
Multiple Object Stitching for Unsupervised Representation Learning 9 Jun 2025 · 1 repository · arXiv:2506.07364
-
Nearness of Neighbors Attention for Regression in Supervised Finetuning 9 Jun 2025 · 1 repository · arXiv:2506.08139
-
OneIG-Bench: Omni-dimensional Nuanced Evaluation for Image Generation 9 Jun 2025 · 1 repository · arXiv:2506.07977
-
Prompt to Protection: A Comparative Study of Multimodal LLMs in Construction Hazard Recognition 9 Jun 2025 · 0 repositories · arXiv:2506.07436
-
Quantum Graph Transformer for NLP Sentiment Classification 9 Jun 2025 · 0 repositories · arXiv:2506.07937
-
Rethinking Cross-Modal Interaction in Multimodal Diffusion Transformers 9 Jun 2025 · 1 repository · arXiv:2506.07986
-
SceneRAG: Scene-level Retrieval-Augmented Generation for Video Understanding 9 Jun 2025 · 0 repositories · arXiv:2506.07600
-
SILK: Smooth InterpoLation frameworK for motion in-betweening A Simplified Computational Approach 9 Jun 2025 · 0 repositories · arXiv:2506.09075
-
SoK: Data Reconstruction Attacks Against Machine Learning Models: Definition, Metrics, and Benchmark 9 Jun 2025 · 0 repositories · arXiv:2506.07888
-
ST-GraphNet: A Spatio-Temporal Graph Neural Network for Understanding and Predicting Automated Vehicle Crash Severity 9 Jun 2025 · 0 repositories · arXiv:2506.08051
-
STAMImputer: Spatio-Temporal Attention MoE for Traffic Data Imputation 9 Jun 2025 · 1 repository · arXiv:2506.08054
-
Text-guided multi-stage cross-perception network for medical image segmentation 9 Jun 2025 · 0 repositories · arXiv:2506.07475
-
Vision Transformers Don't Need Trained Registers 9 Jun 2025 · 1 repository · arXiv:2506.08010Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
When Style Breaks Safety: Defending Language Models Against Superficial Style Alignment 9 Jun 2025 · 1 repository · arXiv:2506.07452Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A dependently-typed calculus of event telicity and culminativity 8 Jun 2025 · 0 repositories · arXiv:2506.06968
-
Accelerating 3D Gaussian Splatting with Neural Sorting and Axis-Oriented Rasterization 8 Jun 2025 · 0 repositories · arXiv:2506.07069
-
Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation 8 Jun 2025 · 0 repositories · arXiv:2506.07214
-
Circuit-Based Modeling Approach for Channel Estimation in RIS-Assisted Communications 8 Jun 2025 · 0 repositories · arXiv:2506.07124
-
Joint Channel and Symbol Estimation for Communication Systems with Movable Antennas 8 Jun 2025 · 0 repositories · arXiv:2506.07183
-
MAGNet: A Multi-Scale Attention-Guided Graph Fusion Network for DRC Violation Detection 8 Jun 2025 · 0 repositories · arXiv:2506.07126
-
MS-TVNet:A Long-Term Time Series Prediction Method Based on Multi-Scale Dynamic Convolution 8 Jun 2025 · 0 repositories · arXiv:2506.17253
-
Physics-Informed Teleconnection-Aware Transformer for Global Subseasonal-to-Seasonal Forecasting 8 Jun 2025 · 0 repositories · arXiv:2506.08049
-
Quality-Diversity Red-Teaming: Automated Generation of High-Quality and Diverse Attackers for Large Language Models 8 Jun 2025 · 0 repositories · arXiv:2506.07121
-
RBA-FE: A Robust Brain-Inspired Audio Feature Extractor for Depression Diagnosis 8 Jun 2025 · 0 repositories · arXiv:2506.07118
-
Reward Model Interpretability via Optimal and Pessimal Tokens 8 Jun 2025 · 0 repositories · arXiv:2506.07326
-
SiliCoN: Simultaneous Nuclei Segmentation and Color Normalization of Histological Images 8 Jun 2025 · 0 repositories · arXiv:2506.07028
-
Breaking Data Silos: Towards Open and Scalable Mobility Foundation Models via Generative Continual Learning 7 Jun 2025 · 0 repositories · arXiv:2506.06694
-
Can In-Context Reinforcement Learning Recover From Reward Poisoning Attacks? 7 Jun 2025 · 0 repositories · arXiv:2506.06891
-
Conditional Denoising Diffusion for ISAC Enhanced Channel Estimation in Cell-Free 6G 7 Jun 2025 · 0 repositories · arXiv:2506.06942
-
Deep Inertial Pose: A deep learning approach for human pose estimation 7 Jun 2025 · 0 repositories · arXiv:2506.06850
-
Do Protein Transformers Have Biological Intelligence? 7 Jun 2025 · 1 repository · arXiv:2506.06701
-
Exploring Length Generalization For Transformer-based Speech Enhancement 7 Jun 2025 · 0 repositories · arXiv:2506.06697
-
Graph Neural Networks in Modern AI-aided Drug Discovery 7 Jun 2025 · 0 repositories · arXiv:2506.06915
-
RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation 7 Jun 2025 · 0 repositories · arXiv:2506.06677
-
Training-Free Identity Preservation in Stylized Image Generation Using Diffusion Models 7 Jun 2025 · 0 repositories · arXiv:2506.06802
-
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning 6 Jun 2025 · 0 repositories · arXiv:2506.06072
-
Direct Behavior Optimization: Unlocking the Potential of Lightweight LLMs 6 Jun 2025 · 0 repositories · arXiv:2506.06401
-
Domain Adaptation in Agricultural Image Analysis: A Comprehensive Review from Shallow Models to Deep Learning 6 Jun 2025 · 0 repositories · arXiv:2506.05972
-
Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey 6 Jun 2025 · 0 repositories · arXiv:2506.11102
-
FPDANet: A Multi-Section Classification Model for Intelligent Screening of Fetal Ultrasound 6 Jun 2025 · 0 repositories · arXiv:2506.06054
-
Joint-GCG: Unified Gradient-Based Poisoning Attacks on Retrieval-Augmented Generation Systems 6 Jun 2025 · 1 repository · arXiv:2506.06151
-
On-board Mission Replanning for Adaptive Cooperative Multi-Robot Systems 6 Jun 2025 · 0 repositories · arXiv:2506.06094
-
RecGPT: A Foundation Model for Sequential Recommendation 6 Jun 2025 · 1 repository · arXiv:2506.06270Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TADA: Training-free Attribution and Out-of-Domain Detection of Audio Deepfakes 6 Jun 2025 · 1 repository · arXiv:2506.05802
-
Textile Analysis for Recycling Automation using Transfer Learning and Zero-Shot Foundation Models 6 Jun 2025 · 0 repositories · arXiv:2506.06569
-
The Lock-in Hypothesis: Stagnation by Algorithm 6 Jun 2025 · 0 repositories · arXiv:2506.06166
-
The Optimization Paradox in Clinical AI Multi-Agent Systems 6 Jun 2025 · 1 repository · arXiv:2506.06574
-
When Better Features Mean Greater Risks: The Performance-Privacy Trade-Off in Contrastive Learning 6 Jun 2025 · 0 repositories · arXiv:2506.05743
-
When to use Graphs in RAG: A Comprehensive Analysis for Graph Retrieval-Augmented Generation 6 Jun 2025 · 1 repository · arXiv:2506.05690Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
DACN: Dual-Attention Convolutional Network for Hyperspectral Image Super-Resolution 5 Jun 2025 · 1 repository · arXiv:2506.05041
-
Nonlinear Causal Discovery for Grouped Data 5 Jun 2025 · 0 repositories · arXiv:2506.05120
-
On the Convergence of Gradient Descent on Learning Transformers with Residual Connections 5 Jun 2025 · 0 repositories · arXiv:2506.05249