Methods › General › Output Functions › Softmax › Papers, page 12
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 12 of 375: papers 1,101 to 1,200 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
The Atlas of In-Context Learning: How Attention Heads Shape In-Context Retrieval Augmentation 21 May 2025 · 1 repository · arXiv:2505.15807Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Time Tracker: Mixture-of-Experts-Enhanced Foundation Time Series Forecasting Model with Decoupled Training Pipelines 21 May 2025 · 0 repositories · arXiv:2505.15151
-
UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning 21 May 2025 · 0 repositories · arXiv:2505.15725
-
Unified Cross-Modal Attention-Mixer Based Structural-Functional Connectomics Fusion for Neuropsychiatric Disorder Diagnosis 21 May 2025 · 0 repositories · arXiv:2505.15139
-
When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning 21 May 2025 · 0 repositories · arXiv:2505.15276
-
FlowBERT: Prompt-tuned BERT for variable flow field prediction 20 May 2025 · 0 repositories · arXiv:2506.08021
-
A Direct Comparison of Simultaneously Recorded Scalp, Around-Ear, and In-Ear EEG for Neural Selective Auditory Attention Decoding to Speech 20 May 2025 · 0 repositories · arXiv:2505.14478
-
A Logic of General Attention Using Edge-Conditioned Event Models (Extended Version) 20 May 2025 · 0 repositories · arXiv:2505.14539
-
AAPO: Enhance the Reasoning Capabilities of LLMs with Advantage Momentum 20 May 2025 · 0 repositories · arXiv:2505.14264
-
AI-empowered Channel Estimation for Block-based Active IRS-enhanced Hybrid-field IoT Network 20 May 2025 · 0 repositories · arXiv:2505.14098
-
Aligning Attention Distribution to Information Flow for Hallucination Mitigation in Large Vision-Language Models 20 May 2025 · 0 repositories · arXiv:2505.14257
-
AppleGrowthVision: A large-scale stereo dataset for phenological analysis, fruit detection, and 3D reconstruction in apple orchards 20 May 2025 · 0 repositories · arXiv:2505.14029
-
Articulatory Feature Prediction from Surface EMG during Speech Production 20 May 2025 · 1 repository · arXiv:2505.13814
-
Automated Quality Evaluation of Cervical Cytopathology Whole Slide Images Based on Content Analysis 20 May 2025 · 0 repositories · arXiv:2505.13875
-
Automatic Dataset Generation for Knowledge Intensive Question Answering Tasks 20 May 2025 · 0 repositories · arXiv:2505.14212
-
Beyond Text: Unveiling Privacy Vulnerabilities in Multi-modal Retrieval-Augmented Generation 20 May 2025 · 0 repositories · arXiv:2505.13957
-
Breaking Bad Tokens: Detoxification of LLMs Using Sparse Autoencoders 20 May 2025 · 0 repositories · arXiv:2505.14536Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CAD-Coder: An Open-Source Vision-Language Model for Computer-Aided Design Code Generation 20 May 2025 · 1 repository · arXiv:2505.14646Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Choosing a Model, Shaping a Future: Comparing LLM Perspectives on Sustainability and its Relationship with AI 20 May 2025 · 0 repositories · arXiv:2505.14435
-
CONSIGN: Conformal Segmentation Informed by Spatial Groupings via Decomposition 20 May 2025 · 0 repositories · arXiv:2505.14113Syntology 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Cost-Augmented Monte Carlo Tree Search for LLM-Assisted Planning 20 May 2025 · 0 repositories · arXiv:2505.14656
-
CSAGC-IDS: A Dual-Module Deep Learning Network Intrusion Detection Model for Complex and Imbalanced Data 20 May 2025 · 0 repositories · arXiv:2505.14027
-
Disentangled Multi-span Evolutionary Network against Temporal Knowledge Graph Reasoning 20 May 2025 · 0 repositories · arXiv:2505.14020
-
Divide by Question, Conquer by Agent: SPLIT-RAG with Question-Driven Graph Partitioning 20 May 2025 · 0 repositories · arXiv:2505.13994
-
Do Language Models Use Their Depth Efficiently? 20 May 2025 · 1 repository · arXiv:2505.13898Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
DSMentor: Enhancing Data Science Agents with Curriculum Learning and Online Knowledge Accumulation 20 May 2025 · 0 repositories · arXiv:2505.14163
-
EEG-to-Text Translation: A Model for Deciphering Human Brain Activity 20 May 2025 · 1 repository · arXiv:2505.13936
-
EfficientLLM: Efficiency in Large Language Models 20 May 2025 · 0 repositories · arXiv:2505.13840
-
Embedded Mean Field Reinforcement Learning for Perimeter-defense Game 20 May 2025 · 0 repositories · arXiv:2505.14209
-
Energy-Efficient Deep Reinforcement Learning with Spiking Transformers 20 May 2025 · 0 repositories · arXiv:2505.14533
-
Enhancing Abstractive Summarization of Scientific Papers Using Structure Information 20 May 2025 · 1 repository · arXiv:2505.14179
-
EVA: Red-Teaming GUI Agents via Evolving Indirect Prompt Injection 20 May 2025 · 0 repositories · arXiv:2505.14289
-
Every Pixel Tells a Story: End-to-End Urdu Newspaper OCR 20 May 2025 · 0 repositories · arXiv:2505.13943
-
Exploring Image Quality Assessment from a New Perspective: Pupil Size 20 May 2025 · 0 repositories · arXiv:2505.13841
-
Exploring Jailbreak Attacks on LLMs through Intent Concealment and Diversion 20 May 2025 · 0 repositories · arXiv:2505.14316
-
FLASH-D: FlashAttention with Hidden Softmax Division 20 May 2025 · 0 repositories · arXiv:2505.14201
-
FlashKAT: Understanding and Addressing Performance Bottlenecks in the Kolmogorov-Arnold Transformer 20 May 2025 · 1 repository · arXiv:2505.13813
-
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning 20 May 2025 · 0 repositories · arXiv:2505.14139
-
Generative AI at the Crossroads: Light Bulb, Dynamo, or Microscope? 20 May 2025 · 0 repositories · arXiv:2505.14588
-
Grouping First, Attending Smartly: Training-Free Acceleration for Diffusion Transformers 20 May 2025 · 1 repository · arXiv:2505.14687
-
HausaNLP: Current Status, Challenges and Future Directions for Hausa Natural Language Processing 20 May 2025 · 0 repositories · arXiv:2505.14311
-
Informatics for Food Processing 20 May 2025 · 1 repository · arXiv:2505.17087
-
JOLT-SQL: Joint Loss Tuning of Text-to-SQL with Confusion-aware Noisy Schema Sampling 20 May 2025 · 1 repository · arXiv:2505.14305Syntology official: no sample here; runs from other or unrecorded repositories · 16 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 3 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 1 pointer-only (licence)
-
Large Language Model-Driven Distributed Integrated Multimodal Sensing and Semantic Communications 20 May 2025 · 0 repositories · arXiv:2505.18194
-
Latent Flow Transformer 20 May 2025 · 1 repository · arXiv:2505.14513
-
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models 20 May 2025 · 0 repositories · arXiv:2505.14759
-
Learning Spatio-Temporal Dynamics for Trajectory Recovery via Time-Aware Transformer 20 May 2025 · 2 repositories · arXiv:2505.13857
-
LOD1 3D City Model from LiDAR: The Impact of Segmentation Accuracy on Quality of Urban 3D Modeling and Morphology Extraction 20 May 2025 · 1 repository · arXiv:2505.14747
-
Low-Cost FlashAttention with Fused Exponential and Multiplication Hardware Operators 20 May 2025 · 0 repositories · arXiv:2505.14314
-
Mechanistic Fine-tuning for In-context Learning 20 May 2025 · 0 repositories · arXiv:2505.14233
-
MGStream: Motion-aware 3D Gaussian for Streamable Dynamic Scene Reconstruction 20 May 2025 · 1 repository · arXiv:2505.13839
-
ModRWKV: Transformer Multimodality in Linear Time 20 May 2025 · 1 repository · arXiv:2505.14505
-
MSDformer: Multi-scale Discrete Transformer For Time Series Generation 20 May 2025 · 0 repositories · arXiv:2505.14202
-
Multi-Channel Swin Transformer Framework for Bearing Remaining Useful Life Prediction 20 May 2025 · 0 repositories · arXiv:2505.14897
-
Multimodal RAG-driven Anomaly Detection and Classification in Laser Powder Bed Fusion using Large Language Models 20 May 2025 · 0 repositories · arXiv:2505.13828
-
OmniStyle: Filtering High Quality Style Transfer Data at Scale 20 May 2025 · 1 repository · arXiv:2505.14028
-
Out-of-Distribution Generalization of In-Context Learning: A Low-Dimensional Subspace Perspective 20 May 2025 · 0 repositories · arXiv:2505.14808
-
Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis 20 May 2025 · 0 repositories · arXiv:2505.14406Syntology 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey 20 May 2025 · 0 repositories · arXiv:2505.14340
-
Polar Sparsity: High Throughput Batched LLM Inferencing with Scalable Contextual Sparsity 20 May 2025 · 1 repository · arXiv:2505.14884Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Predicting Neo-Adjuvant Chemotherapy Response in Triple-Negative Breast Cancer Using Pre-Treatment Histopathologic Images 20 May 2025 · 0 repositories · arXiv:2505.14730
-
Probing BERT for German Compound Semantics 20 May 2025 · 0 repositories · arXiv:2505.14130
-
Process vs. Outcome Reward: Which is Better for Agentic RAG Reinforcement Learning 20 May 2025 · 1 repository · arXiv:2505.14069Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ReactDiff: Latent Diffusion for Facial Reaction Generation 20 May 2025 · 1 repository · arXiv:2505.14151
-
s3: You Don't Need That Much Data to Train a Search Agent via RL 20 May 2025 · 1 repository · arXiv:2505.14146Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
SAE-FiRE: Enhancing Earnings Surprise Predictions Through Sparse Autoencoder Feature Selection 20 May 2025 · 0 repositories · arXiv:2505.14420
-
SAFEPATH: Preventing Harmful Reasoning in Chain-of-Thought via Early Alignment 20 May 2025 · 0 repositories · arXiv:2505.14667
-
Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation 20 May 2025 · 1 repository · arXiv:2505.14821
-
Scale-invariant Attention 20 May 2025 · 0 repositories · arXiv:2505.17083
-
Scaling Laws for State Dynamics in Large Language Models 20 May 2025 · 0 repositories · arXiv:2505.14892
-
SCAN: Semantic Document Layout Analysis for Textual and Visual Retrieval-Augmented Generation 20 May 2025 · 0 repositories · arXiv:2505.14381
-
Selective Structured State Space for Multispectral-fused Small Target Detection 20 May 2025 · 0 repositories · arXiv:2505.14043
-
Self-Reasoning Language Models: Unfold Hidden Reasoning Chains with Few Reasoning Catalyst 20 May 2025 · 0 repositories · arXiv:2505.14116
-
Normalized Cut with Reinforcement Learning in Constrained Action Space 20 May 2025 · 0 repositories · arXiv:2505.13986
-
Spiking Neural Networks with Temporal Attention-Guided Adaptive Fusion for imbalanced Multi-modal Learning 20 May 2025 · 0 repositories · arXiv:2505.14535
-
SSPS: Self-Supervised Positive Sampling for Robust Self-Supervised Speaker Verification 20 May 2025 · 1 repository · arXiv:2505.14561
-
STree: Speculative Tree Decoding for Hybrid State-Space Models 20 May 2025 · 0 repositories · arXiv:2505.14969
-
Subquadratic Algorithms and Hardness for Attention with Any Temperature 20 May 2025 · 0 repositories · arXiv:2505.14840
-
TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis 20 May 2025 · 1 repository · arXiv:2505.14910
-
The Post Double LASSO for Efficiency Analysis 20 May 2025 · 0 repositories · arXiv:2505.14282
-
Towards Efficient Multi-Scale Deformable Attention on NPU 20 May 2025 · 0 repositories · arXiv:2505.14022
-
TRATES: Trait-Specific Rubric-Assisted Cross-Prompt Essay Scoring 20 May 2025 · 0 repositories · arXiv:2505.14577
-
Unlocking the Power of SAM 2 for Few-Shot Segmentation 20 May 2025 · 1 repository · arXiv:2505.14100Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Unsupervised Graph Clustering with Deep Structural Entropy 20 May 2025 · 1 repository · arXiv:2505.14040
-
UPTor: Unified 3D Human Pose Dynamics and Trajectory Prediction for Human-Robot Interaction 20 May 2025 · 0 repositories · arXiv:2505.14866
-
Vulnerability of Transfer-Learned Neural Networks to Data Reconstruction Attacks in Small-Data Regime 20 May 2025 · 1 repository · arXiv:2505.14323
-
XDementNET: An Explainable Attention Based Deep Convolutional Network to Detect Alzheimer Progression from MRI data 20 May 2025 · 1 repository · arXiv:2505.13906
-
A3 : an Analytical Low-Rank Approximation Framework for Attention 19 May 2025 · 0 repositories · arXiv:2505.12942
-
Accelerate TarFlow Sampling with GS-Jacobi Iteration 19 May 2025 · 1 repository · arXiv:2505.12849
-
Accelerating Adaptive Retrieval Augmented Generation via Instruction-Driven Representation Reduction of Retrieval Overlaps 19 May 2025 · 0 repositories · arXiv:2505.12731
-
Adaptive Tokenization: On the Hop-Overpriority Problem in Tokenized Graph Learning Models 19 May 2025 · 0 repositories · arXiv:2505.15845
-
AdaToken-3D: Dynamic Spatial Gating for Efficient 3D Large Multimodal-Models Reasoning 19 May 2025 · 0 repositories · arXiv:2505.12782
-
Adversarial Testing in LLMs: Insights into Decision-Making Vulnerabilities 19 May 2025 · 0 repositories · arXiv:2505.13195
-
AMAQA: A Metadata-based QA Dataset for RAG Systems 19 May 2025 · 0 repositories · arXiv:2505.13557
-
An Attentional Model of Time Discounting 19 May 2025 · 0 repositories · arXiv:2505.13016
-
Are Large Language Models Good at Detecting Propaganda? 19 May 2025 · 0 repositories · arXiv:2505.13706
-
Attention-based clustering 19 May 2025 · 0 repositories · arXiv:2505.13112
-
Benchmarking and Confidence Evaluation of LALMs For Temporal Reasoning 19 May 2025 · 1 repository · arXiv:2505.13115
-
Bridging the Modality Gap: Enhancing Channel Prediction with Semantically Aligned LLMs and Knowledge Distillation 19 May 2025 · 0 repositories · arXiv:2505.12729
-
CALM-PDE: Continuous and Adaptive Convolutions for Latent Space Modeling of Time-dependent PDEs 19 May 2025 · 1 repository · arXiv:2505.12944Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)