Methods › General › Output Functions › Softmax › Papers, page 39
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 39 of 375: papers 3,801 to 3,900 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Pretrained LLMs as Real-Time Controllers for Robot Operated Serial Production Line 5 Mar 2025 · 0 repositories · arXiv:2503.03889
-
Qieemo: Speech Is All You Need in the Emotion Recognition in Conversations 5 Mar 2025 · 0 repositories · arXiv:2503.22687
-
RiskAgent: Autonomous Medical AI Copilot for Generalist Risk Prediction 5 Mar 2025 · 0 repositories · arXiv:2503.03802
-
RGB-Thermal Infrared Fusion for Robust Depth Estimation in Complex Environments 5 Mar 2025 · 0 repositories · arXiv:2503.04821
-
RVAFM: Re-parameterizing Vertical Attention Fusion Module for Handwritten Paragraph Text Recognition 5 Mar 2025 · 0 repositories · arXiv:2503.03104
-
Sarcasm Detection as a Catalyst: Improving Stance Detection with Cross-Target Capabilities 5 Mar 2025 · 0 repositories · arXiv:2503.03787
-
ScaleFusionNet: Transformer-Guided Multi-Scale Feature Fusion for Skin Lesion Segmentation 5 Mar 2025 · 1 repository · arXiv:2503.03327
-
See What You Are Told: Visual Attention Sink in Large Multimodal Models 5 Mar 2025 · 0 repositories · arXiv:2503.03321
-
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation 5 Mar 2025 · 1 repository · arXiv:2503.03308
-
The Signed Two-Space Proximity Model for Learning Representations in Protein-Protein Interaction Networks 5 Mar 2025 · 0 repositories · arXiv:2503.03904
-
A Joint Visual Compression and Perception Framework for Neuralmorphic Spiking Camera 4 Mar 2025 · 0 repositories · arXiv:2503.02725
-
A Transformer Model for Predicting Chemical Reaction Products from Generic Templates 4 Mar 2025 · 0 repositories · arXiv:2503.05810
-
Adapting Decoder-Based Language Models for Diverse Encoder Downstream Tasks 4 Mar 2025 · 0 repositories · arXiv:2503.02656
-
Attention Bootstrapping for Multi-Modal Test-Time Adaptation 4 Mar 2025 · 0 repositories · arXiv:2503.02221
-
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE) 4 Mar 2025 · 0 repositories · arXiv:2503.02360
-
BHViT: Binarized Hybrid Vision Transformer 4 Mar 2025 · 1 repository · arXiv:2503.02394Syntology official (archive's flag): 16 ran · 16 ran (of which 13 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 29 harvested samples)
-
Boltzmann Attention Sampling for Image Analysis with Small Objects 4 Mar 2025 · 0 repositories · arXiv:2503.02841
-
Controllable Motion Generation via Diffusion Modal Coupling 4 Mar 2025 · 1 repository · arXiv:2503.02353
-
CoServe: Efficient Collaboration-of-Experts (CoE) Model Inference with Limited Memory 4 Mar 2025 · 0 repositories · arXiv:2503.02354
-
CrystalFramer: Rethinking the Role of Frames for SE(3)-Invariant Crystal Structure Modeling 4 Mar 2025 · 0 repositories · arXiv:2503.02209
-
Developing a PET/CT Foundation Model for Cross-Modal Anatomical and Functional Imaging 4 Mar 2025 · 0 repositories · arXiv:2503.02824
-
Disentangled Knowledge Tracing for Alleviating Cognitive Bias 4 Mar 2025 · 2 repositories · arXiv:2503.02539
-
Effectively Steer LLM To Follow Preference via Building Confident Directions 4 Mar 2025 · 0 repositories · arXiv:2503.02989
-
LREA: Low-Rank Efficient Attention on Modeling Long-Term User Behaviors for CTR Prediction 4 Mar 2025 · 0 repositories · arXiv:2503.02542
-
Exploring Token-Level Augmentation in Vision Transformer for Semi-Supervised Semantic Segmentation 4 Mar 2025 · 1 repository · arXiv:2503.02459
-
Extrapolating the long-term seasonal component of electricity prices for forecasting in the day-ahead market 4 Mar 2025 · 0 repositories · arXiv:2503.02518
-
Fair Play in the Fast Lane: Integrating Sportsmanship into Autonomous Racing Systems 4 Mar 2025 · 0 repositories · arXiv:2503.03774
-
FourierNAT: A Fourier-Mixing-Based Non-Autoregressive Transformer for Parallel Sequence Generation 4 Mar 2025 · 0 repositories · arXiv:2503.07630
-
Graph Transformer with Disease Subgraph Positional Encoding for Improved Comorbidity Prediction 4 Mar 2025 · 1 repository · arXiv:2503.03046
-
Haste Makes Waste: Evaluating Planning Abilities of LLMs for Efficient and Feasible Multitasking with Time Constraints Between Actions 4 Mar 2025 · 1 repository · arXiv:2503.02238
-
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02917
-
JPDS-NN: Reinforcement Learning-Based Dynamic Task Allocation for Agricultural Vehicle Routing Optimization 4 Mar 2025 · 0 repositories · arXiv:2503.02369
-
LADM: Long-context Training Data Selection with Attention-based Dependency Measurement for LLMs 4 Mar 2025 · 0 repositories · arXiv:2503.02502
-
Learning Precoding in Multi-user Multi-antenna Systems: Transformer or Graph Transformer? 4 Mar 2025 · 0 repositories · arXiv:2503.02998
-
LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning 4 Mar 2025 · 0 repositories · arXiv:2503.04812
-
LLM Misalignment via Adversarial RLHF Platforms 4 Mar 2025 · 0 repositories · arXiv:2503.03039
-
Multilingualism, Transnationality, and K-pop in the Online #StopAsianHate Movement 4 Mar 2025 · 1 repository · arXiv:2503.02707
-
Network Traffic Classification Using Machine Learning, Transformer, and Large Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02141
-
NodeNAS: Node-Specific Graph Neural Architecture Search for Out-of-Distribution Generalization 4 Mar 2025 · 0 repositories · arXiv:2503.02448
-
Numerical methods for two-dimensional G-heat equation 4 Mar 2025 · 0 repositories · arXiv:2503.02395
-
Optimizing open-domain question answering with graph-based retrieval augmented generation 4 Mar 2025 · 0 repositories · arXiv:2503.02922
-
PanguIR Technical Report for NTCIR-18 AEOLLM Task 4 Mar 2025 · 0 repositories · arXiv:2503.04809
-
PennyLang: Pioneering LLM-Based Quantum Code Generation with a Novel PennyLane-Centric Dataset 4 Mar 2025 · 0 repositories · arXiv:2503.02497
-
Q-Filters: Leveraging QK Geometry for Efficient KV Cache Compression 4 Mar 2025 · 1 repository · arXiv:2503.02812
-
RACNN: Residual Attention Convolutional Neural Network for Near-Field Channel Estimation in 6G Wireless Communications 4 Mar 2025 · 1 repository · arXiv:2503.02299
-
Remote Sensing Image Classification Using Convolutional Neural Network (CNN) and Transfer Learning Techniques 4 Mar 2025 · 0 repositories · arXiv:2503.02510
-
Resource-Efficient Affordance Grounding with Complementary Depth and Semantic Prompts 4 Mar 2025 · 1 repository · arXiv:2503.02600
-
Seeing is Understanding: Unlocking Causal Attention into Modality-Mutual Attention for Multimodal LLMs 4 Mar 2025 · 1 repository · arXiv:2503.02597Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Sparse Meets Dense: Unified Generative Recommendations with Cascaded Sparse-Dense Representations 4 Mar 2025 · 0 repositories · arXiv:2503.02453
-
STAA-SNN: Spatial-Temporal Attention Aggregator for Spiking Neural Networks 4 Mar 2025 · 0 repositories · arXiv:2503.02689
-
Tabby: Tabular Data Synthesis with Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02152
-
Target Return Optimizer for Multi-Game Decision Transformer 4 Mar 2025 · 0 repositories · arXiv:2503.02311
-
TeTRA-VPR: A Ternary Transformer Approach for Compact Visual Place Recognition 4 Mar 2025 · 0 repositories · arXiv:2503.02511
-
Towards Robust Multi-UAV Collaboration: MARL with Noise-Resilient Communication and Attention Mechanisms 4 Mar 2025 · 1 repository · arXiv:2503.02913
-
Union of Experts: Adapting Hierarchical Routing to Equivalently Decomposed Transformer 4 Mar 2025 · 1 repository · arXiv:2503.02495
-
Use Me Wisely: AI-Driven Assessment for LLM Prompting Skills Development 4 Mar 2025 · 0 repositories · arXiv:2503.02532
-
Weak-to-Strong Generalization Even in Random Feature Networks, Provably 4 Mar 2025 · 0 repositories · arXiv:2503.02877
-
Wikipedia in the Era of LLMs: Evolution and Risks 4 Mar 2025 · 1 repository · arXiv:2503.02879
-
Wyckoff Transformer: Generation of Symmetric Crystals 4 Mar 2025 · 1 repository · arXiv:2503.02407
-
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders 4 Mar 2025 · 0 repositories · arXiv:2503.02993
-
From Claims to Evidence: A Unified Framework and Critical Analysis of CNN vs. Transformer vs. Mamba in Medical Image Segmentation 3 Mar 2025 · 1 repository · arXiv:2503.01306
-
Neural ODE Transformers: Analyzing Internal Dynamics and Adaptive Fine-tuning 3 Mar 2025 · 0 repositories · arXiv:2503.01329Syntology 0 ran · 1 unverified (of 1 harvested sample)
-
Enhancing Social Media Rumor Detection: A Semantic and Graph Neural Network Approach for the 2024 Global Election 3 Mar 2025 · 0 repositories · arXiv:2503.01394
-
AC-Lite : A Lightweight Image Captioning Model for Low-Resource Assamese Language 3 Mar 2025 · 0 repositories · arXiv:2503.01453
-
SrSv: Integrating Sequential Rollouts with Sequential Value Estimation for Multi-agent Reinforcement Learning 3 Mar 2025 · 0 repositories · arXiv:2503.01458
-
Liger: Linearizing Large Language Models to Gated Recurrent Structures 3 Mar 2025 · 1 repository · arXiv:2503.01496Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
An Efficient Approach to Detecting Lung Nodules Using Swin Transformer 3 Mar 2025 · 0 repositories · arXiv:2503.01592
-
Machine Learners Should Acknowledge the Legal Implications of Large Language Models as Personal Data 3 Mar 2025 · 0 repositories · arXiv:2503.01630
-
SAGE: A Framework of Precise Retrieval for RAG 3 Mar 2025 · 0 repositories · arXiv:2503.01713
-
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01814
-
A Generalized Theory of Mixup for Structure-Preserving Synthetic Data 3 Mar 2025 · 1 repository · arXiv:2503.02645
-
A Hybrid CNN-Transformer Model for Heart Disease Prediction Using Life History Data 3 Mar 2025 · 0 repositories · arXiv:2503.02124
-
ACCORD: Alleviating Concept Coupling through Dependence Regularization for Text-to-Image Diffusion Personalization 3 Mar 2025 · 0 repositories · arXiv:2503.01122
-
Architectural and Inferential Inductive Biases For Exchangeable Sequence Modeling 3 Mar 2025 · 1 repository · arXiv:2503.01215
-
AskToAct: Enhancing LLMs Tool Use via Self-Correcting Clarification 3 Mar 2025 · 0 repositories · arXiv:2503.01940
-
Attention Condensation via Sparsity Induced Regularized Training 3 Mar 2025 · 0 repositories · arXiv:2503.01564
-
Boolean-aware Attention for Dense Retrieval 3 Mar 2025 · 0 repositories · arXiv:2503.01753
-
Cancer Type, Stage and Prognosis Assessment from Pathology Reports using LLMs 3 Mar 2025 · 1 repository · arXiv:2503.01194
-
Dementia Insights: A Context-Based MultiModal Approach 3 Mar 2025 · 0 repositories · arXiv:2503.01226
-
Efficient or Powerful? Trade-offs Between Machine Learning and Deep Learning for Mental Illness Detection on Social Media 3 Mar 2025 · 0 repositories · arXiv:2503.01082
-
Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond 3 Mar 2025 · 0 repositories · arXiv:2503.01210
-
Fault Localization and State Estimation of Power Grid under Parallel Cyber-Physical Attacks 3 Mar 2025 · 0 repositories · arXiv:2503.05797
-
Forgetting Transformer: Softmax Attention with a Forget Gate 3 Mar 2025 · 1 repository · arXiv:2503.02130Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
GRAIN: Exact Graph Reconstruction from Gradients 3 Mar 2025 · 1 repository · arXiv:2503.01838Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
HanDrawer: Leveraging Spatial Information to Render Realistic Hands Using a Conditional Diffusion Model in Single Stage 3 Mar 2025 · 0 repositories · arXiv:2503.02127
-
HeterRec: Heterogeneous Information Transformer for Scalable Sequential Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01469
-
HoH: A Dynamic Benchmark for Evaluating the Impact of Outdated Information on Retrieval-Augmented Generation 3 Mar 2025 · 0 repositories · arXiv:2503.04800Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
HOP: Heterogeneous Topology-based Multimodal Entanglement for Co-Speech Gesture Generation 3 Mar 2025 · 0 repositories · arXiv:2503.01175
-
How simple can you go? An off-the-shelf transformer approach to molecular dynamics 3 Mar 2025 · 1 repository · arXiv:2503.01431
-
Interactive Gadolinium-Free MRI Synthesis: A Transformer with Localization Prompt Learning 3 Mar 2025 · 1 repository · arXiv:2503.01265
-
Label Ranker: Self-Aware Preference for Classification Label Position in Visual Masked Self-Supervised Pre-Trained Model 3 Mar 2025 · 1 repository
-
Linear Representations of Political Perspective Emerge in Large Language Models 3 Mar 2025 · 1 repository · arXiv:2503.02080
-
MAPS: Motivation-Aware Personalized Search via LLM-Driven Consultation Alignment 3 Mar 2025 · 1 repository · arXiv:2503.01711Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
MeshPad: Interactive Sketch-Conditioned Artist-Designed Mesh Generation and Editing 3 Mar 2025 · 0 repositories · arXiv:2503.01425
-
MI-DETR: An Object Detection Model with Multi-time Inquiries Mechanism 3 Mar 2025 · 1 repository · arXiv:2503.01463
-
MRI super-resolution reconstruction using efficient diffusion probabilistic model with residual shifting 3 Mar 2025 · 1 repository · arXiv:2503.01576
-
Object-Aware Video Matting with Cross-Frame Guidance 3 Mar 2025 · 0 repositories · arXiv:2503.01262
-
Open-Set Recognition of Novel Species in Biodiversity Monitoring 3 Mar 2025 · 0 repositories · arXiv:2503.01691
-
Primer C-VAE: An interpretable deep learning primer design method to detect emerging virus variants 3 Mar 2025 · 0 repositories · arXiv:2503.01459
-
Primus: Enforcing Attention Usage for 3D Medical Image Segmentation 3 Mar 2025 · 0 repositories · arXiv:2503.01835