Methods › General › Output Functions › Softmax › Papers, page 34
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 34 of 375: papers 3,301 to 3,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Humanoid Policy ~ Human Policy 17 Mar 2025 · 0 repositories · arXiv:2503.13441
-
In-Context Linear Regression Demystified: Training Dynamics and Mechanistic Interpretability of Multi-Head Softmax Attention 17 Mar 2025 · 1 repository · arXiv:2503.12734
-
Intra-neuronal attention within language models Relationships between activation and semantics 17 Mar 2025 · 0 repositories · arXiv:2503.12992
-
KVShare: An LLM Service System with Efficient and Effective Multi-Tenant KV Cache Reuse 17 Mar 2025 · 0 repositories · arXiv:2503.16525
-
Let Synthetic Data Shine: Domain Reassembly and Soft-Fusion for Single Domain Generalization 17 Mar 2025 · 0 repositories · arXiv:2503.13617
-
MaTVLM: Hybrid Mamba-Transformer for Efficient Vision-Language Modeling 17 Mar 2025 · 1 repository · arXiv:2503.13440Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
MES-RAG: Bringing Multi-modal, Entity-Storage, and Secure Enhancements to RAG 17 Mar 2025 · 1 repository · arXiv:2503.13563
-
Mitigating Visual Forgetting via Take-along Visual Conditioning for Multi-modal Long CoT Reasoning 17 Mar 2025 · 0 repositories · arXiv:2503.13360
-
OKRA: an Explainable, Heterogeneous, Multi-Stakeholder Job Recommender System 17 Mar 2025 · 0 repositories · arXiv:2504.07108
-
OSCAR: Online Soft Compression And Reranking 17 Mar 2025 · 0 repositories · arXiv:2504.07109
-
OSLO-IC: On-the-Sphere Learned Omnidirectional Image Compression with Attention Modules and Spatial Context 17 Mar 2025 · 0 repositories · arXiv:2503.13119
-
PAUSE: Low-Latency and Privacy-Aware Active User Selection for Federated Learning 17 Mar 2025 · 1 repository · arXiv:2503.13173
-
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval 17 Mar 2025 · 0 repositories · arXiv:2503.15548
-
Robust Audio-Visual Segmentation via Audio-Guided Visual Convergent Alignment 17 Mar 2025 · 0 repositories · arXiv:2503.12847
-
SeisRDT: Latent Diffusion Model Based On Representation Learning For Seismic Data Interpolation And Reconstruction 17 Mar 2025 · 0 repositories · arXiv:2503.21791
-
Synchronous vs Asynchronous Reinforcement Learning in a Real World Robot 17 Mar 2025 · 0 repositories · arXiv:2503.14554
-
Towards Scalable Foundation Model for Multi-modal and Hyperspectral Geospatial Data 17 Mar 2025 · 0 repositories · arXiv:2503.12843
-
Unlock Pose Diversity: Accurate and Efficient Implicit Keypoint-based Spatiotemporal Diffusion for Audio-driven Talking Portrait 17 Mar 2025 · 1 repository · arXiv:2503.12963
-
VeriContaminated: Assessing LLM-Driven Verilog Coding for Data Contamination 17 Mar 2025 · 0 repositories · arXiv:2503.13572
-
Atlas: Multi-Scale Attention Improves Long Context Image Modeling 16 Mar 2025 · 1 repository · arXiv:2503.12355
-
Deepfake Detection with Optimized Hybrid Model: EAR Biometric Descriptor via Improved RCNN 16 Mar 2025 · 0 repositories · arXiv:2503.12381
-
Fourier-Based 3D Multistage Transformer for Aberration Correction in Multicellular Specimens 16 Mar 2025 · 2 repositories · arXiv:2503.12593
-
Fragile Mastery: Are Domain-Specific Trade-Offs Undermining On-Device Language Models? 16 Mar 2025 · 0 repositories · arXiv:2503.22698
-
GCBLANE: A graph-enhanced convolutional BiLSTM attention network for improved transcription factor binding site prediction 16 Mar 2025 · 1 repository · arXiv:2503.12377
-
GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation 16 Mar 2025 · 0 repositories · arXiv:2503.12600
-
GS-I³: Gaussian Splatting for Surface Reconstruction from Illumination-Inconsistent Images 16 Mar 2025 · 1 repository · arXiv:2503.12335
-
HyperKAN: Hypergraph Representation Learning with Kolmogorov-Arnold Networks 16 Mar 2025 · 0 repositories · arXiv:2503.12365
-
MambaIC: State Space Models for High-Performance Learned Image Compression 16 Mar 2025 · 1 repository · arXiv:2503.12461
-
MAVEN: Multi-modal Attention for Valence-Arousal Emotion Network 16 Mar 2025 · 1 repository · arXiv:2503.12623
-
Modality-Composable Diffusion Policy via Inference-Time Distribution-level Composition 16 Mar 2025 · 1 repository · arXiv:2503.12466
-
MSCMHMST: A traffic flow prediction model based on Transformer 16 Mar 2025 · 0 repositories · arXiv:2503.13540
-
SAM2-ELNet: Label Enhancement and Automatic Annotation for Remote Sensing Segmentation 16 Mar 2025 · 0 repositories · arXiv:2503.12404
-
Semantic Matters: Multimodal Features for Affective Analysis 16 Mar 2025 · 0 repositories · arXiv:2504.11460
-
State Fourier Diffusion Language Model (SFDLM): A Scalable, Novel Iterative Approach to Language Modeling 16 Mar 2025 · 0 repositories · arXiv:2503.17382
-
TuneNSearch: a hybrid transfer learning and local search approach for solving vehicle routing problems 16 Mar 2025 · 0 repositories · arXiv:2503.12662
-
PA-CFL: Privacy-Adaptive Clustered Federated Learning for Transformer-Based Sales Forecasting on Heterogeneous Retail Data 15 Mar 2025 · 0 repositories · arXiv:2503.12220
-
Applications of Large Language Model Reasoning in Feature Generation 15 Mar 2025 · 0 repositories · arXiv:2503.11989
-
Att-Adapter: A Robust and Precise Domain-Specific Multi-Attributes T2I Diffusion Adapter via Conditional Variational Autoencoder 15 Mar 2025 · 0 repositories · arXiv:2503.11937
-
Changing Base Without Losing Pace: A GPU-Efficient Alternative to MatMul in DNNs 15 Mar 2025 · 0 repositories · arXiv:2503.12211
-
Cognitive Activation and Chaotic Dynamics in Large Language Models: A Quasi-Lyapunov Analysis of Reasoning Mechanisms 15 Mar 2025 · 0 repositories · arXiv:2503.13530
-
Design of an Expression Recognition Solution Employing the Global Channel-Spatial Attention Mechanism 15 Mar 2025 · 0 repositories · arXiv:2503.11935
-
EHNet: An Efficient Hybrid Network for Crowd Counting and Localization 15 Mar 2025 · 0 repositories · arXiv:2503.12061
-
Fast Critical Clearing Time Calculation for Power Systems with Synchronous and Asynchronous Generation 15 Mar 2025 · 0 repositories · arXiv:2503.12132
-
From Laboratory to Real World: A New Benchmark Towards Privacy-Preserved Visible-Infrared Person Re-Identification 15 Mar 2025 · 0 repositories · arXiv:2503.12232
-
United we stand, Divided we fall: Handling Weak Complementary Relationships for Audio-Visual Emotion Recognition in Valence-Arousal Space 15 Mar 2025 · 0 repositories · arXiv:2503.12261
-
Integrating Chain-of-Thought and Retrieval Augmented Generation Enhances Rare Disease Diagnosis from Clinical Notes 15 Mar 2025 · 0 repositories · arXiv:2503.12286
-
Language Models for Automated Classification of Brain MRI Reports and Growth Chart Generation 15 Mar 2025 · 0 repositories · arXiv:2503.12143
-
Leveraging Motion Information for Better Self-Supervised Video Correspondence Learning 15 Mar 2025 · 0 repositories · arXiv:2503.12026
-
LLM & HPC:Benchmarking DeepSeek's Performance in High-Performance Computing Tasks 15 Mar 2025 · 1 repository · arXiv:2504.03665
-
Maritime Mission Planning for Unmanned Surface Vessel using Large Language Model 15 Mar 2025 · 0 repositories · arXiv:2503.12065
-
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models 15 Mar 2025 · 1 repository · arXiv:2503.12096
-
Tailor: An Integrated Text-Driven CG-Ready Human and Garment Generation System 15 Mar 2025 · 0 repositories · arXiv:2503.12052
-
Towards Learning High-Precision Least Squares Algorithms with Sequence Models 15 Mar 2025 · 1 repository · arXiv:2503.12295Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 10 harvested samples)
-
VeriMind: Agentic LLM for Automated Verilog Generation with a Novel Evaluation Metric 15 Mar 2025 · 0 repositories · arXiv:2503.16514
-
VTON 360: High-Fidelity Virtual Try-On from Any Viewing Direction 15 Mar 2025 · 0 repositories · arXiv:2503.12165
-
Weighted Graph Structure Learning with Attention Denoising for Node Classification 15 Mar 2025 · 1 repository · arXiv:2503.12157
-
Winning the MIDST Challenge: New Membership Inference Attacks on Diffusion Models for Tabular Data Synthesis 15 Mar 2025 · 1 repository · arXiv:2503.12008
-
A Neural Network Architecture Based on Attention Gate Mechanism for 3D Magnetotelluric Forward Modeling 14 Mar 2025 · 0 repositories · arXiv:2503.11408
-
A Review of DeepSeek Models' Key Innovative Techniques 14 Mar 2025 · 0 repositories · arXiv:2503.11486
-
A Survey of Cross-domain Graph Learning: Progress and Future Directions 14 Mar 2025 · 1 repository · arXiv:2503.11086
-
Addressing Information Loss and Interaction Collapse: A Dual Enhanced Attention Framework for Feature Interaction 14 Mar 2025 · 0 repositories · arXiv:2503.11233
-
Advanced Deep Learning Methods for Protein Structure Prediction and Design 14 Mar 2025 · 0 repositories · arXiv:2503.13522
-
Advancing 3D Gaussian Splatting Editing with Complementary and Consensus Information 14 Mar 2025 · 0 repositories · arXiv:2503.11601
-
Alzheimer's Disease Classification Using Retinal OCT: TransnetOCT and Swin Transformer Models 14 Mar 2025 · 0 repositories · arXiv:2503.11511
-
APLA: A Simple Adaptation Method for Vision Transformers 14 Mar 2025 · 1 repository · arXiv:2503.11335
-
Asynchronous Sharpness-Aware Minimization For Fast and Accurate Deep Learning 14 Mar 2025 · 0 repositories · arXiv:2503.11147
-
Augmenting Image Annotation: A Human-LMM Collaborative Framework for Efficient Object Selection and Label Generation 14 Mar 2025 · 0 repositories · arXiv:2503.11096
-
BannerAgency: Advertising Banner Design with Multimodal LLM Agents 14 Mar 2025 · 0 repositories · arXiv:2503.11060
-
BEVDiffLoc: End-to-End LiDAR Global Localization in BEV View based on Diffusion Model 14 Mar 2025 · 1 repository · arXiv:2503.11372
-
Bottom-up Iterative Anomalous Diffusion Detector (BI-ADD) 14 Mar 2025 · 1 repository · arXiv:2503.11529
-
Brain Effective Connectivity Estimation via Fourier Spatiotemporal Attention 14 Mar 2025 · 1 repository · arXiv:2503.11283
-
Cardiomyopathy Diagnosis Model from Endomyocardial Biopsy Specimens: Appropriate Feature Space and Class Boundary in Small Sample Size Data 14 Mar 2025 · 0 repositories · arXiv:2503.11331
-
Combining Causal Models for More Accurate Abstractions of Neural Networks 14 Mar 2025 · 1 repository · arXiv:2503.11429Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Context-Aware Rule Mining Using a Dynamic Transformer-Based Framework 14 Mar 2025 · 0 repositories · arXiv:2503.11125
-
DCAT: Dual Cross-Attention Fusion for Disease Classification in Radiological Images with Uncertainty Estimation 14 Mar 2025 · 0 repositories · arXiv:2503.11851
-
Direction-Aware Diagonal Autoregressive Image Generation 14 Mar 2025 · 0 repositories · arXiv:2503.11129
-
Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models 14 Mar 2025 · 0 repositories · arXiv:2503.11154
-
DynRsl-VLM: Enhancing Autonomous Driving Perception with Dynamic Resolution Vision-Language Models 14 Mar 2025 · 0 repositories · arXiv:2503.11265
-
Enhanced Multi-View Pedestrian Detection Using Probabilistic Occupancy Volume 14 Mar 2025 · 0 repositories · arXiv:2503.10982
-
Exploring Competitive and Collusive Behaviors in Algorithmic Pricing with Deep Reinforcement Learning 14 Mar 2025 · 0 repositories · arXiv:2503.11270
-
Exploring the Potential of Large Multimodal Models as Effective Alternatives for Pronunciation Assessment 14 Mar 2025 · 0 repositories · arXiv:2503.11229
-
FMNet: Frequency-Assisted Mamba-Like Linear Attention Network for Camouflaged Object Detection 14 Mar 2025 · 0 repositories · arXiv:2503.11030
-
From Pixels to Histopathology: A Graph-Based Framework for Interpretable Whole Slide Image Analysis 14 Mar 2025 · 1 repository · arXiv:2503.11846
-
GaussianIP: Identity-Preserving Realistic 3D Human Generation via Human-Centric Diffusion Prior 14 Mar 2025 · 1 repository · arXiv:2503.11143
-
Image-Goal Navigation Using Refined Feature Guidance and Scene Graph Enhancement 14 Mar 2025 · 1 repository · arXiv:2503.10986
-
Key, Value, Compress: A Systematic Exploration of KV Cache Compression Techniques 14 Mar 2025 · 0 repositories · arXiv:2503.11816
-
Time and Memory Trade-off of KV-Cache Compression in Tensor Transformer Decoding 14 Mar 2025 · 0 repositories · arXiv:2503.11108
-
LLaVA-MLB: Mitigating and Leveraging Attention Bias for Training-Free Video LLMs 14 Mar 2025 · 0 repositories · arXiv:2503.11205
-
Making Every Step Effective: Jailbreaking Large Vision-Language Models Through Hierarchical KV Equalization 14 Mar 2025 · 0 repositories · arXiv:2503.11750
-
MEET: A Million-Scale Dataset for Fine-Grained Geospatial Scene Classification with Zoom-Free Remote Sensing Imagery 14 Mar 2025 · 0 repositories · arXiv:2503.11219
-
Modeling and Optimization for Flexible Cylindrical Arrays-Enabled Wireless Communications 14 Mar 2025 · 1 repository · arXiv:2503.11123
-
MTV-Inpaint: Multi-Task Long Video Inpainting 14 Mar 2025 · 0 repositories · arXiv:2503.11412
-
Multi-View Industrial Anomaly Detection with Epipolar Constrained Cross-View Fusion 14 Mar 2025 · 0 repositories · arXiv:2503.11088
-
Open3DVQA: A Benchmark for Comprehensive Spatial Reasoning with Multimodal Large Language Model in Open Space 14 Mar 2025 · 1 repository · arXiv:2503.11094
-
PARIC: Probabilistic Attention Regularization for Language Guided Image Classification from Pre-trained Vison Language Models 14 Mar 2025 · 0 repositories · arXiv:2503.11360
-
Prof. Robot: Differentiable Robot Rendering Without Static and Self-Collisions 14 Mar 2025 · 1 repository · arXiv:2503.11269
-
Prompt Sentiment: The Catalyst for LLM Change 14 Mar 2025 · 0 repositories · arXiv:2503.13510
-
Quantifying Interpretability in CLIP Models with Concept Consistency 14 Mar 2025 · 0 repositories · arXiv:2503.11103
-
RAG-KG-IL: A Multi-Agent Hybrid Framework for Reducing Hallucinations and Enhancing LLM Reasoning through RAG and Incremental Knowledge Graph Learning Integration 14 Mar 2025 · 0 repositories · arXiv:2503.13514
-
Relevance Isn't All You Need: Scaling RAG Systems With Inference-Time Compute Via Multi-Criteria Reranking 14 Mar 2025 · 2 repositories · arXiv:2504.07104