Methods › General › Output Functions › Softmax › Papers, page 31
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 31 of 375: papers 3,001 to 3,100 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fundamental Limits of Perfect Concept Erasure 25 Mar 2025 · 1 repository · arXiv:2503.20098
-
Gemma 3 Technical Report 25 Mar 2025 · 0 repositories · arXiv:2503.19786
-
GPT Meets Graphs and KAN Splines: Testing Novel Frameworks on Multitask Fine-Tuned GPT-2 with LoRA 25 Mar 2025 · 0 repositories · arXiv:2504.10490
-
Guidelines For The Choice Of The Baseline in XAI Attribution Methods 25 Mar 2025 · 1 repository · arXiv:2503.19813
-
Hierarchical Attention Network for Interpretable ECG-based Heart Disease Classification 25 Mar 2025 · 0 repositories · arXiv:2504.03703
-
Reverse-Engineering the Retrieval Process in GenIR Models 25 Mar 2025 · 1 repository · arXiv:2503.19715
-
Improved Alignment of Modalities in Large Vision Language Models 25 Mar 2025 · 0 repositories · arXiv:2503.19508
-
iNatAg: Multi-Class Classification Models Enabled by a Large-Scale Benchmark Dataset with 4.7M Images of 2,959 Crop and Weed Species 25 Mar 2025 · 1 repository · arXiv:2503.20068
-
Inference-Time Scaling for Flow Models via Stochastic Generation and Rollover Budget Forcing 25 Mar 2025 · 0 repositories · arXiv:2503.19385
-
Innate Reasoning is Not Enough: In-Context Learning Enhances Reasoning Large Language Models with Less Overthinking 25 Mar 2025 · 0 repositories · arXiv:2503.19602
-
LogQuant: Log-Distributed 2-Bit Quantization of KV Cache with Superior Accuracy Preservation 25 Mar 2025 · 1 repository · arXiv:2503.19950
-
M²CD: A Unified MultiModal Framework for Optical-SAR Change Detection with Mixture of Experts and Self-Distillation 25 Mar 2025 · 0 repositories · arXiv:2503.19406
-
Machine-assisted writing evaluation: Exploring pre-trained language models in analyzing argumentative moves 25 Mar 2025 · 0 repositories · arXiv:2503.19279
-
Mapping Technological Futures: Anticipatory Discourse Through Text Mining 25 Mar 2025 · 0 repositories · arXiv:2504.02853
-
Mask²DiT: Dual Mask-based Diffusion Transformer for Multi-Scene Long Video Generation 25 Mar 2025 · 0 repositories · arXiv:2503.19881
-
MATT-GS: Masked Attention-based 3DGS for Robot Perception and Object Detection 25 Mar 2025 · 0 repositories · arXiv:2503.19330
-
Membership Inference Attacks on Large-Scale Models: A Survey 25 Mar 2025 · 0 repositories · arXiv:2503.19338
-
No Black Box Anymore: Demystifying Clinical Predictive Modeling with Temporal-Feature Cross Attention Mechanism 25 Mar 2025 · 0 repositories · arXiv:2503.19285
-
One Framework to Rule Them All: Unifying RL-Based and RL-Free Methods in RLHF 25 Mar 2025 · 0 repositories · arXiv:2503.19523
-
Peer Disambiguation in Self-Reported Surveys using Graph Attention Networks 25 Mar 2025 · 1 repository · arXiv:2503.20076
-
Prompt-Guided Dual-Path UNet with Mamba for Medical Image Segmentation 25 Mar 2025 · 0 repositories · arXiv:2503.19589
-
RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation 25 Mar 2025 · 0 repositories · arXiv:2503.19510
-
Scaling Down Text Encoders of Text-to-Image Diffusion Models 25 Mar 2025 · 1 repository · arXiv:2503.19897
-
SCI-IDEA: Context-Aware Scientific Ideation Using Token and Sentence Embeddings 25 Mar 2025 · 0 repositories · arXiv:2503.19257
-
Social Network User Profiling for Anomaly Detection Based on Graph Neural Networks 25 Mar 2025 · 0 repositories · arXiv:2503.19380
-
Surg-3M: A Dataset and Foundation Model for Perception in Surgical Settings 25 Mar 2025 · 1 repository · arXiv:2503.19740
-
Taxonomy Inference for Tabular Data Using Large Language Models 25 Mar 2025 · 0 repositories · arXiv:2503.21810
-
The importance of exploration: Modelling site-constant foraging 25 Mar 2025 · 0 repositories · arXiv:2503.20086
-
Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better 25 Mar 2025 · 0 repositories · arXiv:2503.19904
-
TraF-Align: Trajectory-aware Feature Alignment for Asynchronous Multi-agent Perception 25 Mar 2025 · 1 repository · arXiv:2503.19391Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TRIDIS: A Comprehensive Medieval and Early Modern Corpus for HTR and NER 25 Mar 2025 · 0 repositories · arXiv:2503.22714
-
VGAT: A Cancer Survival Analysis Framework Transitioning from Generative Visual Question Answering to Genomic Reconstruction 25 Mar 2025 · 1 repository · arXiv:2503.19367
-
AIM2PC: Aerial Image to 3D Building Point Cloud Reconstruction 24 Mar 2025 · 0 repositories · arXiv:2503.18527
-
AMD-Hummingbird: Towards an Efficient Text-to-Video Model 24 Mar 2025 · 1 repository · arXiv:2503.18559
-
An optimal baseline selection methodology for data-driven damage detection and temperature compensation in acousto-ultrasonics 24 Mar 2025 · 0 repositories · arXiv:2504.03694
-
Analyzing Islamophobic Discourse Using Semi-Coded Terms and LLMs 24 Mar 2025 · 0 repositories · arXiv:2503.18273
-
Chirp Localization via Fine-Tuned Transformer Model: A Proof-of-Concept Study 24 Mar 2025 · 0 repositories · arXiv:2503.22713
-
Coeff-Tuning: A Graph Filter Subspace View for Tuning Attention-Based Large Models 24 Mar 2025 · 1 repository · arXiv:2503.18337Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Construction Identification and Disambiguation Using BERT: A Case Study of NPN 24 Mar 2025 · 0 repositories · arXiv:2503.18751
-
Context-Enhanced Memory-Refined Transformer for Online Action Detection 24 Mar 2025 · 1 repository · arXiv:2503.18359
-
Deep learning-based identification of precipitation clouds from all-sky camera data for observatory safety 24 Mar 2025 · 0 repositories · arXiv:2503.18670
-
Detecting Arbitrary Planted Subgraphs in Random Graphs 24 Mar 2025 · 0 repositories · arXiv:2503.19069
-
DisentTalk: Cross-lingual Talking Face Generation via Semantic Disentangled Diffusion Model 24 Mar 2025 · 0 repositories · arXiv:2503.19001
-
Distil-xLSTM: Learning Attention Mechanisms through Recurrent Structures 24 Mar 2025 · 0 repositories · arXiv:2503.18565
-
Dual-domain Multi-path Self-supervised Diffusion Model for Accelerated MRI Reconstruction 24 Mar 2025 · 0 repositories · arXiv:2503.18836
-
Efficient Self-Supervised Adaptation for Medical Image Analysis 24 Mar 2025 · 1 repository · arXiv:2503.18873
-
Enhancing Recommender Systems Using Textual Embeddings from Pre-trained Language Models 24 Mar 2025 · 0 repositories · arXiv:2504.08746
-
Equivariant Image Modeling 24 Mar 2025 · 1 repository · arXiv:2503.18948
-
Exploring State Space Model in Wavelet Domain: An Infrared and Visible Image Fusion Network via Wavelet Transform and State Space Model 24 Mar 2025 · 0 repositories · arXiv:2503.18378
-
Exploring the Integration of Key-Value Attention Into Pure and Hybrid Transformers for Semantic Segmentation 24 Mar 2025 · 0 repositories · arXiv:2503.18862
-
Exploring Training and Inference Scaling Laws in Generative Retrieval 24 Mar 2025 · 1 repository · arXiv:2503.18941
-
FFN Fusion: Rethinking Sequential Computation in Large Language Models 24 Mar 2025 · 0 repositories · arXiv:2503.18908
-
Frequency Dynamic Convolution for Dense Image Prediction 24 Mar 2025 · 1 repository · arXiv:2503.18783Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Global-Local Tree Search in VLMs for 3D Indoor Scene Generation 24 Mar 2025 · 1 repository · arXiv:2503.18476Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
How to Capture and Study Conversations Between Research Participants and ChatGPT: GPT for Researchers (g4r.org) 24 Mar 2025 · 0 repositories · arXiv:2503.18303
-
HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation 24 Mar 2025 · 0 repositories · arXiv:2503.18860
-
Image-to-Text for Medical Reports Using Adaptive Co-Attention and Triple-LSTM Module 24 Mar 2025 · 0 repositories · arXiv:2503.18297
-
Improving RAG for Personalization with Author Features and Contrastive Examples 24 Mar 2025 · 1 repository · arXiv:2504.08745
-
InPO: Inversion Preference Optimization with Reparametrized DDIM for Efficient Diffusion Model Alignment 24 Mar 2025 · 1 repository · arXiv:2503.18454
-
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models 24 Mar 2025 · 1 repository · arXiv:2503.18556
-
Language Model Uncertainty Quantification with Attention Chain 24 Mar 2025 · 1 repository · arXiv:2503.19168
-
LeanStereo: A Leaner Backbone based Stereo Network 24 Mar 2025 · 1 repository · arXiv:2503.18557
-
Learning a Class of Mixed Linear Regressions: Global Convergence under General Data Conditions 24 Mar 2025 · 0 repositories · arXiv:2503.18500
-
Learning to segment anatomy and lesions from disparately labeled sources in brain MRI 24 Mar 2025 · 0 repositories · arXiv:2503.18840
-
LGI-DETR: Local-Global Interaction for UAV Object Detection 24 Mar 2025 · 0 repositories · arXiv:2503.18785
-
LiDAR Remote Sensing Meets Weak Supervision: Concepts, Methods, and Perspectives 24 Mar 2025 · 0 repositories · arXiv:2503.18384
-
Linguistics-aware Masked Image Modeling for Self-supervised Scene Text Recognition 24 Mar 2025 · 1 repository · arXiv:2503.18746Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
LoTUS: Large-Scale Machine Unlearning with a Taste of Uncertainty 24 Mar 2025 · 1 repository · arXiv:2503.18314
-
Mechanistic Interpretability of Fine-Tuned Vision Transformers on Distorted Images: Decoding Attention Head Behavior for Transparent and Trustworthy AI 24 Mar 2025 · 0 repositories · arXiv:2503.18762
-
Mitigating Cache Noise in Test-Time Adaptation for Large Vision-Language Models 24 Mar 2025 · 0 repositories · arXiv:2503.18334
-
Oaken: Fast and Efficient LLM Serving with Online-Offline Hybrid KV Cache Quantization 24 Mar 2025 · 0 repositories · arXiv:2503.18599
-
On the Perception Bottleneck of VLMs for Chart Understanding 24 Mar 2025 · 1 repository · arXiv:2503.18435
-
Predicting the Road Ahead: A Knowledge Graph based Foundation Model for Scene Understanding in Autonomous Driving 24 Mar 2025 · 0 repositories · arXiv:2503.18730
-
Quantum Complex-Valued Self-Attention Model 24 Mar 2025 · 0 repositories · arXiv:2503.19002
-
REALM: A Dataset of Real-World LLM Use Cases 24 Mar 2025 · 0 repositories · arXiv:2503.18792
-
SPMTrack: Spatio-Temporal Parameter-Efficient Fine-Tuning with Mixture of Experts for Scalable Visual Tracking 24 Mar 2025 · 1 repository · arXiv:2503.18338Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SyncVP: Joint Diffusion for Synchronous Multi-Modal Video Prediction 24 Mar 2025 · 1 repository · arXiv:2503.18933
-
Synthetic Function Demonstrations Improve Generation in Low-Resource Programming Languages 24 Mar 2025 · 0 repositories · arXiv:2503.18760
-
Target-Aware Video Diffusion Models 24 Mar 2025 · 0 repositories · arXiv:2503.18950
-
The Fragility of the Reverse Facilitation Effect in the Stroop Task: A Dynamic Neurocognitive Model 24 Mar 2025 · 0 repositories · arXiv:2503.19128
-
TopV: Compatible Token Pruning with Inference Time Optimization for Fast and Low-Memory Multimodal Vision Language Model 24 Mar 2025 · 0 repositories · arXiv:2503.18278
-
U-REPA: Aligning Diffusion U-Nets to ViTs 24 Mar 2025 · 1 repository · arXiv:2503.18414
-
Video-XL-Pro: Reconstructive Token Compression for Extremely Long Video Understanding 24 Mar 2025 · 0 repositories · arXiv:2503.18478
-
xKV: Cross-Layer SVD for KV-Cache Compression 24 Mar 2025 · 1 repository · arXiv:2503.18893Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples)
-
Your ViT is Secretly an Image Segmentation Model 24 Mar 2025 · 1 repository · arXiv:2503.19108Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
ZeroLM: Data-Free Transformer Architecture Search for Language Models 24 Mar 2025 · 0 repositories · arXiv:2503.18646
-
Adaptive Rank Allocation: Speeding Up Modern Transformers with RaNA Adapters 23 Mar 2025 · 1 repository · arXiv:2503.18216Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Cat-AIR: Content and Task-Aware All-in-One Image Restoration 23 Mar 2025 · 0 repositories · arXiv:2503.17915
-
End-to-End Implicit Neural Representations for Classification 23 Mar 2025 · 1 repository · arXiv:2503.18123
-
ExpertRAG: Efficient RAG with Mixture of Experts -- Optimizing Context Retrieval for Adaptive LLM Responses 23 Mar 2025 · 0 repositories · arXiv:2504.08744
-
FS-SS: Few-Shot Learning for Fast and Accurate Spike Sorting of High-channel Count Probes 23 Mar 2025 · 0 repositories · arXiv:2503.18040
-
Investigating Recent Large Language Models for Vietnamese Machine Reading Comprehension 23 Mar 2025 · 0 repositories · arXiv:2503.18062
-
LakotaBERT: A Transformer-based Model for Low Resource Lakota Language 23 Mar 2025 · 0 repositories · arXiv:2503.18212
-
M3Net: Multimodal Multi-task Learning for 3D Detection, Segmentation, and Occupancy Prediction in Autonomous Driving 23 Mar 2025 · 1 repository · arXiv:2503.18100
-
Multi-Disease-Aware Training Strategy for Cardiac MR Image Segmentation 23 Mar 2025 · 0 repositories · arXiv:2503.17896
-
PanopticSplatting: End-to-End Panoptic Gaussian Splatting 23 Mar 2025 · 0 repositories · arXiv:2503.18073
-
PathoHR: Breast Cancer Survival Prediction on High-Resolution Pathological Images 23 Mar 2025 · 1 repository · arXiv:2503.17970
-
Payload-Aware Intrusion Detection with CMAE and Large Language Models 23 Mar 2025 · 0 repositories · arXiv:2503.20798
-
Real-World Remote Sensing Image Dehazing: Benchmark and Baseline 23 Mar 2025 · 1 repository · arXiv:2503.17966
-
Retrieval Augmented Generation and Understanding in Vision: A Survey and New Outlook 23 Mar 2025 · 1 repository · arXiv:2503.18016