Methods › General › Output Functions › Softmax › Papers, page 99
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 99 of 375: papers 9,801 to 9,900 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Simplified priors for Object-Centric Learning 1 Oct 2024 · 0 repositories · arXiv:2410.00728
-
Softmax is not Enough (for Sharp Size Generalisation) 1 Oct 2024 · 0 repositories · arXiv:2410.01104
-
Sparse Attention Decomposition Applied to Circuit Tracing 1 Oct 2024 · 1 repository · arXiv:2410.00340
-
Spatial Action Unit Cues for Interpretable Deep Facial Expression Recognition 1 Oct 2024 · 1 repository · arXiv:2410.01848Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
STGformer: Efficient Spatiotemporal Graph Transformer for Traffic Forecasting 1 Oct 2024 · 1 repository · arXiv:2410.00385
-
Graph-Based Representation Learning of Neuronal Dynamics and Behavior 1 Oct 2024 · 1 repository · arXiv:2410.00665
-
TFCT-I2P: Three stream fusion network with color aware transformer for image-to-point cloud registration 1 Oct 2024 · 1 repository · arXiv:2410.00360
-
TPN: Transferable Proto-Learning Network towards Few-shot Document-Level Relation Extraction 1 Oct 2024 · 1 repository · arXiv:2410.00412
-
TransResNet: Integrating the Strengths of ViTs and CNNs for High Resolution Medical Image Segmentation via Feature Grafting 1 Oct 2024 · 1 repository · arXiv:2410.00986
-
Y-CA-Net: A Convolutional Attention Based Network for Volumetric Medical Image Segmentation 1 Oct 2024 · 0 repositories · arXiv:2410.01003
-
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions 30 Sep 2024 · 0 repositories · arXiv:2409.20303
-
A Methodology for Explainable Large Language Models with Integrated Gradients and Linguistic Analysis in Text Classification 30 Sep 2024 · 0 repositories · arXiv:2410.00250
-
ACE: All-round Creator and Editor Following Instructions via Diffusion Transformer 30 Sep 2024 · 0 repositories · arXiv:2410.00086
-
Adapting LLMs for the Medical Domain in Portuguese: A Study on Fine-Tuning and Model Evaluation 30 Sep 2024 · 0 repositories · arXiv:2410.00163
-
ASQuery: A Query-based Model for Action Segmentation 30 Sep 2024 · 1 repository
-
BSharedRAG: Backbone Shared Retrieval-Augmented Generation for the E-commerce Domain 30 Sep 2024 · 0 repositories · arXiv:2409.20075
-
CBAM-SwinT-BL: Small Rail Surface Defect Detection Method Based on Swin Transformer with Block Level CBAM Enhancement 30 Sep 2024 · 0 repositories · arXiv:2409.20113
-
Characterizing and Efficiently Accelerating Multimodal Generation Model Inference 30 Sep 2024 · 0 repositories · arXiv:2410.00215
-
CliMB: An AI-enabled Partner for Clinical Predictive Modeling 30 Sep 2024 · 1 repository · arXiv:2410.03736Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
CLR-GAN: Improving GANs Stability and Quality via Consistent Latent Representation and Reconstruction 30 Sep 2024 · 1 repository
-
Contrastive Token Learning with Similarity Decay for Repetition Suppression in Machine Translation 30 Sep 2024 · 0 repositories · arXiv:2409.19877
-
Depression detection in social media posts using transformer-based models and auxiliary features 30 Sep 2024 · 0 repositories · arXiv:2409.20048
-
Enhancing Romanian Offensive Language Detection through Knowledge Distillation, Multi-Task Learning, and Data Augmentation 30 Sep 2024 · 0 repositories · arXiv:2409.20498
-
Evaluating the fairness of task-adaptive pretraining on unlabeled test data before few-shot text classification 30 Sep 2024 · 1 repository · arXiv:2410.00179
-
FreeMask: Rethinking the Importance of Attention Masks for Zero-Shot Video Editing 30 Sep 2024 · 0 repositories · arXiv:2409.20500
-
MoCoLSK: Modality Conditioned High-Resolution Downscaling for Land Surface Temperature 30 Sep 2024 · 1 repository · arXiv:2409.19835
-
GTransPDM: A Graph-embedded Transformer with Positional Decoupling for Pedestrian Crossing Intention Prediction 30 Sep 2024 · 0 repositories · arXiv:2409.20223
-
HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding 30 Sep 2024 · 1 repository · arXiv:2409.20429
-
ImmersePro: End-to-End Stereo Video Synthesis Via Implicit Disparity Learning 30 Sep 2024 · 1 repository · arXiv:2410.00262
-
Ingest-And-Ground: Dispelling Hallucinations from Continually-Pretrained LLMs with RAG 30 Sep 2024 · 0 repositories · arXiv:2410.02825
-
Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis 30 Sep 2024 · 0 repositories · arXiv:2409.20059
-
KV-Compress: Paged KV-Cache Compression with Variable Compression Rates per Attention Head 30 Sep 2024 · 1 repository · arXiv:2410.00161Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Learning Multimodal Latent Generative Models with Energy-Based Prior 30 Sep 2024 · 1 repository · arXiv:2409.19862
-
Maia-2: A Unified Model for Human-AI Alignment in Chess 30 Sep 2024 · 2 repositories · arXiv:2409.20553Syntology official (archive's flag): 4 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified (of 18 harvested samples) · 11 pointer-only (licence)
-
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation 30 Sep 2024 · 0 repositories · arXiv:2409.19937
-
Mechanism Design with Endogenous Perception 30 Sep 2024 · 0 repositories · arXiv:2409.19853
-
Modelando procesos cognitivos de la lectura natural con GPT-2 30 Sep 2024 · 0 repositories · arXiv:2409.20174
-
Numerically Robust Fixed-Point Smoothing Without State Augmentation 30 Sep 2024 · 1 repository · arXiv:2409.20004
-
Exploring Social Media Image Categorization Using Large Models with Different Adaptation Methods: A Case Study on Cultural Nature's Contributions to People 30 Sep 2024 · 0 repositories · arXiv:2410.00275
-
On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability 30 Sep 2024 · 2 repositories · arXiv:2409.19924
-
QAEncoder: Towards Aligned Representation Learning in Question Answering System 30 Sep 2024 · 1 repository · arXiv:2409.20434
-
Social Conjuring: Multi-User Runtime Collaboration with AI in Building Virtual 3D Worlds 30 Sep 2024 · 0 repositories · arXiv:2410.00274
-
SWIM: Short-Window CNN Integrated with Mamba for EEG-Based Auditory Spatial Attention Decoding 30 Sep 2024 · 1 repository · arXiv:2409.19884
-
Systemic Risk Asymptotics in a Renewal Model with Multiple Business Lines and Heterogeneous Claims 30 Sep 2024 · 0 repositories · arXiv:2410.00158
-
T-KAER: Transparency-enhanced Knowledge-Augmented Entity Resolution Framework 30 Sep 2024 · 1 repository · arXiv:2410.00218
-
The age of spiritual machines: Language quietus induces synthetic altered states of consciousness in artificial intelligence 30 Sep 2024 · 0 repositories · arXiv:2410.00257
-
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels 30 Sep 2024 · 0 repositories · arXiv:2409.19846
-
Whole-Graph Representation Learning For the Classification of Signed Networks 30 Sep 2024 · 1 repository · arXiv:2409.20073
-
2D-TPE: Two-Dimensional Positional Encoding Enhances Table Understanding for Large Language Models 29 Sep 2024 · 1 repository · arXiv:2409.19700
-
A multimodal LLM for the non-invasive decoding of spoken text from brain recordings 29 Sep 2024 · 0 repositories · arXiv:2409.19710
-
Abstractive Summarization of Low resourced Nepali language using Multilingual Transformers 29 Sep 2024 · 0 repositories · arXiv:2409.19566
-
Adversarial Examples for DNA Classification 29 Sep 2024 · 0 repositories · arXiv:2409.19788
-
Automated Disease Diagnosis in Pumpkin Plants Using Advanced CNN Models 29 Sep 2024 · 0 repositories · arXiv:2410.00062
-
Black-Box Segmentation of Electronic Medical Records 29 Sep 2024 · 0 repositories · arXiv:2409.19796
-
Can Models Learn Skill Composition from Examples? 29 Sep 2024 · 0 repositories · arXiv:2409.19808
-
Causal Deciphering and Inpainting in Spatio-Temporal Dynamics via Diffusion Model 29 Sep 2024 · 0 repositories · arXiv:2409.19608
-
CRScore: Grounding Automated Evaluation of Code Review Comments in Code Claims and Smells 29 Sep 2024 · 0 repositories · arXiv:2409.19801
-
Differentially Private Bilevel Optimization 29 Sep 2024 · 0 repositories · arXiv:2409.19800
-
DIIT: A Domain-Invariant Information Transfer Method for Industrial Cross-Domain Recommendation 29 Sep 2024 · 0 repositories · arXiv:2410.10835
-
Discerning the Chaos: Detecting Adversarial Perturbations while Disentangling Intentional from Unintentional Noises 29 Sep 2024 · 0 repositories · arXiv:2409.19619
-
Does RAG Introduce Unfairness in LLMs? Evaluating Fairness in Retrieval-Augmented Generation Systems 29 Sep 2024 · 1 repository · arXiv:2409.19804Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Dual-Attention Frequency Fusion at Multi-Scale for Joint Segmentation and Deformable Medical Image Registration 29 Sep 2024 · 0 repositories · arXiv:2409.19658
-
Federated Learning from Vision-Language Foundation Models: Theoretical Analysis and Method 29 Sep 2024 · 1 repository · arXiv:2409.19610Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Flipped Classroom: Aligning Teacher Attention with Student in Generalized Category Discovery 29 Sep 2024 · 0 repositories · arXiv:2409.19659
-
Fully Aligned Network for Referring Image Segmentation 29 Sep 2024 · 0 repositories · arXiv:2409.19569
-
GenTel-Safe: A Unified Benchmark and Shielding Framework for Defending Against Prompt Injection Attacks 29 Sep 2024 · 0 repositories · arXiv:2409.19521
-
DATransNet: Dynamic Attention Transformer Network for Infrared Small Target Detection 29 Sep 2024 · 1 repository · arXiv:2409.19599
-
Hybrid Mamba for Few-Shot Segmentation 29 Sep 2024 · 1 repository · arXiv:2409.19613Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 9 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
IDEA: An Inverse Domain Expert Adaptation Based Active DNN IP Protection Method 29 Sep 2024 · 0 repositories · arXiv:2410.00059
-
Identifying Knowledge Editing Types in Large Language Models 29 Sep 2024 · 1 repository · arXiv:2409.19663Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
InfantCryNet: A Data-driven Framework for Intelligent Analysis of Infant Cries 29 Sep 2024 · 0 repositories · arXiv:2409.19689
-
Instruction Embedding: Latent Representations of Instructions Towards Task Identification 29 Sep 2024 · 0 repositories · arXiv:2409.19680
-
Investigating the Effect of Network Pruning on Performance and Interpretability 29 Sep 2024 · 1 repository · arXiv:2409.19727
-
Learning Attentional Mixture of LoRAs for Language Model Continual Learning 29 Sep 2024 · 0 repositories · arXiv:2409.19611
-
MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models 29 Sep 2024 · 0 repositories · arXiv:2409.19492
-
Natural Language Generation for Visualizations: State of the Art, Challenges and Future Directions 29 Sep 2024 · 0 repositories · arXiv:2409.19747
-
Neural-Polyptych: Content Controllable Painting Recreation for Diverse Genres 29 Sep 2024 · 0 repositories · arXiv:2409.19690
-
PEAR: Position-Embedding-Agnostic Attention Re-weighting Enhances Retrieval-Augmented Generation with Zero Inference Overhead 29 Sep 2024 · 0 repositories · arXiv:2409.19745
-
See then Tell: Enhancing Key Information Extraction with Vision Grounding 29 Sep 2024 · 0 repositories · arXiv:2409.19573
-
Spiking Transformer with Spatial-Temporal Attention 29 Sep 2024 · 1 repository · arXiv:2409.19764
-
Analog In-Memory Computing Attention Mechanism for Fast and Energy-Efficient Large Language Models 28 Sep 2024 · 1 repository · arXiv:2409.19315
-
Beyond Euclidean: Dual-Space Representation Learning for Weakly Supervised Video Violence Detection 28 Sep 2024 · 0 repositories · arXiv:2409.19252
-
Decoding Android Malware with a Fraction of Features: An Attention-Enhanced MLP-SVM Approach 28 Sep 2024 · 0 repositories · arXiv:2409.19234
-
DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning 28 Sep 2024 · 0 repositories · arXiv:2409.19255
-
Efficient Federated Intrusion Detection in 5G ecosystem using optimized BERT-based model 28 Sep 2024 · 1 repository · arXiv:2409.19390
-
INSIGHTBUDDY-AI: Medication Extraction and Entity Linking using Large Language Models and Ensemble Learning 28 Sep 2024 · 2 repositories · arXiv:2409.19467
-
Membership Privacy Evaluation in Deep Spiking Neural Networks 28 Sep 2024 · 0 repositories · arXiv:2409.19413
-
Multi-Atlas Brain Network Classification through Consistency Distillation and Complementary Information Fusion 28 Sep 2024 · 0 repositories · arXiv:2410.08228
-
NeuralQP: A General Hypergraph-based Optimization Framework for Large-scale QCQPs 28 Sep 2024 · 0 repositories · arXiv:2410.03720
-
RMLR: Extending Multinomial Logistic Regression into General Geometries 28 Sep 2024 · 2 repositories · arXiv:2409.19433Syntology official (archive's flag): 21 ran · 23 ran (of which 8 constructed an object rather than computing a result; 22 with no instrument failure: 0 honoured, 8 violated, 14 with no contract checked; 1 where Syntology's instrument failed) · 14 unverified (of 37 harvested samples) · 37 pointer-only (licence)
-
Unveil Benign Overfitting for Transformer in Vision: Training Dynamics, Convergence, and Generalization 28 Sep 2024 · 0 repositories · arXiv:2409.19345
-
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation 28 Sep 2024 · 1 repository · arXiv:2409.19342Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
AIPatient: Simulating Patients with EHRs and LLM Powered Agentic Workflow 27 Sep 2024 · 0 repositories · arXiv:2409.18924
-
Charting the Future: Using Chart Question-Answering for Scalable Evaluation of LLM-Driven Data Visualizations 27 Sep 2024 · 0 repositories · arXiv:2409.18764
-
Cottention: Linear Transformers With Cosine Attention 27 Sep 2024 · 1 repository · arXiv:2409.18747Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Cross-video Identity Correlating for Person Re-identification Pre-training 27 Sep 2024 · 1 repository · arXiv:2409.18569Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Deep Hybrid Architecture for Very Low-Resolution Image Classification Using Capsule Attention 27 Sep 2024 · 1 repository
-
Dynamic Competition for Attention 27 Sep 2024 · 0 repositories · arXiv:2409.18595
-
Experimental Evaluation of Machine Learning Models for Goal-oriented Customer Service Chatbot with Pipeline Architecture 27 Sep 2024 · 0 repositories · arXiv:2409.18568
-
Exploring Token Pruning in Vision State Space Models 27 Sep 2024 · 0 repositories · arXiv:2409.18962