Methods › General › Output Functions › Softmax › Papers, page 6
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 6 of 375: papers 501 to 600 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
FreeScene: Mixed Graph Diffusion for 3D Scene Synthesis from Free Prompts 3 Jun 2025 · 0 repositories · arXiv:2506.02781
-
FuXi-Ocean: A Global Ocean Forecasting System with Sub-Daily Resolution 3 Jun 2025 · 0 repositories · arXiv:2506.03210
-
HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference 3 Jun 2025 · 1 repository · arXiv:2506.02572
-
High Performance Space Debris Tracking in Complex Skylight Backgrounds with a Large-Scale Dataset 3 Jun 2025 · 0 repositories · arXiv:2506.02614
-
Learning Pyramid-structured Long-range Dependencies for 3D Human Pose Estimation 3 Jun 2025 · 1 repository · arXiv:2506.02853
-
Multi-modal brain MRI synthesis based on SwinUNETR 3 Jun 2025 · 0 repositories · arXiv:2506.02467
-
Overcoming Challenges of Partial Client Participation in Federated Learning : A Comprehensive Review 3 Jun 2025 · 0 repositories · arXiv:2506.02887
-
Rethinking Machine Unlearning in Image Generation Models 3 Jun 2025 · 1 repository · arXiv:2506.02761Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Rethinking the effects of data contamination in Code Intelligence 3 Jun 2025 · 0 repositories · arXiv:2506.02791
-
Revisiting End-to-End Learning with Slide-level Supervision in Computational Pathology 3 Jun 2025 · 2 repositories · arXiv:2506.02408Syntology official (archive's flag): 5 ran · 6 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Speaker Diarization with Overlapping Community Detection Using Graph Attention Networks and Label Propagation Algorithm 3 Jun 2025 · 1 repository · arXiv:2506.02610
-
Tactile MNIST: Benchmarking Active Tactile Perception 3 Jun 2025 · 0 repositories · arXiv:2506.06361
-
Talk2SAM: Text-Guided Semantic Enhancement for Complex-Shaped Object Segmentation 3 Jun 2025 · 0 repositories · arXiv:2506.05396
-
TL;DR: Too Long, Do Re-weighting for Efficient LLM Reasoning Compression 3 Jun 2025 · 1 repository · arXiv:2506.02678
-
Towards a Japanese Full-duplex Spoken Dialogue System 3 Jun 2025 · 0 repositories · arXiv:2506.02979
-
WeightLoRA: Keep Only Necessary Adapters 3 Jun 2025 · 0 repositories · arXiv:2506.02724
-
Are Mamba-based Audio Foundation Models the Best Fit for Non-Verbal Emotion Recognition? 2 Jun 2025 · 0 repositories · arXiv:2506.02258
-
Automatic Stage Lighting Control: Is it a Rule-Driven Process or Generative Task? 2 Jun 2025 · 1 repository · arXiv:2506.01482
-
Bayes optimal learning of attention-indexed models 2 Jun 2025 · 1 repository · arXiv:2506.01582
-
EfficientFER: EfficientNetv2 Based Deep Learning Approach for Facial Expression Recognition 2 Jun 2025 · 1 repository
-
Enhancing Speech Emotion Recognition with Graph-Based Multimodal Fusion and Prosodic Features for the Speech Emotion Recognition in Naturalistic Conditions Challenge at Interspeech 2025 2 Jun 2025 · 0 repositories · arXiv:2506.02088
-
GenDMR: A dynamic multimodal role-swapping network for identifying risk gene phenotypes 2 Jun 2025 · 0 repositories · arXiv:2506.01456
-
Hybrid AI for Responsive Multi-Turn Online Conversations with Novel Dynamic Routing and Feedback Adaptation 2 Jun 2025 · 0 repositories · arXiv:2506.02097
-
Life Sequence Transformer: Generative Modelling for Counterfactual Simulation 2 Jun 2025 · 0 repositories · arXiv:2506.01874
-
LLM in the Loop: Creating the PARADEHATE Dataset for Hate Speech Detoxification 2 Jun 2025 · 0 repositories · arXiv:2506.01484
-
LLMs as World Models: Data-Driven and Human-Centered Pre-Event Simulation for Disaster Impact Assessment 2 Jun 2025 · 0 repositories · arXiv:2506.06355
-
LongDWM: Cross-Granularity Distillation for Building a Long-Term Driving World Model 2 Jun 2025 · 0 repositories · arXiv:2506.01546
-
MoCA: Multi-modal Cross-masked Autoencoder for Digital Health Measurements 2 Jun 2025 · 0 repositories · arXiv:2506.02260
-
On-device Streaming Discrete Speech Units 2 Jun 2025 · 1 repository · arXiv:2506.01845
-
Optimal Coordination of Flexible DERs in Local Energy and Flexibility Markets to Ensure Social Equity 2 Jun 2025 · 0 repositories · arXiv:2506.02179
-
Retrieval-Augmented Generation of Ontologies from Relational Databases 2 Jun 2025 · 0 repositories · arXiv:2506.01232
-
SIL Allocation for Mitigation Safety Functions 2 Jun 2025 · 0 repositories · arXiv:2506.02309
-
Sparse Imagination for Efficient Visual World Model Planning 2 Jun 2025 · 0 repositories · arXiv:2506.01392
-
The Promise of Spiking Neural Networks for Ubiquitous Computing: A Survey and New Perspectives 2 Jun 2025 · 0 repositories · arXiv:2506.01737
-
A Graph-Retrieval-Augmented Generation Framework Enhances Decision-Making in the Circular Economy 1 Jun 2025 · 0 repositories · arXiv:2506.04252
-
Beyond Attention: Learning Spatio-Temporal Dynamics with Emergent Interpretable Topologies 1 Jun 2025 · 0 repositories · arXiv:2506.00770
-
CLAP-ART: Automated Audio Captioning with Semantic-rich Audio Representation Tokenizer 1 Jun 2025 · 0 repositories · arXiv:2506.00800
-
Explainable-AI powered stock price prediction using time series transformers: A Case Study on BIST100 1 Jun 2025 · 0 repositories · arXiv:2506.06345
-
GigaAM: Efficient Self-Supervised Learner for Speech Recognition 1 Jun 2025 · 1 repository · arXiv:2506.01192
-
How Neural Networks Organize Concepts: Introducing Concept Trajectory Analysis for Deep Learning Interpretability 1 Jun 2025 · 1 repository
-
Humanoid World Models: Open World Foundation Models for Humanoid Robotics 1 Jun 2025 · 0 repositories · arXiv:2506.01182
-
Leveraging AM and FM Rhythm Spectrograms for Dementia Classification and Assessment 1 Jun 2025 · 1 repository · arXiv:2506.00861
-
RARE: Retrieval-Aware Robustness Evaluation for Retrieval-Augmented Generation Systems 1 Jun 2025 · 1 repository · arXiv:2506.00789
-
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching 1 Jun 2025 · 0 repositories · arXiv:2506.01014
-
SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models 1 Jun 2025 · 1 repository · arXiv:2506.01062
-
TRUST -- Transformer-Driven U-Net for Sparse Target Recovery 1 Jun 2025 · 0 repositories · arXiv:2506.01112
-
A Foundation Model for Non-Destructive Defect Identification from Vibrational Spectra 31 May 2025 · 1 repository · arXiv:2506.00725
-
Assortment of Attention Heads: Accelerating Federated PEFT with Head Pruning and Strategic Client Selection 31 May 2025 · 0 repositories · arXiv:2506.00743
-
Attention-Aided MMSE for OFDM Channel Estimation: Learning Linear Filters with Attention 31 May 2025 · 0 repositories · arXiv:2506.00452
-
Blockchain-Enabled Privacy-Preserving Second-Order Federated Edge Learning in Personalized Healthcare 31 May 2025 · 0 repositories · arXiv:2506.00416
-
Channel-Imposed Fusion: A Simple yet Effective Method for Medical Time Series Classification 31 May 2025 · 0 repositories · arXiv:2506.00337
-
Evaluating Robot Policies in a World Model 31 May 2025 · 0 repositories · arXiv:2506.00613
-
FinBERT2: A Specialized Bidirectional Encoder for Bridging the Gap in Finance-Specific Deployment of Large Language Models 31 May 2025 · 0 repositories · arXiv:2506.06335
-
Machine vs Machine: Using AI to Tackle Generative AI Threats in Assessment 31 May 2025 · 0 repositories · arXiv:2506.02046
-
Multi-Objective Neural Network Assisted Design Optimization of Soft Fin-Ray Grippers for Enhanced Grasping Performance 31 May 2025 · 0 repositories · arXiv:2506.00494
-
Position: Olfaction Standardization is Essential for the Advancement of Embodied Artificial Intelligence 31 May 2025 · 0 repositories · arXiv:2506.00398
-
Power-of-Two (PoT) Weights in Large Language Models (LLMs) 31 May 2025 · 0 repositories · arXiv:2506.00315
-
Towards Graph-Based Privacy-Preserving Federated Learning: ModelNet -- A ResNet-based Model Classification Dataset 31 May 2025 · 0 repositories · arXiv:2506.00476
-
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations 31 May 2025 · 1 repository · arXiv:2506.00748
-
Using Diffusion Ensembles to Estimate Uncertainty for End-to-End Autonomous Driving 31 May 2025 · 0 repositories · arXiv:2506.00560
-
3D Gaussian Splat Vulnerabilities 30 May 2025 · 1 repository · arXiv:2506.00280
-
Adaptive LoRA Merge with Parameter Pruning for Low-Resource Generation 30 May 2025 · 1 repository · arXiv:2505.24174
-
Adversarial Threat Vectors and Risk Mitigation for Retrieval-Augmented Generation Systems 30 May 2025 · 0 repositories · arXiv:2506.00281
-
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks 30 May 2025 · 1 repository · arXiv:2505.24876
-
Bayesian Data Sketching for Varying Coefficient Regression Models 30 May 2025 · 0 repositories · arXiv:2506.00270
-
Cloud Optical Thickness Retrievals Using Angle Invariant Attention Based Deep Learning Models 30 May 2025 · 0 repositories · arXiv:2505.24638
-
ClueAnchor: Clue-Anchored Knowledge Reasoning Exploration and Optimization for Retrieval-Augmented Generation 30 May 2025 · 1 repository · arXiv:2505.24388
-
Cross-Attention Speculative Decoding 30 May 2025 · 0 repositories · arXiv:2505.24544
-
D2AF: A Dual-Driven Annotation and Filtering Framework for Visual Grounding 30 May 2025 · 0 repositories · arXiv:2505.24372
-
Decoding Knowledge Attribution in Mixture-of-Experts: A Framework of Basic-Refinement Collaboration and Efficiency Analysis 30 May 2025 · 0 repositories · arXiv:2505.24593
-
Deformable Attention Mechanisms Applied to Object Detection, case of Remote Sensing 30 May 2025 · 0 repositories · arXiv:2505.24489
-
E^2GraphRAG: Streamlining Graph-based RAG for High Efficiency and Effectiveness 30 May 2025 · 0 repositories · arXiv:2505.24226
-
Efficient Text Encoders for Labor Market Analysis 30 May 2025 · 0 repositories · arXiv:2505.24640
-
Explainable Depression Detection using Masked Hard Instance Mining 30 May 2025 · 0 repositories · arXiv:2505.24609
-
From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models 30 May 2025 · 0 repositories · arXiv:2505.24232
-
HELM: Hyperbolic Large Language Models via Mixture-of-Curvature Experts 30 May 2025 · 1 repository · arXiv:2505.24722Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Interactive Video Generation via Domain Adaptation 30 May 2025 · 0 repositories · arXiv:2505.24253
-
Interpretable phenotyping of Heart Failure patients with Dutch discharge letters 30 May 2025 · 0 repositories · arXiv:2505.24619
-
Interpreting Large Text-to-Image Diffusion Models with Dictionary Learning 30 May 2025 · 1 repository · arXiv:2505.24360
-
Large Language Models are Locally Linear Mappings 30 May 2025 · 1 repository · arXiv:2505.24293
-
Leveraging Intermediate Features of Vision Transformer for Face Anti-Spoofing 30 May 2025 · 0 repositories · arXiv:2505.24402
-
Lightweight Relational Embedding in Task-Interpolated Few-Shot Networks for Enhanced Gastrointestinal Disease Classification 30 May 2025 · 0 repositories · arXiv:2505.24792
-
LPASS: Linear Probes as Stepping Stones for vulnerability detection using compressed LLMs 30 May 2025 · 0 repositories · arXiv:2505.24451
-
Mamba Knockout for Unraveling Factual Information Flow 30 May 2025 · 1 repository · arXiv:2505.24244Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer 30 May 2025 · 1 repository · arXiv:2505.24378
-
Mixture-of-Experts for Personalized and Semantic-Aware Next Location Prediction 30 May 2025 · 0 repositories · arXiv:2505.24597
-
Model-Guided Network with Cluster-Based Operators for Spatio-Spectral Super-Resolution 30 May 2025 · 1 repository · arXiv:2505.24605
-
MOFGPT: Generative Design of Metal-Organic Frameworks using Language Models 30 May 2025 · 1 repository · arXiv:2506.00198
-
Optimal Weighted Convolution for Classification and Denosing 30 May 2025 · 2 repositories · arXiv:2505.24558
-
PCIE_Pose Solution for EgoExo4D Pose and Proficiency Estimation Challenge 30 May 2025 · 0 repositories · arXiv:2505.24411
-
PersianMedQA: Language-Centric Evaluation of LLMs in the Persian Medical Domain 30 May 2025 · 0 repositories · arXiv:2506.00250
-
RealDrive: Retrieval-Augmented Driving with Diffusion Models 30 May 2025 · 0 repositories · arXiv:2505.24808
-
ReCalKV: Low-Rank KV Cache Compression via Head Reordering and Offline Calibration 30 May 2025 · 1 repository · arXiv:2505.24357
-
S3CE-Net: Spike-guided Spatiotemporal Semantic Coupling and Expansion Network for Long Sequence Event Re-Identification 30 May 2025 · 1 repository · arXiv:2505.24401
-
SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling 30 May 2025 · 1 repository · arXiv:2505.24179
-
SPPSFormer: High-quality Superpoint-based Transformer for Roof Plane Instance Segmentation from Point Clouds 30 May 2025 · 0 repositories · arXiv:2505.24475
-
STAR-Net: An Interpretable Model-Aided Network for Remote Sensing Image Denoising 30 May 2025 · 1 repository · arXiv:2505.24327
-
The Hype Index: an NLP-driven Measure of Market News Attention 30 May 2025 · 0 repositories · arXiv:2506.06329
-
Transformers Are Universally Consistent 30 May 2025 · 0 repositories · arXiv:2505.24531
-
Two failure modes of deep transformers and how to avoid them: a unified theory of signal propagation at initialisation 30 May 2025 · 0 repositories · arXiv:2505.24333