Methods › General › Output Functions › Softmax › Papers, page 14
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 14 of 375: papers 1,301 to 1,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Learning High-Order Relationships with Hypergraph Attention-based Spatio-Temporal Aggregation for Brain Disease Analysis 17 May 2025 · 0 repositories · arXiv:2505.12068
-
Learning to Dissipate Energy in Oscillatory State-Space Models 17 May 2025 · 1 repository · arXiv:2505.12171
-
Let's have a chat with the EU AI Act 17 May 2025 · 0 repositories · arXiv:2505.11946
-
Lightweight Spatio-Temporal Attention Network with Graph Embedding and Rotational Position Encoding for Traffic Forecasting 17 May 2025 · 0 repositories · arXiv:2505.12136
-
LoRASuite: Efficient LoRA Adaptation Across Large Language Model Upgrades 17 May 2025 · 0 repositories · arXiv:2505.13515
-
MedVKAN: Efficient Feature Extraction with Mamba and KAN for Medical Image Segmentation 17 May 2025 · 1 repository · arXiv:2505.11797
-
Mixture of Decoding: An Attention-Inspired Adaptive Decoding Strategy to Mitigate Hallucinations in Large Vision-Language Models 17 May 2025 · 1 repository · arXiv:2505.17061
-
Neuro-Symbolic Query Compiler 17 May 2025 · 1 repository · arXiv:2505.11932
-
SpatialCrafter: Unleashing the Imagination of Video Diffusion Models for Scene Reconstruction from Limited Observations 17 May 2025 · 0 repositories · arXiv:2505.11992
-
Telco-oRAG: Optimizing Retrieval-augmented Generation for Telecom Queries via Hybrid Retrieval and Neural Routing 17 May 2025 · 0 repositories · arXiv:2505.11856
-
The Logical Expressiveness of Temporal GNNs via Two-Dimensional Product Logics 17 May 2025 · 0 repositories · arXiv:2505.11930
-
Towards Comprehensive Argument Analysis in Education: Dataset, Tasks, and Method 17 May 2025 · 0 repositories · arXiv:2505.12028
-
Unveiling Knowledge Utilization Mechanisms in LLM-based Retrieval-Augmented Generation 17 May 2025 · 0 repositories · arXiv:2505.11995
-
VeriReason: Reinforcement Learning with Testbench Feedback for Reasoning-Enhanced Verilog Generation 17 May 2025 · 1 repository · arXiv:2505.11849
-
Why Not Act on What You Know? Unleashing Safety Potential of LLMs via Self-Aware Guard Enhancement 17 May 2025 · 1 repository · arXiv:2505.12060Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Attention-Based Reward Shaping for Sparse and Delayed Rewards 16 May 2025 · 1 repository · arXiv:2505.10802
-
A High-Performance Thermal Infrared Object Detection Framework with Centralized Regulation 16 May 2025 · 0 repositories · arXiv:2505.10825
-
RefPose: Leveraging Reference Geometric Correspondences for Accurate 6D Pose Estimation of Unseen Objects 16 May 2025 · 0 repositories · arXiv:2505.10841
-
Have Multimodal Large Language Models (MLLMs) Really Learned to Tell the Time on Analog Clocks? 16 May 2025 · 0 repositories · arXiv:2505.10862
-
CTP: A hybrid CNN-Transformer-PINN model for ocean front forecasting 16 May 2025 · 0 repositories · arXiv:2505.10894
-
Automated Identification of Logical Errors in Programs: Advancing Scalable Analysis of Student Misconceptions 16 May 2025 · 0 repositories · arXiv:2505.10913
-
Connecting the Dots: A Chain-of-Collaboration Prompting Framework for LLM Agents 16 May 2025 · 0 repositories · arXiv:2505.10936
-
SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache 16 May 2025 · 0 repositories · arXiv:2505.10951
-
Relational Graph Transformer 16 May 2025 · 1 repository · arXiv:2505.10960Syntology official (archive's flag): 2 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
RAGSynth: Synthetic Data for Robust and Faithful RAG Component Optimization 16 May 2025 · 1 repository · arXiv:2505.10989
-
Illusion or Algorithm? Investigating Memorization, Emergence, and Symbolic Processing in In-Context Learning 16 May 2025 · 1 repository · arXiv:2505.11004
-
Rethinking the Mean Teacher Strategy from the Perspective of Self-paced Learning 16 May 2025 · 0 repositories · arXiv:2505.11018
-
Deep Latent Variable Model based Vertical Federated Learning with Flexible Alignment and Labeling Scenarios 16 May 2025 · 0 repositories · arXiv:2505.11035
-
Efficient Attention via Pre-Scoring: Prioritizing Informative Keys in Transformers 16 May 2025 · 1 repository · arXiv:2505.11040
-
ShiQ: Bringing back Bellman to LLMs 16 May 2025 · 0 repositories · arXiv:2505.11081
-
Fault Diagnosis across Heterogeneous Domains via Self-Adaptive Temporal-Spatial Attention and Sample Generation 16 May 2025 · 1 repository · arXiv:2505.11083
-
Redundancy-Aware Pretraining of Vision-Language Foundation Models in Remote Sensing 16 May 2025 · 0 repositories · arXiv:2505.11121
-
GraphOracle: A Foundation Model for Knowledge Graph Reasoning 16 May 2025 · 0 repositories · arXiv:2505.11125
-
STEP: A Unified Spiking Transformer Evaluation Platform for Fair and Reproducible Benchmarking 16 May 2025 · 1 repository · arXiv:2505.11151
-
Attention on the Sphere 16 May 2025 · 1 repository · arXiv:2505.11157
-
Maximizing Asynchronicity in Event-based Neural Networks 16 May 2025 · 0 repositories · arXiv:2505.11165
-
CheX-DS: Improving Chest X-ray Image Classification with Ensemble Learning Based on DenseNet and Swin Transformer 16 May 2025 · 0 repositories · arXiv:2505.11168
-
mmRAG: A Modular Benchmark for Retrieval-Augmented Generation over Text, Tables, and Knowledge Graphs 16 May 2025 · 1 repository · arXiv:2505.11180
-
DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling 16 May 2025 · 1 repository · arXiv:2505.11196Syntology official (archive's flag): 4 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
NoPE: The Counting Power of Transformers with No Positional Encodings 16 May 2025 · 0 repositories · arXiv:2505.11199
-
GLOVA: Global and Local Variation-Aware Analog Circuit Design with Risk-Sensitive Reinforcement Learning 16 May 2025 · 0 repositories · arXiv:2505.11208
-
Delta Attention: Fast and Accurate Sparse Attention Inference by Delta Correction 16 May 2025 · 0 repositories · arXiv:2505.11254
-
Multiclass threshold-based classification 16 May 2025 · 0 repositories · arXiv:2505.11276
-
LegoSLM: Connecting LLM with Speech Encoder using CTC Posteriors 16 May 2025 · 0 repositories · arXiv:2505.11352
-
Fractal Graph Contrastive Learning 16 May 2025 · 0 repositories · arXiv:2505.11356
-
LGBQPC: Local Granular-Ball Quality Peaks Clustering 16 May 2025 · 0 repositories · arXiv:2505.11359
-
Towards Cultural Bridge by Bahnaric-Vietnamese Translation Using Transfer Learning of Sequence-To-Sequence Pre-training Language Model 16 May 2025 · 0 repositories · arXiv:2505.11421
-
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs 16 May 2025 · 0 repositories · arXiv:2505.11423
-
MegaScale-MoE: Large-Scale Communication-Efficient Training of Mixture-of-Experts Models in Production 16 May 2025 · 0 repositories · arXiv:2505.11432
-
SurgPose: Generalisable Surgical Instrument Pose Estimation using Zero-Shot Learning and Stereo Vision 16 May 2025 · 0 repositories · arXiv:2505.11439
-
A Classical View on Benign Overfitting: The Role of Sample Size 16 May 2025 · 0 repositories · arXiv:2505.11621
-
ACSE-Eval: Can LLMs threat model real-world cloud infrastructure? 16 May 2025 · 1 repository · arXiv:2505.11565
-
AI-Driven Digital Transformation and Firm Performance in Chinese Industrial Enterprises: Mediating Role of Green Digital Innovation and Moderating Effects of Human-AI Collaboration 16 May 2025 · 0 repositories · arXiv:2505.11558
-
Beyond KL-divergence: Risk Aware Control Through Cross Entropy and Adversarial Entropy Regularization 16 May 2025 · 0 repositories · arXiv:2505.11068
-
Can an Easy-to-Hard Curriculum Make Reasoning Emerge in Small Language Models? Evidence from a Four-Stage Curriculum on GPT-2 16 May 2025 · 0 repositories · arXiv:2505.11643
-
EcoSafeRAG: Efficient Security through Context Analysis in Retrieval-Augmented Generation 16 May 2025 · 0 repositories · arXiv:2505.13506
-
Enhancing Mathematics Learning for Hard-of-Hearing Students Through Real-Time Palestinian Sign Language Recognition: A New Dataset 16 May 2025 · 0 repositories · arXiv:2505.17055
-
Finetune-RAG: Fine-Tuning Language Models to Resist Hallucination in Retrieval-Augmented Generation 16 May 2025 · 1 repository · arXiv:2505.10792
-
Flash Invariant Point Attention 16 May 2025 · 1 repository · arXiv:2505.11580
-
Heart2Mind: Human-Centered Contestable Psychiatric Disorder Diagnosis System using Wearable ECG Monitors 16 May 2025 · 1 repository · arXiv:2505.11612
-
Let the Trial Begin: A Mock-Court Approach to Vulnerability Detection using LLM-Based Agents 16 May 2025 · 0 repositories · arXiv:2505.10961
-
Masking in Multi-hop QA: An Analysis of How Language Models Perform with Context Permutation 16 May 2025 · 1 repository · arXiv:2505.11754
-
Mathematical models for the EP2 and EP4 signaling pathways and their crosstalk 16 May 2025 · 0 repositories · arXiv:2505.11712
-
Optimal Control for Transformer Architectures: Enhancing Generalization, Robustness and Efficiency 16 May 2025 · 0 repositories · arXiv:2505.13499
-
Phi: Leveraging Pattern-based Hierarchical Sparsity for High-Efficiency Spiking Neural Networks 16 May 2025 · 0 repositories · arXiv:2505.10909
-
SageAttention3: Microscaling FP4 Attention for Inference and An Exploration of 8-Bit Training 16 May 2025 · 1 repository · arXiv:2505.11594Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
THELMA: Task Based Holistic Evaluation of Large Language Model Applications-RAG Question Answering 16 May 2025 · 0 repositories · arXiv:2505.11626
-
Transforming Decoder-Only Transformers for Accurate WiFi-Telemetry Based Indoor Localization 16 May 2025 · 0 repositories · arXiv:2505.15835
-
Vaiage: A Multi-Agent Solution to Personalized Travel Planning 16 May 2025 · 0 repositories · arXiv:2505.10922
-
ZeroTuning: Unlocking the Initial Token's Power to Enhance Large Language Models Without Training 16 May 2025 · 0 repositories · arXiv:2505.11739
-
ARFC-WAHNet: Adaptive Receptive Field Convolution and Wavelet-Attentive Hierarchical Network for Infrared Small Target Detection 15 May 2025 · 1 repository · arXiv:2505.10595
-
SRMamba: Mamba for Super-Resolution of LiDAR Point Clouds 15 May 2025 · 0 repositories · arXiv:2505.10601
-
Continuity and Isolation Lead to Doubts or Dilemmas in Large Language Models 15 May 2025 · 0 repositories · arXiv:2505.10606
-
MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly 15 May 2025 · 1 repository · arXiv:2505.10610Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Artificial Intelligence Bias on English Language Learners in Automatic Scoring 15 May 2025 · 0 repositories · arXiv:2505.10643
-
Advancing Multiple Instance Learning with Continual Learning for Whole Slide Imaging 15 May 2025 · 0 repositories · arXiv:2505.10649
-
Seasonal Forecasting of Pan-Arctic Sea Ice with State Space Model 15 May 2025 · 1 repository · arXiv:2505.10665
-
GaussianFormer3D: Multi-Modal Gaussian-based Semantic Occupancy Prediction with 3D Deformable Attention 15 May 2025 · 0 repositories · arXiv:2505.10685
-
A Modular Approach for Clinical SLMs Driven by Synthetic Data with Pre-Instruction Tuning, Model Merging, and Clinical-Tasks Alignment 15 May 2025 · 0 repositories · arXiv:2505.10717
-
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation 15 May 2025 · 0 repositories · arXiv:2505.10743
-
Advances in Radiance Field for Dynamic Scene: From Neural Field to Gaussian Field 15 May 2025 · 1 repository · arXiv:2505.10049
-
AI Agents vs. Agentic AI: A Conceptual Taxonomy, Applications and Challenges 15 May 2025 · 0 repositories · arXiv:2505.10468
-
All You Need Is Synthetic Task Augmentation 15 May 2025 · 0 repositories · arXiv:2505.10120
-
Are Sparse Autoencoders Useful for Java Function Bug Detection? 15 May 2025 · 1 repository · arXiv:2505.10375
-
Assessing Collective Reasoning in Multi-Agent LLMs via Hidden Profile Tasks 15 May 2025 · 0 repositories · arXiv:2505.11556
-
Automating Security Audit Using Large Language Model based Agent: An Exploration Experiment 15 May 2025 · 0 repositories · arXiv:2505.10732
-
Avocado Price Prediction Using a Hybrid Deep Learning Model: TCN-MLP-Attention Architecture 15 May 2025 · 0 repositories · arXiv:2505.09907
-
CAFE: Retrieval Head-based Coarse-to-Fine Information Seeking to Enhance Multi-Document QA Capability 15 May 2025 · 0 repositories · arXiv:2505.10063
-
CL-RAG: Bridging the Gap in Retrieval-Augmented Generation with Curriculum Learning 15 May 2025 · 0 repositories · arXiv:2505.10493
-
Comparing LLM Text Annotation Skills: A Study on Human Rights Violations in Social Media Data 15 May 2025 · 1 repository · arXiv:2505.10260
-
ComplexFormer: Disruptively Advancing Transformer Inference Ability via Head-Specific Complex Vector Attention 15 May 2025 · 1 repository · arXiv:2505.10222
-
Defending the Edge: Representative-Attention for Mitigating Backdoor Attacks in Federated Learning 15 May 2025 · 0 repositories · arXiv:2505.10297
-
Does Scaling Law Apply in Time Series Forecasting? 15 May 2025 · 0 repositories · arXiv:2505.10172
-
Exploring Implicit Visual Misunderstandings in Multimodal Large Language Models through Attention Analysis 15 May 2025 · 1 repository · arXiv:2505.10541
-
Hierarchical Document Refinement for Long-context Retrieval-augmented Generation 15 May 2025 · 1 repository · arXiv:2505.10413
-
ILIF: Temporal Inhibitory Leaky Integrate-and-Fire Neuron for Overactivation in Spiking Neural Networks 15 May 2025 · 1 repository · arXiv:2505.10371
-
Leveraging Graph Retrieval-Augmented Generation to Support Learners' Understanding of Knowledge Concepts in MOOCs 15 May 2025 · 0 repositories · arXiv:2505.10074
-
MASS: Multi-Agent Simulation Scaling for Portfolio Construction 15 May 2025 · 1 repository · arXiv:2505.10278
-
MSCI: Addressing CLIP's Inherent Limitations for Compositional Zero-Shot Learning 15 May 2025 · 1 repository · arXiv:2505.10289Syntology official (archive's flag): 22 ran · 23 ran (of which 13 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 28 harvested samples) · 28 pointer-only (licence)
-
MTVCrafter: 4D Motion Tokenization for Open-World Human Image Animation 15 May 2025 · 1 repository · arXiv:2505.10238