Methods › General › Output Functions › Softmax › Papers, page 27
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 27 of 375: papers 2,601 to 2,700 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards Principled Learning for Re-ranking in Recommender Systems 5 Apr 2025 · 1 repository · arXiv:2504.04188
-
Transformer representation learning is necessary for dynamic multi-modal physiological data on small-cohort patients 5 Apr 2025 · 0 repositories · arXiv:2504.04120
-
Video4DGen: Enhancing Video and 4D Generation through Mutual Optimization 5 Apr 2025 · 1 repository · arXiv:2504.04153
-
Adaptive Classification of Interval-Valued Time Series 4 Apr 2025 · 0 repositories · arXiv:2504.03318
-
AdaViT: Adaptive Vision Transformer for Flexible Pretrain and Finetune with Variable 3D Medical Image Modalities 4 Apr 2025 · 0 repositories · arXiv:2504.03589
-
Beyond Progress Measures: Theoretical Insights into the Mechanism of Grokking 4 Apr 2025 · 1 repository · arXiv:2504.03162
-
Block Toeplitz Sparse Precision Matrix Estimation for Large-Scale Interval-Valued Time Series Forecasting 4 Apr 2025 · 0 repositories · arXiv:2504.03322
-
Detecting underdetermination in parameterized quantum circuits 4 Apr 2025 · 0 repositories · arXiv:2504.03315
-
Do LLM Evaluators Prefer Themselves for a Reason? 4 Apr 2025 · 1 repository · arXiv:2504.03846
-
DP-LET: An Efficient Spatio-Temporal Network Traffic Prediction Framework 4 Apr 2025 · 0 repositories · arXiv:2504.03792
-
Dynamic Importance in Diffusion U-Net for Enhanced Image Synthesis 4 Apr 2025 · 1 repository · arXiv:2504.03471
-
Efficient Dynamic Clustering-Based Document Compression for Retrieval-Augmented-Generation 4 Apr 2025 · 1 repository · arXiv:2504.03165
-
Electromyography-Based Gesture Recognition: Hierarchical Feature Extraction for Enhanced Spatial-Temporal Dynamics 4 Apr 2025 · 0 repositories · arXiv:2504.03221
-
FADConv: A Frequency-Aware Dynamic Convolution for Farmland Non-agriculturalization Identification and Segmentation 4 Apr 2025 · 0 repositories · arXiv:2504.03510
-
FaR: Enhancing Multi-Concept Text-to-Image Diffusion via Concept Fusion and Localized Refinement 4 Apr 2025 · 0 repositories · arXiv:2504.03292
-
Generating ensembles of spatially-coherent in-situ forecasts using flow matching 4 Apr 2025 · 0 repositories · arXiv:2504.03463
-
Generative AI Enhanced Financial Risk Management Information Retrieval 4 Apr 2025 · 1 repository · arXiv:2504.06293
-
HeterMoE: Efficient Training of Mixture-of-Experts Models on Heterogeneous GPUs 4 Apr 2025 · 0 repositories · arXiv:2504.03871
-
HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration 4 Apr 2025 · 0 repositories · arXiv:2504.03536
-
Inherent and emergent liability issues in LLM-based agentic systems: a principal-agent perspective 4 Apr 2025 · 0 repositories · arXiv:2504.03255
-
JanusDDG: A Thermodynamics-Compliant Model for Sequence-Based Protein Stability via Two-Fronts Multi-Head Attention 4 Apr 2025 · 1 repository · arXiv:2504.03278
-
Joint Retrieval of Cloud properties using Attention-based Deep Learning Models 4 Apr 2025 · 0 repositories · arXiv:2504.03133
-
Learning Natural Language Constraints for Safe Reinforcement Learning of Language Agents 4 Apr 2025 · 0 repositories · arXiv:2504.03185
-
Mamba as a Bridge: Where Vision Foundation Models Meet Vision Language Models for Domain-Generalized Semantic Segmentation 4 Apr 2025 · 1 repository · arXiv:2504.03193
-
Meta-DAN: towards an efficient prediction strategy for page-level handwritten text recognition 4 Apr 2025 · 1 repository · arXiv:2504.03349
-
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT 4 Apr 2025 · 0 repositories · arXiv:2504.07982
-
Model Reveals What to Cache: Profiling-Based Feature Reuse for Video Diffusion Models 4 Apr 2025 · 1 repository · arXiv:2504.03140Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-encoder nnU-Net outperforms Transformer models with self-supervised pretraining 4 Apr 2025 · 0 repositories · arXiv:2504.03474
-
Multi-Granularity Vision Fastformer with Fusion Mechanism for Skin Lesion Segmentation 4 Apr 2025 · 0 repositories · arXiv:2504.03108
-
Multilingual Retrieval-Augmented Generation for Knowledge-Intensive Task 4 Apr 2025 · 0 repositories · arXiv:2504.03616
-
NAACL2025 Tutorial: Adaptation of Large Language Models 4 Apr 2025 · 0 repositories · arXiv:2504.03931
-
Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models 4 Apr 2025 · 0 repositories · arXiv:2504.03624
-
Practical Poisoning Attacks against Retrieval-Augmented Generation 4 Apr 2025 · 0 repositories · arXiv:2504.03957
-
Rotation Invariance in Floor Plan Digitization using Zernike Moments 4 Apr 2025 · 0 repositories · arXiv:2504.03241
-
Structured Extraction of Process Structure Properties Relationships in Materials Science 4 Apr 2025 · 0 repositories · arXiv:2504.03979
-
TQD-Track: Temporal Query Denoising for 3D Multi-Object Tracking 4 Apr 2025 · 0 repositories · arXiv:2504.03258
-
VISTA-OCR: Towards generative and interactive end to end OCR models 4 Apr 2025 · 0 repositories · arXiv:2504.03621
-
ZFusion: An Effective Fuser of Camera and 4D Radar for 3D Object Perception in Autonomous Driving 4 Apr 2025 · 0 repositories · arXiv:2504.03438
-
A Framework for Situating Innovations, Opportunities, and Challenges in Advancing Vertical Systems with Large AI Models 3 Apr 2025 · 0 repositories · arXiv:2504.02793
-
A Sensorimotor Vision Transformer 3 Apr 2025 · 0 repositories · arXiv:2504.02536
-
AC-LoRA: Auto Component LoRA for Personalized Artistic Style Image Generation 3 Apr 2025 · 0 repositories · arXiv:2504.02231
-
AD-GPT: Large Language Models in Alzheimer's Disease 3 Apr 2025 · 0 repositories · arXiv:2504.03071
-
Adapting Large Language Models for Multi-Domain Retrieval-Augmented-Generation 3 Apr 2025 · 0 repositories · arXiv:2504.02411
-
Attention-Aware Multi-View Pedestrian Tracking 3 Apr 2025 · 0 repositories · arXiv:2504.03047
-
Beyond Conventional Transformers: The Medical X-ray Attention (MXA) Block for Improved Multi-Label Diagnosis Using Knowledge Distillation 3 Apr 2025 · 1 repository · arXiv:2504.02277
-
Cognitive Memory in Large Language Models 3 Apr 2025 · 0 repositories · arXiv:2504.02441
-
CoLa -- Learning to Interactively Collaborate with Large LMs 3 Apr 2025 · 0 repositories · arXiv:2504.02965
-
Computing High-dimensional Confidence Sets for Arbitrary Distributions 3 Apr 2025 · 0 repositories · arXiv:2504.02723
-
Deep Reinforcement Learning via Object-Centric Attention 3 Apr 2025 · 1 repository · arXiv:2504.03024
-
F-ViTA: Foundation Model Guided Visible to Thermal Translation 3 Apr 2025 · 1 repository · arXiv:2504.02801
-
FT-Transformer: Resilient and Reliable Transformer with End-to-End Fault Tolerant Attention 3 Apr 2025 · 0 repositories · arXiv:2504.02211
-
GPTAQ: Efficient Finetuning-Free Quantization for Asymmetric Calibration 3 Apr 2025 · 2 repositories · arXiv:2504.02692Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Graph Attention for Heterogeneous Graphs with Positional Encoding 3 Apr 2025 · 1 repository · arXiv:2504.02938
-
Graphs are everywhere -- Psst! In Music Recommendation too 3 Apr 2025 · 0 repositories · arXiv:2504.02598
-
Group-based Distinctive Image Captioning with Memory Difference Encoding and Attention 3 Apr 2025 · 0 repositories · arXiv:2504.02496
-
HGFormer: Topology-Aware Vision Transformer with HyperGraph Learning 3 Apr 2025 · 0 repositories · arXiv:2504.02440
-
HQViT: Hybrid Quantum Vision Transformer for Image Classification 3 Apr 2025 · 0 repositories · arXiv:2504.02730
-
Hummus: A Dataset of Humorous Multimodal Metaphor Use 3 Apr 2025 · 1 repository · arXiv:2504.02983
-
HyperRAG: Enhancing Quality-Efficiency Tradeoffs in Retrieval-Augmented Generation with Reranker KV-Cache Reuse 3 Apr 2025 · 0 repositories · arXiv:2504.02921
-
Hyperspectral Remote Sensing Images Salient Object Detection: The First Benchmark Dataset and Baseline 3 Apr 2025 · 1 repository · arXiv:2504.02416
-
LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models 3 Apr 2025 · 0 repositories · arXiv:2504.02327
-
Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval 3 Apr 2025 · 0 repositories · arXiv:2504.02397
-
Localized Definitions and Distributed Reasoning: A Proof-of-Concept Mechanistic Interpretability Study via Activation Patching 3 Apr 2025 · 1 repository · arXiv:2504.02976
-
MegaScale-Infer: Serving Mixture-of-Experts at Scale with Disaggregated Expert Parallelism 3 Apr 2025 · 0 repositories · arXiv:2504.02263
-
MMTL-UniAD: A Unified Framework for Multimodal and Multi-Task Learning in Assistive Driving Perception 3 Apr 2025 · 1 repository · arXiv:2504.02264Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-Head Adaptive Graph Convolution Network for Sparse Point Cloud-Based Human Activity Recognition 3 Apr 2025 · 1 repository · arXiv:2504.02778
-
On Vanishing Variance in Transformer Length Generalization 3 Apr 2025 · 0 repositories · arXiv:2504.02827
-
QID: Efficient Query-Informed ViTs in Data-Scarce Regimes for OCR-free Visual Document Understanding 3 Apr 2025 · 0 repositories · arXiv:2504.02971
-
Secure Generalization through Stochastic Bidirectional Parameter Updates Using Dual-Gradient Mechanism 3 Apr 2025 · 0 repositories · arXiv:2504.02213
-
Semiconductor Wafer Map Defect Classification with Tiny Vision Transformers 3 Apr 2025 · 0 repositories · arXiv:2504.02494
-
SLACK: Attacking LiDAR-based SLAM with Adversarial Point Injections 3 Apr 2025 · 0 repositories · arXiv:2504.03089
-
Spline-based Transformers 3 Apr 2025 · 0 repositories · arXiv:2504.02797
-
Task as Context Prompting for Accurate Medical Symptom Coding Using Large Language Models 3 Apr 2025 · 1 repository · arXiv:2504.03051
-
Towards Computation- and Communication-efficient Computational Pathology 3 Apr 2025 · 0 repositories · arXiv:2504.02628
-
Why do LLMs attend to the first token? 3 Apr 2025 · 1 repository · arXiv:2504.02732Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
WonderTurbo: Generating Interactive 3D World in 0.72 Seconds 3 Apr 2025 · 0 repositories · arXiv:2504.02261
-
A Prefixed Patch Time Series Transformer for Two-Point Boundary Value Problems in Three-Body Problems 2 Apr 2025 · 0 repositories · arXiv:2504.01464
-
A thorough benchmark of automatic text classification: From traditional approaches to large language models 2 Apr 2025 · 1 repository · arXiv:2504.01930
-
Analysis of an Idealized Stochastic Polyak Method and its Application to Black-Box Model Distillation 2 Apr 2025 · 0 repositories · arXiv:2504.01898
-
Attention Mamba: Time Series Modeling with Adaptive Pooling Acceleration and Receptive Field Enhancements 2 Apr 2025 · 0 repositories · arXiv:2504.02013
-
BioAtt: Anatomical Prior Driven Low-Dose CT Denoising 2 Apr 2025 · 0 repositories · arXiv:2504.01662
-
Biomedical Question Answering via Multi-Level Summarization on a Local Knowledge Graph 2 Apr 2025 · 0 repositories · arXiv:2504.01309
-
BlenderGym: Benchmarking Foundational Model Systems for Graphics Editing 2 Apr 2025 · 1 repository · arXiv:2504.01786
-
BOLDSimNet: Examining Brain Network Similarity between Task and Resting-State fMRI 2 Apr 2025 · 0 repositories · arXiv:2504.01274
-
Breaking BERT: Gradient Attack on Twitter Sentiment Analysis for Targeted Misclassification 2 Apr 2025 · 1 repository · arXiv:2504.01345
-
Chain of Correction for Full-text Speech Recognition with Large Language Models 2 Apr 2025 · 0 repositories · arXiv:2504.01519
-
Coarse-to-Fine Semantic Communication Systems for Text Transmission 2 Apr 2025 · 0 repositories · arXiv:2504.01442
-
Context-Aware Toxicity Detection in Multiplayer Games: Integrating Domain-Adaptive Pretraining and Match Metadata 2 Apr 2025 · 1 repository · arXiv:2504.01534
-
CoRAG: Collaborative Retrieval-Augmented Generation 2 Apr 2025 · 0 repositories · arXiv:2504.01883
-
Decoding Covert Speech from EEG Using a Functional Areas Spatio-Temporal Transformer 2 Apr 2025 · 1 repository · arXiv:2504.03762
-
Deep Representation Learning for Unsupervised Clustering of Myocardial Fiber Trajectories in Cardiac Diffusion Tensor Imaging 2 Apr 2025 · 0 repositories · arXiv:2504.01953
-
Dual-stream Transformer-GCN Model with Contextualized Representations Learning for Monocular 3D Human Pose Estimation 2 Apr 2025 · 1 repository · arXiv:2504.01764
-
Efficient Model Selection for Time Series Forecasting via LLMs 2 Apr 2025 · 0 repositories · arXiv:2504.02119
-
Enhancing Traffic Sign Recognition On The Performance Based On Yolov8 2 Apr 2025 · 0 repositories · arXiv:2504.02884
-
GaussianLSS -- Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting 2 Apr 2025 · 0 repositories · arXiv:2504.01957
-
Geometric Reasoning in the Embedding Space 2 Apr 2025 · 0 repositories · arXiv:2504.02018
-
GeoRAG: A Question-Answering Approach from a Geographical Perspective 2 Apr 2025 · 0 repositories · arXiv:2504.01458
-
GPT Adoption and the Impact of Disclosure Policies 2 Apr 2025 · 0 repositories · arXiv:2504.01566
-
GTR: Graph-Table-RAG for Cross-Table Question Answering 2 Apr 2025 · 0 repositories · arXiv:2504.01346
-
InvFussion: Bridging Supervised and Zero-shot Diffusion for Inverse Problems 2 Apr 2025 · 1 repository · arXiv:2504.01689