Methods › General › Output Functions › Softmax › Papers, page 76
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 76 of 375: papers 7,501 to 7,600 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Geometric Point Attention Transformer for 3D Shape Reassembly 26 Nov 2024 · 0 repositories · arXiv:2411.17788
-
"Give me the code" -- Log Analysis of First-Year CS Students' Interactions With GPT 26 Nov 2024 · 0 repositories · arXiv:2411.17855
-
GMFlow: Global Motion-Guided Recurrent Flow for 6D Object Pose Estimation 26 Nov 2024 · 0 repositories · arXiv:2411.17174
-
H³Fusion: Helpful, Harmless, Honest Fusion of Aligned LLMs 26 Nov 2024 · 1 repository · arXiv:2411.17792
-
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning 26 Nov 2024 · 1 repository · arXiv:2411.17100
-
LampMark: Proactive Deepfake Detection via Training-Free Landmark Perceptual Watermarks 26 Nov 2024 · 1 repository · arXiv:2411.17209
-
Learning Chemical Reaction Representation with Reactant-Product Alignment 26 Nov 2024 · 0 repositories · arXiv:2411.17629
-
Learning Monotonic Attention in Transducer for Streaming Generation 26 Nov 2024 · 1 repository · arXiv:2411.17170
-
Leveraging Large Language Models and Topic Modeling for Toxicity Classification 26 Nov 2024 · 1 repository · arXiv:2411.17876
-
LiteVAR: Compressing Visual Autoregressive Modelling with Efficient Attention and Quantization 26 Nov 2024 · 1 repository · arXiv:2411.17178
-
MADE: Graph Backdoor Defense with Masked Unlearning 26 Nov 2024 · 0 repositories · arXiv:2411.18648
-
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation 26 Nov 2024 · 2 repositories · arXiv:2411.17945Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
MAT: Multi-Range Attention Transformer for Efficient Image Super-Resolution 26 Nov 2024 · 1 repository · arXiv:2411.17214Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Mixed-State Quantum Denoising Diffusion Probabilistic Model 26 Nov 2024 · 0 repositories · arXiv:2411.17608
-
Multimodal Outer Arithmetic Block Dual Fusion of Whole Slide Images and Omics Data for Precision Oncology 26 Nov 2024 · 0 repositories · arXiv:2411.17418
-
MWFormer: Multi-Weather Image Restoration Using Degradation-Aware Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17226
-
On Limitations of LLM as Annotator for Low Resource Languages 26 Nov 2024 · 0 repositories · arXiv:2411.17637
-
One Mind, Many Tongues: A Deep Dive into Language-Agnostic Knowledge Neurons in Large Language Models 26 Nov 2024 · 0 repositories · arXiv:2411.17401
-
ΩSFormer: Dual-Modal Ω-like Super-Resolution Transformer Network for Cross-scale and High-accuracy Terraced Field Vectorization Extraction 26 Nov 2024 · 0 repositories · arXiv:2411.17088
-
Pretrained LLM Adapted with LoRA as a Decision Transformer for Offline RL in Quantitative Trading 26 Nov 2024 · 1 repository · arXiv:2411.17900
-
Push the Limit of Multi-modal Emotion Recognition by Prompting LLMs with Receptive-Field-Aware Attention Weighting 26 Nov 2024 · 0 repositories · arXiv:2411.17674
-
SatVision-TOA: A Geospatial Foundation Model for Coarse-Resolution All-Sky Remote Sensing Imagery 26 Nov 2024 · 1 repository · arXiv:2411.17000
-
Scalable iterative pruning of large language and vision models using block coordinate descent 26 Nov 2024 · 0 repositories · arXiv:2411.17796
-
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation 26 Nov 2024 · 0 repositories · arXiv:2411.17061
-
SoftmAP: Software-Hardware Co-design for Integer-Only Softmax on Associative Processors 26 Nov 2024 · 0 repositories · arXiv:2411.17847
-
Star Attention: Efficient LLM Inference over Long Sequences 26 Nov 2024 · 1 repository · arXiv:2411.17116Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Structure-Guided MR-to-CT Synthesis with Spatial and Semantic Alignments for Attenuation Correction of Whole-Body PET/MR Imaging 26 Nov 2024 · 0 repositories · arXiv:2411.17488
-
TAFM-Net: A Novel Approach to Skin Lesion Segmentation Using Transformer Attention and Focal Modulation 26 Nov 2024 · 0 repositories · arXiv:2411.17556
-
TED-VITON: Transformer-Empowered Diffusion Models for Virtual Try-On 26 Nov 2024 · 1 repository · arXiv:2411.17017
-
TinyViM: Frequency Decoupling for Tiny Hybrid Vision Mamba 26 Nov 2024 · 1 repository · arXiv:2411.17473
-
What Differentiates Educational Literature? A Multimodal Fusion Approach of Transformers and Computational Linguistics 26 Nov 2024 · 0 repositories · arXiv:2411.17593
-
What's in the Image? A Deep-Dive into the Vision of Vision Language Models 26 Nov 2024 · 0 repositories · arXiv:2411.17491
-
Adaptive Circuit Behavior and Generalization in Mechanistic Interpretability 25 Nov 2024 · 0 repositories · arXiv:2411.16105
-
Are Transformers Truly Foundational for Robotics? 25 Nov 2024 · 0 repositories · arXiv:2411.16917
-
AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning 25 Nov 2024 · 1 repository · arXiv:2411.16495Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering 25 Nov 2024 · 1 repository · arXiv:2411.16863Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Boosting 3D Object Generation through PBR Materials 25 Nov 2024 · 0 repositories · arXiv:2411.16080
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
CARE Transformer: Mobile-Friendly Linear Visual Transformer via Decoupled Dual Interaction 25 Nov 2024 · 0 repositories · arXiv:2411.16170
-
CATP-LLM: Empowering Large Language Models for Cost-Aware Tool Planning 25 Nov 2024 · 0 repositories · arXiv:2411.16313
-
CMAViT: Integrating Climate, Managment, and Remote Sensing Data for Crop Yield Estimation with Multimodel Vision Transformers 25 Nov 2024 · 0 repositories · arXiv:2411.16989
-
CoCoNO: Attention Contrast-and-Complete for Initial Noise Optimization in Text-to-Image Synthesis 25 Nov 2024 · 0 repositories · arXiv:2411.16783
-
Comparative Analysis of Machine Learning Models for Short-Term Distribution System Load Forecasting 25 Nov 2024 · 0 repositories · arXiv:2411.16118
-
Contrastive Multi-graph Learning with Neighbor Hierarchical Sifting for Semi-supervised Text Classification 25 Nov 2024 · 1 repository · arXiv:2411.16787
-
DF-GNN: Dynamic Fusion Framework for Attention Graph Neural Networks on GPUs 25 Nov 2024 · 1 repository · arXiv:2411.16127
-
DreamRunner: Fine-Grained Storytelling Video Generation with Retrieval-Augmented Motion Adaptation 25 Nov 2024 · 0 repositories · arXiv:2411.16657
-
Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16991
-
Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16797
-
Enhancing Fluorescence Lifetime Parameter Estimation Accuracy with Differential Transformer Based Deep Learning Model Incorporating Pixelwise Instrument Response Function 25 Nov 2024 · 0 repositories · arXiv:2411.16896
-
Enhancing Multi-Agent Consensus through Third-Party LLM Integration: Analyzing Uncertainty and Mitigating Hallucinations in Large Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16189
-
Even Sparser Graph Transformers 25 Nov 2024 · 1 repository · arXiv:2411.16278Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
Explainable AI Approach using Near Misses Analysis 25 Nov 2024 · 0 repositories · arXiv:2411.16895
-
Factorized Visual Tokenization and Generation 25 Nov 2024 · 0 repositories · arXiv:2411.16681
-
Fine-Tuning LLMs with Noisy Data for Political Argument Generation and Post Guidance 25 Nov 2024 · 0 repositories · arXiv:2411.16813
-
Harnessing Superclasses for Learning from Hierarchical Databases 25 Nov 2024 · 1 repository · arXiv:2411.16438
-
Human-Calibrated Automated Testing and Validation of Generative Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16391
-
Image Generation Diversity Issues and How to Tame Them 25 Nov 2024 · 1 repository · arXiv:2411.16171
-
Interpreting Object-level Foundation Models via Visual Precision Search 25 Nov 2024 · 2 repositories · arXiv:2411.16198
-
J-CaPA : Joint Channel and Pyramid Attention Improves Medical Image Segmentation 25 Nov 2024 · 0 repositories · arXiv:2411.16568
-
LaB-RAG: Label Boosted Retrieval Augmented Generation for Radiology Report Generation 25 Nov 2024 · 1 repository · arXiv:2411.16523
-
Local and Global Feature Attention Fusion Network for Face Recognition 25 Nov 2024 · 0 repositories · arXiv:2411.16169
-
MarketGPT: Developing a Pre-trained transformer (GPT) for Modeling Financial Time Series 25 Nov 2024 · 1 repository · arXiv:2411.16585Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Noise Diffusion for Enhancing Semantic Faithfulness in Text-to-Image Synthesis 25 Nov 2024 · 0 repositories · arXiv:2411.16503
-
NormXLogit: The Head-on-Top Never Lies 25 Nov 2024 · 0 repositories · arXiv:2411.16252
-
Predictive Power of LLMs in Financial Markets 25 Nov 2024 · 0 repositories · arXiv:2411.16569
-
Privacy Protection in Personalized Diffusion Models via Targeted Cross-Attention Adversarial Attack 25 Nov 2024 · 0 repositories · arXiv:2411.16437
-
A Physics-Inspired Deep Learning Framework with Polar Coordinate Attention for Ptychographic Imaging 25 Nov 2024 · 1 repository · arXiv:2412.06806
-
Quadratic Gaussian Splatting for Efficient and Detailed Surface Reconstruction 25 Nov 2024 · 0 repositories · arXiv:2411.16392
-
Scaling Spike-driven Transformer with Efficient Spike Firing Approximation Training 25 Nov 2024 · 1 repository · arXiv:2411.16061Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
SEMU-Net: A Segmentation-based Corrector for Fabrication Process Variations of Nanophotonics with Microscopic Images 25 Nov 2024 · 0 repositories · arXiv:2411.16973
-
Soft-TransFormers for Continual Learning 25 Nov 2024 · 1 repository · arXiv:2411.16073
-
Solaris: A Foundation Model of the Sun 25 Nov 2024 · 0 repositories · arXiv:2411.16339
-
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training 25 Nov 2024 · 0 repositories · arXiv:2411.16618
-
Swin fMRI Transformer Predicts Early Neurodevelopmental Outcomes from Neonatal fMRI 25 Nov 2024 · 0 repositories · arXiv:2412.07783
-
Tree Transformers are an Ineffective Model of Syntactic Constituency 25 Nov 2024 · 0 repositories · arXiv:2411.16993
-
UltraSam: A Foundation Model for Ultrasound using Large Open-Access Segmentation Datasets 25 Nov 2024 · 1 repository · arXiv:2411.16222
-
Unlocking the Potential of Text-to-Image Diffusion with PAC-Bayesian Theory 25 Nov 2024 · 0 repositories · arXiv:2411.17472
-
Unraveling Arithmetic in Large Language Models: The Role of Algebraic Structures 25 Nov 2024 · 0 repositories · arXiv:2411.16260
-
VICON: Vision In-Context Operator Networks for Multi-Physics Fluid Dynamics Prediction 25 Nov 2024 · 1 repository · arXiv:2411.16063
-
VIRES: Video Instance Repainting via Sketch and Text Guided Generation 25 Nov 2024 · 0 repositories · arXiv:2411.16199
-
VQ-SGen: A Vector Quantized Stroke Representation for Creative Sketch Generation 25 Nov 2024 · 0 repositories · arXiv:2411.16446
-
WTDUN: Wavelet Tree-Structured Sampling and Deep Unfolding Network for Image Compressed Sensing 25 Nov 2024 · 0 repositories · arXiv:2411.16336
-
A General Sensing-assisted Channel Estimation Framework in Distributed MIMO Network 24 Nov 2024 · 0 repositories · arXiv:2411.15995
-
A Method for Building Large Language Models with Predefined KV Cache Capacity 24 Nov 2024 · 0 repositories · arXiv:2411.15785
-
Beyond adaptive gradient: Fast-Controlled Minibatch Algorithm for large-scale optimization 24 Nov 2024 · 1 repository · arXiv:2411.15795
-
Development of Pre-Trained Transformer-based Models for the Nepali Language 24 Nov 2024 · 0 repositories · arXiv:2411.15734
-
FastTrackTr:Towards Fast Multi-Object Tracking with Transformers 24 Nov 2024 · 0 repositories · arXiv:2411.15811
-
Fixing the Perspective: A Critical Examination of Zero-1-to-3 24 Nov 2024 · 0 repositories · arXiv:2411.15706
-
Gradient Norm Regularization Second-Order Algorithms for Solving Nonconvex-Strongly Concave Minimax Problems 24 Nov 2024 · 0 repositories · arXiv:2411.15769
-
Investigating Factuality in Long-Form Text Generation: The Roles of Self-Known and Self-Unknown 24 Nov 2024 · 0 repositories · arXiv:2411.15993
-
LetsTalk: Latent Diffusion Transformer for Talking Video Synthesis 24 Nov 2024 · 0 repositories · arXiv:2411.16748
-
LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training 24 Nov 2024 · 1 repository · arXiv:2411.15708Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LTCF-Net: A Transformer-Enhanced Dual-Channel Fourier Framework for Low-Light Image Restoration 24 Nov 2024 · 0 repositories · arXiv:2411.15740
-
Medical Slice Transformer: Improved Diagnosis and Explainability on 3D Medical Images with DINOv2 24 Nov 2024 · 1 repository · arXiv:2411.15802
-
Nimbus: Secure and Efficient Two-Party Inference for Transformers 24 Nov 2024 · 1 repository · arXiv:2411.15707Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
PR-MIM: Delving Deeper into Partial Reconstruction in Masked Image Modeling 24 Nov 2024 · 0 repositories · arXiv:2411.15746
-
RAMIE: Retrieval-Augmented Multi-task Information Extraction with Large Language Models on Dietary Supplements 24 Nov 2024 · 0 repositories · arXiv:2411.15700
-
ResCLIP: Residual Attention for Training-free Dense Vision-language Inference 24 Nov 2024 · 1 repository · arXiv:2411.15851
-
Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation 24 Nov 2024 · 1 repository · arXiv:2411.15869Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Test-time Alignment-Enhanced Adapter for Vision-Language Models 24 Nov 2024 · 1 repository · arXiv:2411.15735