Methods › General › Output Functions › Softmax › Papers, page 48
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 48 of 375: papers 4,701 to 4,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Next Block Prediction: Video Generation via Semi-Auto-Regressive Modeling 11 Feb 2025 · 0 repositories · arXiv:2502.07737
-
On Training-Conditional Conformal Prediction and Binomial Proportion Confidence Intervals 11 Feb 2025 · 0 repositories · arXiv:2502.07497
-
OpenGrok: Enhancing SNS Data Processing with Distilled Knowledge and Mask-like Mechanisms 11 Feb 2025 · 1 repository · arXiv:2502.07312
-
Optimizing Knowledge Distillation in Transformers: Enabling Multi-Head Attention without Alignment Barriers 11 Feb 2025 · 0 repositories · arXiv:2502.07436
-
Pippo: High-Resolution Multi-View Humans from a Single Image 11 Feb 2025 · 0 repositories · arXiv:2502.07785
-
Rethinking Timing Residuals: Advancing PET Detectors with Explicit TOF Corrections 11 Feb 2025 · 0 repositories · arXiv:2502.07630
-
Robust Indoor Localization in Dynamic Environments: A Multi-source Unsupervised Domain Adaptation Framework 11 Feb 2025 · 0 repositories · arXiv:2502.07246
-
Small Language Model Makes an Effective Long Text Extractor 11 Feb 2025 · 1 repository · arXiv:2502.07286
-
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer 11 Feb 2025 · 0 repositories · arXiv:2502.07216
-
Spatial Degradation-Aware and Temporal Consistent Diffusion Model for Compressed Video Super-Resolution 11 Feb 2025 · 0 repositories · arXiv:2502.07381
-
TextAtlas5M: A Large-scale Dataset for Dense Text Image Generation 11 Feb 2025 · 1 repository · arXiv:2502.07870
-
Towards More Accurate Full-Atom Antibody Co-Design 11 Feb 2025 · 0 repositories · arXiv:2502.19391
-
Tractable Transformers for Flexible Conditional Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07616
-
TransMLA: Multi-Head Latent Attention Is All You Need 11 Feb 2025 · 1 repository · arXiv:2502.07864
-
Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification 11 Feb 2025 · 0 repositories · arXiv:2502.09647
-
VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07531
-
A Simple yet Effective DDG Predictor is An Unsupervised Antibody Optimizer and Explainer 10 Feb 2025 · 1 repository · arXiv:2502.06913
-
Active Inference through Incentive Design in Markov Decision Processes 10 Feb 2025 · 0 repositories · arXiv:2502.07065
-
An Appearance Defect Detection Method for Cigarettes Based on C-CenterNet 10 Feb 2025 · 0 repositories · arXiv:2502.06119
-
C-3PO: Compact Plug-and-Play Proxy Optimization to Achieve Human-like Retrieval-Augmented Generation 10 Feb 2025 · 0 repositories · arXiv:2502.06205Syntology 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Col-OLHTR: A Novel Framework for Multimodal Online Handwritten Text Recognition 10 Feb 2025 · 0 repositories · arXiv:2502.06100
-
Completeness of Datasets Documentation on ML/AI repositories: an Empirical Investigation 10 Feb 2025 · 0 repositories · arXiv:2503.13463
-
Conditional diffusion model with spatial attention and latent embedding for medical image segmentation 10 Feb 2025 · 1 repository · arXiv:2502.06997
-
ConMeC: A Dataset for Metonymy Resolution with Common Nouns 10 Feb 2025 · 1 repository · arXiv:2502.06087
-
CustomVideoX: 3D Reference Attention Driven Dynamic Adaptation for Zero-Shot Customized Video Diffusion Transformers 10 Feb 2025 · 0 repositories · arXiv:2502.06527
-
DebateBench: A Challenging Long Context Reasoning Benchmark For Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06279
-
Deep Learning in Automated Power Line Inspection: A Review 10 Feb 2025 · 0 repositories · arXiv:2502.07826
-
Do Attention Heads Compete or Cooperate during Counting? 10 Feb 2025 · 0 repositories · arXiv:2502.06923
-
Early Operative Difficulty Assessment in Laparoscopic Cholecystectomy via Snapshot-Centric Video Analysis 10 Feb 2025 · 1 repository · arXiv:2502.07008
-
Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile 10 Feb 2025 · 1 repository · arXiv:2502.06155
-
Event Vision Sensor: A Review 10 Feb 2025 · 0 repositories · arXiv:2502.06116
-
Facial Analysis Systems and Down Syndrome 10 Feb 2025 · 0 repositories · arXiv:2502.06341
-
Find Central Dogma Again: Leveraging Multilingual Transfer in Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06253
-
Finding Words Associated with DIF: Predicting Differential Item Functioning using LLMs and Explainable AI 10 Feb 2025 · 0 repositories · arXiv:2502.07017
-
Foundation Model of Electronic Medical Records for Adaptive Risk Estimation 10 Feb 2025 · 1 repository · arXiv:2502.06124
-
Fully Exploiting Vision Foundation Model's Profound Prior Knowledge for Generalizable RGB-Depth Driving Scene Parsing 10 Feb 2025 · 0 repositories · arXiv:2502.06219
-
FunduSAM: A Specialized Deep Learning Model for Enhanced Optic Disc and Cup Segmentation in Fundus Images 10 Feb 2025 · 0 repositories · arXiv:2502.06220
-
History-Guided Video Diffusion 10 Feb 2025 · 1 repository · arXiv:2502.06764
-
Inventory Consensus Control in Supply Chain Networks using Dissipativity-Based Control and Topology Co-Design 10 Feb 2025 · 0 repositories · arXiv:2502.06580
-
Is Long Range Sequential Modeling Necessary For Colorectal Tumor Segmentation? 10 Feb 2025 · 0 repositories · arXiv:2502.07120
-
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights 10 Feb 2025 · 0 repositories · arXiv:2502.07049
-
Leveraging GPT-4o Efficiency for Detecting Rework Anomaly in Business Processes 10 Feb 2025 · 0 repositories · arXiv:2502.06918
-
Marginal Mechanisms For Balanced Exchange 10 Feb 2025 · 0 repositories · arXiv:2502.06499
-
Multimodal Task Representation Memory Bank vs. Catastrophic Forgetting in Anomaly Detection 10 Feb 2025 · 0 repositories · arXiv:2502.06194
-
Neighborhood-Order Learning Graph Attention Network for Fake News Detection 10 Feb 2025 · 1 repository · arXiv:2502.06927
-
No Trick, No Treat: Pursuits and Challenges Towards Simulation-free Training of Neural Samplers 10 Feb 2025 · 0 repositories · arXiv:2502.06685
-
Optimizing Knowledge Integration in Retrieval-Augmented Generation with Self-Selection 10 Feb 2025 · 0 repositories · arXiv:2502.06148
-
Powerformer: A Transformer with Weighted Causal Attention for Time-series Forecasting 10 Feb 2025 · 1 repository · arXiv:2502.06151
-
Progressive Collaborative and Semantic Knowledge Fusion for Generative Recommendation 10 Feb 2025 · 0 repositories · arXiv:2502.06269
-
Prompt-SID: Learning Structural Representation Prompt via Latent Diffusion for Single-Image Denoising 10 Feb 2025 · 1 repository · arXiv:2502.06432
-
RALLRec: Improving Retrieval Augmented Large Language Model Recommendation with Representation Learning 10 Feb 2025 · 1 repository · arXiv:2502.06101
-
RelGNN: Composite Message Passing for Relational Deep Learning 10 Feb 2025 · 1 repository · arXiv:2502.06784Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Structural Reformation of Large Language Model Neuron Encapsulation for Divergent Information Aggregation 10 Feb 2025 · 0 repositories · arXiv:2502.07124
-
Systematic Outliers in Large Language Models 10 Feb 2025 · 1 repository · arXiv:2502.06415Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
The exponential distribution of the orders of demonstrative, numeral, adjective and noun 10 Feb 2025 · 0 repositories · arXiv:2502.06342
-
Towards bandit-based prompt-tuning for in-the-wild foundation agents 10 Feb 2025 · 0 repositories · arXiv:2502.06358
-
Towards Copyright Protection for Knowledge Bases of Retrieval-augmented Language Models via Reasoning 10 Feb 2025 · 0 repositories · arXiv:2502.10440
-
Unconstrained Body Recognition at Altitude and Range: Comparing Four Approaches 10 Feb 2025 · 0 repositories · arXiv:2502.07130
-
Unleashing the Potential of Pre-Trained Diffusion Models for Generalizable Person Re-Identification 10 Feb 2025 · 1 repository · arXiv:2502.06619
-
Utilizing Novelty-based Evolution Strategies to Train Transformers in Reinforcement Learning 10 Feb 2025 · 0 repositories · arXiv:2502.06301
-
ViSIR: Vision Transformer Single Image Reconstruction Method for Earth System Models 10 Feb 2025 · 0 repositories · arXiv:2502.06741
-
Wandering around: A bioinspired approach to visual attention through object motion sensitivity 10 Feb 2025 · 1 repository · arXiv:2502.06747
-
Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models 9 Feb 2025 · 0 repositories · arXiv:2502.06039
-
Emergence of Episodic Memory in Transformers: Characterizing Changes in Temporal Structure of Attention Scores During Training 9 Feb 2025 · 0 repositories · arXiv:2502.06902
-
Enhancing Financial Time-Series Forecasting with Retrieval-Augmented Large Language Models 9 Feb 2025 · 0 repositories · arXiv:2502.05878
-
HyLiFormer: Hyperbolic Linear Attention for Skeleton-based Human Action Recognition 9 Feb 2025 · 0 repositories
-
Injecting Universal Jailbreak Backdoors into LLMs in Minutes 9 Feb 2025 · 1 repository · arXiv:2502.10438Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Investigating Compositional Reasoning in Time Series Foundation Models 9 Feb 2025 · 1 repository · arXiv:2502.06037
-
Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End" 9 Feb 2025 · 0 repositories · arXiv:2502.06898
-
HSI: Head-Specific Intervention Can Induce Misaligned AI Coordination in Large Language Models 9 Feb 2025 · 1 repository · arXiv:2502.05945
-
Linear Attention Modeling for Learned Image Compression 9 Feb 2025 · 0 repositories · arXiv:2502.05741Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
LM2: Large Memory Models 9 Feb 2025 · 2 repositories · arXiv:2502.06049
-
MoEMba: A Mamba-based Mixture of Experts for High-Density EMG-based Hand Gesture Recognition 9 Feb 2025 · 0 repositories · arXiv:2502.17457
-
On the use of Performer and Agent Attention for Spoken Language Identification 9 Feb 2025 · 0 repositories · arXiv:2502.05841
-
Performance Analysis of Traditional VQA Models Under Limited Computational Resources 9 Feb 2025 · 0 repositories · arXiv:2502.05738
-
Provably Overwhelming Transformer Models with Designed Inputs 9 Feb 2025 · 0 repositories · arXiv:2502.06038
-
Satellite Observations Guided Diffusion Model for Accurate Meteorological States at Arbitrary Resolution 9 Feb 2025 · 0 repositories · arXiv:2502.07814
-
Saving 77% of the Parameters in Large Language Models Technical Report 9 Feb 2025 · 1 repository
-
ScaffoldGPT: A Scaffold-based GPT Model for Drug Optimization 9 Feb 2025 · 0 repositories · arXiv:2502.06891
-
SNAT-YOLO: Efficient Cross-Layer Aggregation Network for Edge-Oriented Gangue Detection 9 Feb 2025 · 0 repositories · arXiv:2502.05988
-
SphereFusion: Efficient Panorama Depth Estimation via Gated Fusion 9 Feb 2025 · 0 repositories · arXiv:2502.05859
-
Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding 9 Feb 2025 · 1 repository · arXiv:2502.06020
-
The Curse of Depth in Large Language Models 9 Feb 2025 · 0 repositories · arXiv:2502.05795
-
VFX Creator: Animated Visual Effect Generation with Controllable Diffusion Transformer 9 Feb 2025 · 0 repositories · arXiv:2502.05979
-
Dynamic Noise Preference Optimization for LLM Self-Improvement via Synthetic Data 8 Feb 2025 · 0 repositories · arXiv:2502.05400
-
A Novel Convolutional-Free Method for 3D Medical Imaging Segmentation 8 Feb 2025 · 0 repositories · arXiv:2502.05396
-
APE: Faster and Longer Context-Augmented Generation via Adaptive Parallel Encoding 8 Feb 2025 · 1 repository · arXiv:2502.05431Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Bridging Traffic State and Trajectory for Dynamic Road Network and Trajectory Representation Learning 8 Feb 2025 · 1 repository · arXiv:2502.06870
-
Demystifying Catastrophic Forgetting in Two-Stage Incremental Object Detector 8 Feb 2025 · 0 repositories · arXiv:2502.05540
-
Diffusion Model for Interest Refinement in Multi-Interest Recommendation 8 Feb 2025 · 0 repositories · arXiv:2502.05561
-
Event Stream-based Visual Object Tracking: HDETrack V2 and A High-Definition Benchmark 8 Feb 2025 · 1 repository · arXiv:2502.05574
-
Flow-based Conformal Prediction for Multi-dimensional Time Series 8 Feb 2025 · 0 repositories · arXiv:2502.05709
-
Flowing Through Layers: A Continuous Dynamical Systems Perspective on Transformers 8 Feb 2025 · 0 repositories · arXiv:2502.05656
-
Forbidden Science: Dual-Use AI Challenge Benchmark and Scientific Refusal Tests 8 Feb 2025 · 0 repositories · arXiv:2502.06867
-
Graph Neural Network Enabled Pinching Antennas 8 Feb 2025 · 0 repositories · arXiv:2502.05447
-
GWRF: A Generalizable Wireless Radiance Field for Wireless Signal Propagation Modeling 8 Feb 2025 · 0 repositories · arXiv:2502.05708
-
Hierarchical Lexical Manifold Projection in Large Language Models: A Novel Mechanism for Multi-Scale Semantic Representation 8 Feb 2025 · 0 repositories · arXiv:2502.05395
-
Knowledge Graph-Guided Retrieval Augmented Generation 8 Feb 2025 · 1 repository · arXiv:2502.06864
-
Lossless Acceleration of Large Language Models with Hierarchical Drafting based on Temporal Locality in Speculative Decoding 8 Feb 2025 · 1 repository · arXiv:2502.05609
-
MoFM: A Large-Scale Human Motion Foundation Model 8 Feb 2025 · 0 repositories · arXiv:2502.05432