Methods › General › Output Functions › Softmax › Papers, page 50
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 50 of 375: papers 4,901 to 5,000 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Learning to Synthesize Compatible Fashion Items Using Semantic Alignment and Collocation Classification: An Outfit Generation Framework 5 Feb 2025 · 0 repositories · arXiv:2502.06827
-
MARAGE: Transferable Multi-Model Adversarial Attack for Retrieval-Augmented Generation Data Extraction 5 Feb 2025 · 0 repositories · arXiv:2502.04360
-
Maximizing the Position Embedding for Vision Transformers with Global Average Pooling 5 Feb 2025 · 0 repositories · arXiv:2502.02919
-
Multimodal Brain-Computer Interfaces: AI-powered Decoding Methodologies 5 Feb 2025 · 0 repositories · arXiv:2502.02830
-
Multimodal Transformer Models for Turn-taking Prediction: Effects on Conversational Dynamics of Human-Agent Interaction during Cooperative Gameplay 5 Feb 2025 · 0 repositories · arXiv:2503.16432
-
Omni-DNA: A Unified Genomic Foundation Model for Cross-Modal and Multi-Task Learning 5 Feb 2025 · 0 repositories · arXiv:2502.03499
-
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices 5 Feb 2025 · 1 repository · arXiv:2502.04363
-
On Zero-Initialized Attention: Optimal Prompt and Gating Factor Estimation 5 Feb 2025 · 0 repositories · arXiv:2502.03029
-
OPTIC: Optimizing Patient-Provider Triaging & Improving Communications in Clinical Operations using GPT-4 Data Labeling and Model Distillation 5 Feb 2025 · 0 repositories · arXiv:2503.05701
-
Optimizing Robustness and Accuracy in Mixture of Experts: A Dual-Model Approach 5 Feb 2025 · 0 repositories · arXiv:2502.06832
-
Path Planning for Masked Diffusion Model Sampling 5 Feb 2025 · 0 repositories · arXiv:2502.03540
-
Scaling laws in wearable human activity recognition 5 Feb 2025 · 0 repositories · arXiv:2502.03364
-
Single Antenna Terahertz Sensing using Preconfigured Metasurfaces 5 Feb 2025 · 0 repositories · arXiv:2502.03291
-
SpaceGNN: Multi-Space Graph Neural Network for Node Anomaly Detection with Extremely Limited Labels 5 Feb 2025 · 1 repository · arXiv:2502.03201
-
Structured Token Retention and Computational Memory Paths in Large Language Models 5 Feb 2025 · 0 repositories · arXiv:2502.03102
-
TruePose: Human-Parsing-guided Attention Diffusion for Full-ID Preserving Pose Transfer 5 Feb 2025 · 0 repositories · arXiv:2502.03426
-
Type 2 Tobit Sample Selection Models with Bayesian Additive Regression Trees 5 Feb 2025 · 0 repositories · arXiv:2502.03600
-
ZISVFM: Zero-Shot Object Instance Segmentation in Indoor Robotic Environments with Vision Foundation Models 5 Feb 2025 · 0 repositories · arXiv:2502.03266
-
A Training-Free Length Extrapolation Approach for LLMs: Greedy Attention Logit Interpolation (GALI) 4 Feb 2025 · 1 repository · arXiv:2502.02659
-
AAD-DCE: An Aggregated Multimodal Attention Mechanism for Early and Late Dynamic Contrast Enhanced Prostate MRI Synthesis 4 Feb 2025 · 1 repository · arXiv:2502.02555
-
Adaptive Voxel-Weighted Loss Using L1 Norms in Deep Neural Networks for Detection and Segmentation of Prostate Cancer Lesions in PET/CT Images 4 Feb 2025 · 1 repository · arXiv:2502.02756
-
Aligning Human and Machine Attention for Enhanced Supervised Learning 4 Feb 2025 · 0 repositories · arXiv:2502.06811
-
Astromer 2 4 Feb 2025 · 0 repositories · arXiv:2502.02717
-
AutoGUI: Scaling GUI Grounding with Automatic Functionality Annotations from LLMs 4 Feb 2025 · 0 repositories · arXiv:2502.01977
-
Can LLMs Maintain Fundamental Abilities under KV Cache Compression? 4 Feb 2025 · 0 repositories · arXiv:2502.01941
-
CoAT: Chain-of-Associated-Thoughts Framework for Enhancing Large Language Models Reasoning 4 Feb 2025 · 0 repositories · arXiv:2502.02390
-
CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance 4 Feb 2025 · 1 repository · arXiv:2502.04350
-
Constrained belief updates explain geometric structures in transformer representations 4 Feb 2025 · 0 repositories · arXiv:2502.01954
-
Contextual Memory Reweaving in Large Language Models Using Layered Latent State Reconstruction 4 Feb 2025 · 0 repositories · arXiv:2502.02046
-
Conversation AI Dialog for Medicare powered by Finetuning and Retrieval Augmented Generation 4 Feb 2025 · 0 repositories · arXiv:2502.02249
-
Diff9D: Diffusion-Based Domain-Generalized Category-Level 9-DoF Object Pose Estimation 4 Feb 2025 · 1 repository · arXiv:2502.02525
-
Diffusion Instruction Tuning 4 Feb 2025 · 0 repositories · arXiv:2502.06814
-
Distribution Transformers: Fast Approximate Bayesian Inference With On-The-Fly Prior Adaptation 4 Feb 2025 · 0 repositories · arXiv:2502.02463
-
Dual Ensembled Multiagent Q-Learning with Hypernet Regularizer 4 Feb 2025 · 1 repository · arXiv:2502.02018
-
Dual-Flow: Transferable Multi-Target, Instance-Agnostic Attacks via In-the-wild Cascading Flow Optimization 4 Feb 2025 · 0 repositories · arXiv:2502.02096Syntology 0 ran · 3 unverified (of 3 harvested samples)
-
EdgeGFL: Rethinking Edge Information in Graph Feature Preference Learning 4 Feb 2025 · 0 repositories · arXiv:2502.02302
-
Evaluating the Effectiveness of LLMs in Fixing Maintainability Issues in Real-World Projects 4 Feb 2025 · 0 repositories · arXiv:2502.02368
-
Exact Sequence Classification with Hardmax Transformers 4 Feb 2025 · 0 repositories · arXiv:2502.02270
-
Exploiting Ensemble Learning for Cross-View Isolated Sign Language Recognition 4 Feb 2025 · 1 repository · arXiv:2502.02196
-
Exploring the Panorama of Anxiety Levels: A Multi-Scenario Study Based on Human-Centric Anxiety Level Detection and Personalized Guidance 4 Feb 2025 · 0 repositories · arXiv:2503.15527
-
FewTopNER: Integrating Few-Shot Learning with Topic Modeling and Named Entity Recognition in a Multilingual Framework 4 Feb 2025 · 1 repository · arXiv:2502.02391
-
IMDPrompter: Adapting SAM to Image Manipulation Detection by Cross-View Automated Prompt Learning 4 Feb 2025 · 0 repositories · arXiv:2502.02454
-
IncepFormerNet: A multi-scale multi-head attention network for SSVEP classification 4 Feb 2025 · 1 repository · arXiv:2502.13972
-
LLMER: Crafting Interactive Extended Reality Worlds with JSON Data Generated by Large Language Models 4 Feb 2025 · 1 repository · arXiv:2502.02441
-
LV-XAttn: Distributed Cross-Attention for Long Visual Inputs in Multimodal Large Language Models 4 Feb 2025 · 0 repositories · arXiv:2502.02406
-
Mass-Editing Memory with Attention in Transformers: A cross-lingual exploration of knowledge 4 Feb 2025 · 1 repository · arXiv:2502.02173
-
MATCNN: Infrared and Visible Image Fusion Method Based on Multi-scale CNN with Attention Transformer 4 Feb 2025 · 1 repository · arXiv:2502.01959
-
Memory Efficient Transformer Adapter for Dense Predictions 4 Feb 2025 · 0 repositories · arXiv:2502.01962
-
Mind the Gap: Evaluating Patch Embeddings from General-Purpose and Histopathology Foundation Models for Cell Segmentation and Classification 4 Feb 2025 · 1 repository · arXiv:2502.02471
-
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration 4 Feb 2025 · 0 repositories · arXiv:2502.01969
-
MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm 4 Feb 2025 · 0 repositories · arXiv:2502.02358
-
On the Emergence of Position Bias in Transformers 4 Feb 2025 · 0 repositories · arXiv:2502.01951
-
One Diffusion Step to Real-World Super-Resolution via Flow Trajectory Distillation 4 Feb 2025 · 1 repository · arXiv:2502.01993
-
Open Foundation Models in Healthcare: Challenges, Paradoxes, and Opportunities with GenAI Driven Personalized Prescription 4 Feb 2025 · 0 repositories · arXiv:2502.04356
-
OverThink: Slowdown Attacks on Reasoning LLMs 4 Feb 2025 · 1 repository · arXiv:2502.02542
-
PANDAS: Improving Many-shot Jailbreaking via Positive Affirmation, Negative Demonstration, and Adaptive Sampling 4 Feb 2025 · 1 repository · arXiv:2502.01925Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Peri-LN: Revisiting Layer Normalization in the Transformer Architecture 4 Feb 2025 · 0 repositories · arXiv:2502.02732
-
Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation 4 Feb 2025 · 1 repository · arXiv:2502.02464
-
Rethinking Homogeneity of Vision and Text Tokens in Large Vision-and-Language Models 4 Feb 2025 · 0 repositories · arXiv:2502.01906
-
Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign 4 Feb 2025 · 0 repositories · arXiv:2502.02068
-
SAISA: Towards Multimodal Large Language Models with Both Training and Inference Efficiency 4 Feb 2025 · 1 repository · arXiv:2502.02458
-
SimBEV: A Synthetic Multi-Task Multi-Sensor Driving Data Generation Tool and Dataset 4 Feb 2025 · 1 repository · arXiv:2502.01894
-
Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions 4 Feb 2025 · 0 repositories · arXiv:2502.18470
-
Spatio-temporal transformer to support automatic sign language translation 4 Feb 2025 · 0 repositories · arXiv:2502.02587
-
Stable Port-Hamiltonian Neural Networks 4 Feb 2025 · 0 repositories · arXiv:2502.02480
-
The Skin Game: Revolutionizing Standards for AI Dermatology Model Comparison 4 Feb 2025 · 1 repository · arXiv:2502.02500
-
Topic Modeling in Marathi 4 Feb 2025 · 0 repositories · arXiv:2502.02100
-
Transfer Risk Map: Mitigating Pixel-level Negative Transfer in Medical Segmentation 4 Feb 2025 · 0 repositories · arXiv:2502.02340
-
RIE-SenseNet: Riemannian Manifold Embedding of Multi-Source Industrial Sensor Signals for Robust Pattern Recognition 4 Feb 2025 · 0 repositories · arXiv:2502.02428
-
Twilight: Adaptive Attention Sparsity with Hierarchical Top-p Pruning 4 Feb 2025 · 0 repositories · arXiv:2502.02770
-
UniGaze: Towards Universal Gaze Estimation via Large-scale Pre-Training 4 Feb 2025 · 0 repositories · arXiv:2502.02307
-
UNIP: Rethinking Pre-trained Attention Patterns for Infrared Semantic Segmentation 4 Feb 2025 · 1 repository · arXiv:2502.02257Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
VerteNet -- A Multi-Context Hybrid CNN Transformer for Accurate Vertebral Landmark Localization in Lateral Spine DXA Images 4 Feb 2025 · 1 repository · arXiv:2502.02097
-
Wavelet-based Positional Representation for Long Context 4 Feb 2025 · 0 repositories · arXiv:2502.02004
-
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation 3 Feb 2025 · 0 repositories · arXiv:2502.01697
-
Message-Passing GNNs Fail to Approximate Sparse Triangular Factorizations 3 Feb 2025 · 0 repositories · arXiv:2502.01397
-
COVE: COntext and VEracity prediction for out-of-context images 3 Feb 2025 · 2 repositories · arXiv:2502.01194Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Docking-Aware Attention: Dynamic Protein Representations through Molecular Context Integration 3 Feb 2025 · 0 repositories · arXiv:2502.01461
-
Explaining Context Length Scaling and Bounds for Language Models 3 Feb 2025 · 1 repository · arXiv:2502.01481
-
FastKV: KV Cache Compression for Fast Long-Context Processing with Token-Selective Propagation 3 Feb 2025 · 1 repository · arXiv:2502.01068
-
Fine-Tuning Discrete Diffusion Models with Policy Gradient Methods 3 Feb 2025 · 1 repository · arXiv:2502.01384Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
GauCho: Gaussian Distributions with Cholesky Decomposition for Oriented Object Detection 3 Feb 2025 · 0 repositories · arXiv:2502.01565
-
Generating Multi-Image Synthetic Data for Text-to-Image Customization 3 Feb 2025 · 0 repositories · arXiv:2502.01720
-
GFM-RAG: Graph Foundation Model for Retrieval Augmented Generation 3 Feb 2025 · 1 repository · arXiv:2502.01113Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
GNN-DT: Graph Neural Network Enhanced Decision Transformer for Efficient Optimization in Dynamic Environments 3 Feb 2025 · 1 repository · arXiv:2502.01778
-
Hamming Attention Distillation: Binarizing Keys and Queries for Efficient Long-Context Transformers 3 Feb 2025 · 0 repositories · arXiv:2502.01770
-
Harmonic Loss Trains Interpretable AI Models 3 Feb 2025 · 1 repository · arXiv:2502.01628
-
Hybrid Machine Learning Model for Detecting Bangla Smishing Text Using BERT and Character-Level CNN 3 Feb 2025 · 0 repositories · arXiv:2502.01518
-
Joint Localization and Activation Editing for Low-Resource Fine-Tuning 3 Feb 2025 · 1 repository · arXiv:2502.01179
-
Knowing When to Stop: Dynamic Context Cutoff for Large Language Models 3 Feb 2025 · 0 repositories · arXiv:2502.01025
-
Polynomial, trigonometric, and tropical activations 3 Feb 2025 · 1 repository · arXiv:2502.01247
-
Massive Values in Self-Attention Modules are the Key to Contextual Knowledge Understanding 3 Feb 2025 · 1 repository · arXiv:2502.01563Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Meursault as a Data Point 3 Feb 2025 · 0 repositories · arXiv:2502.01364
-
Molecular Odor Prediction Based on Multi-Feature Graph Attention Networks 3 Feb 2025 · 0 repositories · arXiv:2502.01430
-
Multimodal Inverse Attention Network with Intrinsic Discriminant Feature Exploitation for Fake News Detection 3 Feb 2025 · 0 repositories · arXiv:2502.01699
-
Partial Channel Network: Compute Fewer, Perform Better 3 Feb 2025 · 1 repository · arXiv:2502.01303
-
Preference Leakage: A Contamination Problem in LLM-as-a-judge 3 Feb 2025 · 1 repository · arXiv:2502.01534Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Scalable Language Models with Posterior Inference of Latent Thought Vectors 3 Feb 2025 · 0 repositories · arXiv:2502.01567
-
Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity 3 Feb 2025 · 0 repositories · arXiv:2502.01776
-
SPFFNet: Strip Perception and Feature Fusion Spatial Pyramid Pooling for Fabric Defect Detection 3 Feb 2025 · 0 repositories · arXiv:2502.01445