Methods › General › Output Functions › Softmax › Papers, page 54
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 54 of 375: papers 5,301 to 5,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BOLDreams: Dreaming with pruned in-silico fMRI Encoding Models of the Visual Cortex 24 Jan 2025 · 0 repositories · arXiv:2501.14854
-
Causal Graphs Meet Thoughts: Enhancing Complex Reasoning in Graph-Augmented LLMs 24 Jan 2025 · 1 repository · arXiv:2501.14892
-
Chain-of-Retrieval Augmented Generation 24 Jan 2025 · 0 repositories · arXiv:2501.14342
-
Characteristic-Specific Partial Fine-Tuning for Efficient Emotion and Speaker Adaptation in Codec Language Text-to-Speech Models 24 Jan 2025 · 0 repositories · arXiv:2501.14273
-
DarkMind: Latent Chain-of-Thought Backdoor in Customized LLMs 24 Jan 2025 · 0 repositories · arXiv:2501.18617
-
DepressionX: Knowledge Infused Residual Attention for Explainable Depression Severity Assessment 24 Jan 2025 · 1 repository · arXiv:2501.14985
-
Diffusion based Text-to-Music Generation with Global and Local Text based Conditioning 24 Jan 2025 · 0 repositories · arXiv:2501.14680
-
Dynamic Token Reduction during Generation for Vision Language Models 24 Jan 2025 · 0 repositories · arXiv:2501.14204
-
Effective Defect Detection Using Instance Segmentation for NDI 24 Jan 2025 · 0 repositories · arXiv:2501.14149
-
Fast Think-on-Graph: Wider, Deeper and Faster Reasoning of Large Language Model on Knowledge Graph 24 Jan 2025 · 1 repository · arXiv:2501.14300
-
Geometric Mean Improves Loss For Few-Shot Learning 24 Jan 2025 · 0 repositories · arXiv:2501.14593
-
Global Semantic-Guided Sub-image Feature Weight Allocation in High-Resolution Large Vision-Language Models 24 Jan 2025 · 0 repositories · arXiv:2501.14276
-
GraPPI: A Retrieve-Divide-Solve GraphRAG Framework for Large-scale Protein-protein Interaction Exploration 24 Jan 2025 · 1 repository · arXiv:2501.16382
-
HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation 24 Jan 2025 · 1 repository · arXiv:2501.14729Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Idiom Detection in Sorani Kurdish Texts 24 Jan 2025 · 0 repositories · arXiv:2501.14528
-
Iterative Feature Space Optimization through Incremental Adaptive Evaluation 24 Jan 2025 · 0 repositories · arXiv:2501.14889
-
Light3R-SfM: Towards Feed-forward Structure-from-Motion 24 Jan 2025 · 0 repositories · arXiv:2501.14914
-
Low-rank Prompt Interaction for Continual Vision-Language Retrieval 24 Jan 2025 · 1 repository · arXiv:2501.14369
-
Motion-enhancement to Echocardiography Segmentation via Inserting a Temporal Attention Module: An Efficient, Adaptable, and Scalable Approach 24 Jan 2025 · 0 repositories · arXiv:2501.14929
-
On the locality bias and results in the Long Range Arena 24 Jan 2025 · 0 repositories · arXiv:2501.14850
-
Optimizing Human Pose Estimation Through Focused Human and Joint Regions 24 Jan 2025 · 0 repositories · arXiv:2501.14439
-
Post-hoc Spurious Correlation Neutralization with Single-Weight Fictitious Class Unlearning 24 Jan 2025 · 0 repositories · arXiv:2501.14182
-
Predictive Position Estimation for Remote Surgery under Packet Loss Using the Informer Framework 24 Jan 2025 · 0 repositories · arXiv:2501.14664
-
Prompt-Based Cost-Effective Evaluation and Operation of ChatGPT as a Computer Programming Teaching Assistant 24 Jan 2025 · 0 repositories · arXiv:2501.17176
-
Rethinking Table Instruction Tuning 24 Jan 2025 · 1 repository · arXiv:2501.14693
-
Surface Vision Mamba: Leveraging Bidirectional State Space Model for Efficient Spherical Manifold Representation 24 Jan 2025 · 0 repositories · arXiv:2501.14679
-
Test-Time Code-Switching for Cross-lingual Aspect Sentiment Triplet Extraction 24 Jan 2025 · 0 repositories · arXiv:2501.14144
-
TFG-Flow: Training-free Guidance in Multimodal Generative Flow 24 Jan 2025 · 1 repository · arXiv:2501.14216
-
UDiTQC: U-Net-Style Diffusion Transformer for Quantum Circuit Synthesis 24 Jan 2025 · 0 repositories · arXiv:2501.16380
-
UltraLightSqueezeNet: A Deep Learning Architecture for Malaria Classification with up to 54x fewer trainable parameters for resource constrained devices 24 Jan 2025 · 0 repositories · arXiv:2501.14172
-
VarDrop: Enhancing Training Efficiency by Reducing Variate Redundancy in Periodic Time Series Forecasting 24 Jan 2025 · 1 repository · arXiv:2501.14183
-
VideoShield: Regulating Diffusion-based Video Generation Models via Watermarking 24 Jan 2025 · 1 repository · arXiv:2501.14195Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
ZETA: Leveraging Z-order Curves for Efficient Top-k Attention 24 Jan 2025 · 0 repositories · arXiv:2501.14577
-
5G LDPC Linear Transformer for Channel Decoding 23 Jan 2025 · 1 repository · arXiv:2501.14102
-
A Study of the Plausibility of Attention between RNN Encoders in Natural Language Inference 23 Jan 2025 · 0 repositories · arXiv:2501.13735
-
A Transformer-based Autoregressive Decoder Architecture for Hierarchical Text Classification 23 Jan 2025 · 1 repository · arXiv:2501.13598
-
An Efficient Diffusion-based Non-Autoregressive Solver for Traveling Salesman Problem 23 Jan 2025 · 1 repository · arXiv:2501.13767Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
BMG-Q: Localized Bipartite Match Graph Attention Q-Learning for Ride-Pooling Order Dispatch 23 Jan 2025 · 0 repositories · arXiv:2501.13448
-
CAPRAG: A Large Language Model Solution for Customer Service and Automatic Reporting using Vector and Graph Retrieval-Augmented Generation 23 Jan 2025 · 0 repositories · arXiv:2501.13993
-
Chain of Grounded Objectives: Bridging Process and Goal-oriented Prompting for Code Generation 23 Jan 2025 · 0 repositories · arXiv:2501.13978
-
Contrast: A Hybrid Architecture of Transformers and State Space Models for Low-Level Vision 23 Jan 2025 · 0 repositories · arXiv:2501.13353
-
Diffusion-based Perceptual Neural Video Compression with Temporal Diffusion Information Reuse 23 Jan 2025 · 0 repositories · arXiv:2501.13528
-
Document-Level Sentiment Analysis of Urdu Text Using Deep Learning Techniques 23 Jan 2025 · 0 repositories · arXiv:2501.17175
-
EgoHand: Ego-centric Hand Pose Estimation and Gesture Recognition with Head-mounted Millimeter-wave Radar and IMUs 23 Jan 2025 · 1 repository · arXiv:2501.13805
-
Enhanced PEC-YOLO for Detecting Improper Safety Gear Wearing Among Power Line Workers 23 Jan 2025 · 0 repositories · arXiv:2501.13981
-
Enhancing Biomedical Relation Extraction with Directionality 23 Jan 2025 · 1 repository · arXiv:2501.14079
-
Ensuring Medical AI Safety: Explainable AI-Driven Detection and Mitigation of Spurious Model Behavior and Associated Data 23 Jan 2025 · 1 repository · arXiv:2501.13818
-
Eye Gaze as a Signal for Conveying User Attention in Contextual AI Systems 23 Jan 2025 · 0 repositories · arXiv:2501.13878
-
FreEformer: Frequency Enhanced Transformer for Multivariate Time Series Forecasting 23 Jan 2025 · 1 repository · arXiv:2501.13989Syntology official (archive's flag): 7 ran · 7 ran (of which 7 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 7 samples that ran constructed an object rather than computing a result (of 10 harvested samples) · 10 pointer-only (licence)
-
GraphRAG under Fire 23 Jan 2025 · 0 repositories · arXiv:2501.14050
-
Improving Contextual Faithfulness of Large Language Models via Retrieval Heads-Induced Optimization 23 Jan 2025 · 0 repositories · arXiv:2501.13573
-
KAA: Kolmogorov-Arnold Attention for Enhancing Attentive Graph Neural Networks 23 Jan 2025 · 1 repository · arXiv:2501.13456Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Knowledge-Informed Multi-Agent Trajectory Prediction at Signalized Intersections for Infrastructure-to-Everything 23 Jan 2025 · 0 repositories · arXiv:2501.13461
-
LLMs are Vulnerable to Malicious Prompts Disguised as Scientific Language 23 Jan 2025 · 0 repositories · arXiv:2501.14073
-
LLMs Can Plan Only If We Tell Them 23 Jan 2025 · 0 repositories · arXiv:2501.13545
-
M3PT: A Transformer for Multimodal, Multi-Party Social Signal Prediction with Person-aware Blockwise Attention 23 Jan 2025 · 1 repository · arXiv:2501.13416
-
MambaQuant: Quantizing the Mamba Family with Variance Aligned Rotation Methods 23 Jan 2025 · 0 repositories · arXiv:2501.13484
-
ME-CPT: Multi-Task Enhanced Cross-Temporal Point Transformer for Urban 3D Change Detection 23 Jan 2025 · 1 repository · arXiv:2501.14004
-
Multi-Level Attention and Contrastive Learning for Enhanced Text Classification with an Optimized Transformer 23 Jan 2025 · 0 repositories · arXiv:2501.13467
-
On Storage Neural Network Augmented Approximate Nearest Neighbor Search 23 Jan 2025 · 0 repositories · arXiv:2501.16375
-
Polyhedra Encoding Transformers: Enhancing Diffusion MRI Analysis Beyond Voxel and Volumetric Embedding 23 Jan 2025 · 0 repositories · arXiv:2501.13352
-
PromptMono: Cross Prompting Attention for Self-Supervised Monocular Depth Estimation in Challenging Environments 23 Jan 2025 · 0 repositories · arXiv:2501.13796
-
QMamba: Post-Training Quantization for Vision State Space Models 23 Jan 2025 · 0 repositories · arXiv:2501.13624
-
Quantized Spike-driven Transformer 23 Jan 2025 · 1 repository · arXiv:2501.13492Syntology official (archive's flag): 11 ran · 13 ran (of which 9 constructed an object rather than computing a result; 13 with no instrument failure: 2 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 12 unverified (of 25 harvested samples) · 25 pointer-only (licence)
-
Question Answering on Patient Medical Records with Private Fine-Tuned LLMs 23 Jan 2025 · 0 repositories · arXiv:2501.13687
-
RAMQA: A Unified Framework for Retrieval-Augmented Multi-Modal Question Answering 23 Jan 2025 · 1 repository · arXiv:2501.13297
-
Regularizing cross entropy loss via minimum entropy and K-L divergence 23 Jan 2025 · 1 repository · arXiv:2501.13709
-
Retrievals Can Be Detrimental: A Contrastive Backdoor Attack Paradigm on Retrieval-Augmented Diffusion Models 23 Jan 2025 · 0 repositories · arXiv:2501.13340
-
RPO: Retrieval Preference Optimization for Robust Retrieval-Augmented Generation 23 Jan 2025 · 0 repositories · arXiv:2501.13726
-
SAFR: Neuron Redistribution for Interpretability 23 Jan 2025 · 1 repository · arXiv:2501.16374
-
Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models 23 Jan 2025 · 0 repositories · arXiv:2501.13629
-
Skin Disease Detection and Classification of Actinic Keratosis and Psoriasis Utilizing Deep Transfer Learning 23 Jan 2025 · 0 repositories · arXiv:2501.13713
-
Softplus Attention with Re-weighting Boosts Length Extrapolation in Large Language Models 23 Jan 2025 · 0 repositories · arXiv:2501.13428
-
StreamingRAG: Real-time Contextual Retrieval and Generation Framework 23 Jan 2025 · 0 repositories · arXiv:2501.14101
-
Text-driven Online Action Detection 23 Jan 2025 · 1 repository · arXiv:2501.13518
-
Unlearning Clients, Features and Samples in Vertical Federated Learning 23 Jan 2025 · 0 repositories · arXiv:2501.13683
-
Unveiling Discrete Clues: Superior Healthcare Predictions for Rare Diseases 23 Jan 2025 · 1 repository · arXiv:2501.16373
-
Utilizing Evolution Strategies to Train Transformers in Reinforcement Learning 23 Jan 2025 · 1 repository · arXiv:2501.13883Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Novel Scene Coupling Semantic Mask Network for Remote Sensing Image Segmentation 22 Jan 2025 · 1 repository · arXiv:2501.13130Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Adaptive Retrieval Without Self-Knowledge? Bringing Uncertainty Back Home 22 Jan 2025 · 0 repositories · arXiv:2501.12835Syntology 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 5 pointer-only (licence)
-
An Ensemble Model with Attention Based Mechanism for Image Captioning 22 Jan 2025 · 0 repositories · arXiv:2501.14828
-
Applications and Challenges of AI and Microscopy in Life Science Research: A Review 22 Jan 2025 · 0 repositories · arXiv:2501.13135
-
Attention-Driven Hierarchical Reinforcement Learning with Particle Filtering for Source Localization in Dynamic Fields 22 Jan 2025 · 0 repositories · arXiv:2501.13084
-
Computational modelling of biological systems now and then: revisiting tools and visions from the beginning of the century 22 Jan 2025 · 0 repositories · arXiv:2501.13142
-
Dynamics of Toxicity in Political Podcasts 22 Jan 2025 · 0 repositories · arXiv:2501.12640
-
Efficient Prompt Compression with Evaluator Heads for Long-Context Transformer Inference 22 Jan 2025 · 0 repositories · arXiv:2501.12959
-
Ehrenfeucht-Haussler Rank and Chain of Thought 22 Jan 2025 · 0 repositories · arXiv:2501.12997
-
EmoFormer: A Text-Independent Speech Emotion Recognition using a Hybrid Transformer-CNN model 22 Jan 2025 · 0 repositories · arXiv:2501.12682
-
EvidenceMap: Learning Evidence Analysis to Unleash the Power of Small Language Models for Biomedical Question Answering 22 Jan 2025 · 0 repositories · arXiv:2501.12746
-
Explicit Eigenvalue Regularization Improves Sharpness-Aware Minimization 22 Jan 2025 · 1 repository · arXiv:2501.12666
-
Exploring GPT's Ability as a Judge in Music Understanding 22 Jan 2025 · 1 repository · arXiv:2501.13261
-
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana 22 Jan 2025 · 0 repositories · arXiv:2501.12789
-
GRAMA: Adaptive Graph Autoregressive Moving Average Models 22 Jan 2025 · 0 repositories · arXiv:2501.12732
-
Hybridization of Attention UNet with Repeated Atrous Spatial Pyramid Pooling for Improved Brain Tumour Segmentation 22 Jan 2025 · 0 repositories · arXiv:2501.13129
-
Learning Graph Node Embeddings by Smooth Pair Sampling 22 Jan 2025 · 1 repository · arXiv:2501.12884
-
Let SSMs be ConvNets: State-space Modeling with Optimal Tensor Contractions 22 Jan 2025 · 0 repositories · arXiv:2501.13230
-
LiT: Delving into a Simplified Linear Diffusion Transformer for Image Generation 22 Jan 2025 · 0 repositories · arXiv:2501.12976
-
Multi-Instance Partial-Label Learning with Margin Adjustment 22 Jan 2025 · 1 repository · arXiv:2501.12597Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Multimodal AI on Wound Images and Clinical Notes for Home Patient Referral 22 Jan 2025 · 0 repositories · arXiv:2501.13247
-
On Tradeoffs in Learning-Augmented Algorithms 22 Jan 2025 · 0 repositories · arXiv:2501.12770