Methods › General › Output Functions › Softmax › Papers, page 65
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 65 of 375: papers 6,401 to 6,500 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Leveraging Convolutional Neural Network-Transformer Synergy for Predictive Modeling in Risk-Based Applications 24 Dec 2024 · 0 repositories · arXiv:2412.18222
-
Leveraging Deep Learning with Multi-Head Attention for Accurate Extraction of Medicine from Handwritten Prescriptions 24 Dec 2024 · 0 repositories · arXiv:2412.18199
-
Molly: Making Large Language Model Agents Solve Python Problem More Logically 24 Dec 2024 · 0 repositories · arXiv:2412.18093
-
Multi-View Fusion Neural Network for Traffic Demand Prediction 24 Dec 2024 · 0 repositories · arXiv:2412.19839
-
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English 24 Dec 2024 · 1 repository · arXiv:2412.18415
-
Multimodal joint prediction of traffic spatial-temporal data with graph sparse attention mechanism and bidirectional temporal convolutional network 24 Dec 2024 · 0 repositories · arXiv:2412.19842
-
Pirates of the RAG: Adaptively Attacking LLMs to Leak Knowledge Bases 24 Dec 2024 · 0 repositories · arXiv:2412.18295
-
Re-assessing ImageNet: How aligned is its single-label assumption with its multi-label nature? 24 Dec 2024 · 0 repositories · arXiv:2412.18409
-
Research on the Proximity Relationships of Psychosomatic Disease Knowledge Graph Modules Extracted by Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18419
-
Sampling Bag of Views for Open-Vocabulary Object Detection 24 Dec 2024 · 0 repositories · arXiv:2412.18273
-
SDM-Car: A Dataset for Small and Dim Moving Vehicles Detection in Satellite Videos 24 Dec 2024 · 1 repository · arXiv:2412.18214
-
Segment-Based Attention Masking for GPTs 24 Dec 2024 · 1 repository · arXiv:2412.18487
-
Semi-supervised Credit Card Fraud Detection via Attribute-Driven Graph Representation 24 Dec 2024 · 2 repositories · arXiv:2412.18287
-
SlimGPT: Layer-wise Structured Pruning for Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18110
-
Smooth-Foley: Creating Continuous Sound for Video-to-Audio Generation Under Semantic Guidance 24 Dec 2024 · 0 repositories · arXiv:2412.18157
-
TAB: Transformer Attention Bottlenecks enable User Intervention and Debugging in Vision-Language Models 24 Dec 2024 · 1 repository · arXiv:2412.18675
-
Tackling the Dynamicity in a Production LLM Serving System with SOTA Optimizations via Hybrid Prefill/Decode/Verify Scheduling on Efficient Meta-kernels 24 Dec 2024 · 0 repositories · arXiv:2412.18106
-
TimelyLLM: Segmented LLM Serving System for Time-sensitive Robotic Applications 24 Dec 2024 · 0 repositories · arXiv:2412.18695
-
Towards understanding how attention mechanism works in deep learning 24 Dec 2024 · 0 repositories · arXiv:2412.18288
-
Underwater Image Restoration via Polymorphic Large Kernel CNNs 24 Dec 2024 · 1 repository · arXiv:2412.18459
-
Unlocking the Hidden Treasures: Enhancing Recommendations with Unlabeled Data 24 Dec 2024 · 1 repository · arXiv:2412.18170
-
Unlocking the Potential of Multiple BERT Models for Bangla Question Answering in NCTB Textbooks 24 Dec 2024 · 0 repositories · arXiv:2412.18440
-
Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach 24 Dec 2024 · 0 repositories · arXiv:2412.18108
-
Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models 24 Dec 2024 · 1 repository · arXiv:2412.18609
-
VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis 24 Dec 2024 · 1 repository · arXiv:2412.18178
-
A Bias-Free Training Paradigm for More General AI-generated Image Detection 23 Dec 2024 · 0 repositories · arXiv:2412.17671
-
A Coalition Game for On-demand Multi-modal 3D Automated Delivery System 23 Dec 2024 · 0 repositories · arXiv:2412.17252
-
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression 23 Dec 2024 · 0 repositories · arXiv:2412.17483
-
A Survey of Query Optimization in Large Language Models 23 Dec 2024 · 0 repositories · arXiv:2412.17558
-
An Intrinsically Explainable Approach to Detecting Vertebral Compression Fractures in CT Scans via Neurosymbolic Modeling 23 Dec 2024 · 0 repositories · arXiv:2412.17258
-
Balanced 3DGS: Gaussian-wise Parallelism Rendering with Fine-Grained Tiling 23 Dec 2024 · 0 repositories · arXiv:2412.17378
-
Cech Complex Generation with Homotopy Equivalence Framework for Myocardial Infarction Diagnosis using Electrocardiogram Signals 23 Dec 2024 · 0 repositories · arXiv:2412.17370
-
CiteBART: Learning to Generate Citations for Local Citation Recommendation 23 Dec 2024 · 1 repository · arXiv:2412.17534Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
Comparative Analysis of Document-Level Embedding Methods for Similarity Scoring on Shakespeare Sonnets and Taylor Swift Lyrics 23 Dec 2024 · 0 repositories · arXiv:2412.17552
-
Comprehensive Multi-Modal Prototypes are Simple and Effective Classifiers for Vast-Vocabulary Object Detection 23 Dec 2024 · 1 repository · arXiv:2412.17800
-
Contemporary implementations of spiking bio-inspired neural networks 23 Dec 2024 · 0 repositories · arXiv:2412.17926
-
DiffFormer: a Differential Spatial-Spectral Transformer for Hyperspectral Image Classification 23 Dec 2024 · 1 repository · arXiv:2412.17350
-
Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders 23 Dec 2024 · 1 repository · arXiv:2412.17808Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
DreamFit: Garment-Centric Human Generation via a Lightweight Anything-Dressing Encoder 23 Dec 2024 · 1 repository · arXiv:2412.17644
-
Edge-AI for Agriculture: Lightweight Vision Models for Disease Detection in Resource-Limited Settings 23 Dec 2024 · 0 repositories · arXiv:2412.18635
-
Efficient fine-tuning methodology of text embedding models for information retrieval: contrastive learning penalty (clp) 23 Dec 2024 · 1 repository · arXiv:2412.17364
-
Enhancing Multi-Text Long Video Generation Consistency without Tuning: Time-Frequency Analysis, Prompt Alignment, and Theory 23 Dec 2024 · 0 repositories · arXiv:2412.17254
-
Fast Gradient Computation for RoPE Attention in Almost Linear Time 23 Dec 2024 · 0 repositories · arXiv:2412.17316
-
Feature Based Methods in Domain Adaptation for Object Detection: A Review Paper 23 Dec 2024 · 0 repositories · arXiv:2412.17325
-
Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization 23 Dec 2024 · 1 repository · arXiv:2412.17739
-
Free-viewpoint Human Animation with Pose-correlated Reference Selection 23 Dec 2024 · 0 repositories · arXiv:2412.17290
-
Friends-MMC: A Dataset for Multi-modal Multi-party Conversation Understanding 23 Dec 2024 · 1 repository · arXiv:2412.17295
-
Guided Real Image Dehazing using YCbCr Color Space 23 Dec 2024 · 1 repository · arXiv:2412.17496
-
HPCNeuroNet: A Neuromorphic Approach Merging SNN Temporal Dynamics with Transformer Attention for FPGA-based Particle Physics 23 Dec 2024 · 0 repositories · arXiv:2412.17571
-
LASE: Learned Adjacency Spectral Embeddings 23 Dec 2024 · 1 repository · arXiv:2412.17734
-
LayerDropBack: A Universally Applicable Approach for Accelerating Training of Deep Networks 23 Dec 2024 · 1 repository · arXiv:2412.18027
-
Learning Dynamic Local Context Representations for Infrared Small Target Detection 23 Dec 2024 · 0 repositories · arXiv:2412.17401
-
MRANet: A Modified Residual Attention Networks for Lung and Colon Cancer Classification 23 Dec 2024 · 0 repositories · arXiv:2412.17700
-
Multi-view Fuzzy Graph Attention Networks for Enhanced Graph Learning 23 Dec 2024 · 0 repositories · arXiv:2412.17271
-
Multimodal Preference Data Synthetic Alignment with Reward Model 23 Dec 2024 · 1 repository · arXiv:2412.17417
-
Personalized Large Vision-Language Models 23 Dec 2024 · 0 repositories · arXiv:2412.17610
-
Predicting Satisfied User and Machine Ratio for Compressed Images: A Unified Approach 23 Dec 2024 · 0 repositories · arXiv:2412.17477
-
QTSeg: A Query Token-Based Architecture for Efficient 2D Medical Image Segmentation 23 Dec 2024 · 1 repository · arXiv:2412.17241
-
Revisiting Multimodal Fusion for 3D Anomaly Detection from an Architectural Perspective 23 Dec 2024 · 1 repository · arXiv:2412.17297
-
STAHGNet: Modeling Hybrid-grained Heterogenous Dependency Efficiently for Traffic Prediction 23 Dec 2024 · 0 repositories · arXiv:2412.17524
-
STeInFormer: Spatial-Temporal Interaction Transformer Architecture for Remote Sensing Change Detection 23 Dec 2024 · 1 repository · arXiv:2412.17247
-
Theoretical Constraints on the Expressive Power of RoPE-based Tensor Attention Transformers 23 Dec 2024 · 0 repositories · arXiv:2412.18040
-
Token Statistics Transformer: Linear-Time Attention via Variational Rate Reduction 23 Dec 2024 · 1 repository · arXiv:2412.17810
-
Towards Unsupervised Model Selection for Domain Adaptive Object Detection 23 Dec 2024 · 1 repository · arXiv:2412.17284Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Uncertainty-Participation Context Consistency Learning for Semi-supervised Semantic Segmentation 23 Dec 2024 · 1 repository · arXiv:2412.17331
-
URoadNet: Dual Sparse Attentive U-Net for Multiscale Road Network Extraction 23 Dec 2024 · 0 repositories · arXiv:2412.17573
-
xPatch: Dual-Stream Time Series Forecasting with Exponential Seasonal-Trend Decomposition 23 Dec 2024 · 1 repository · arXiv:2412.17323
-
A Parameter-Efficient Quantum Anomaly Detection Method on a Superconducting Quantum Processor 22 Dec 2024 · 2 repositories · arXiv:2412.16867
-
A Reality Check on Context Utilisation for Retrieval-Augmented Generation 22 Dec 2024 · 1 repository · arXiv:2412.17031
-
Adapting Image-to-Video Diffusion Models for Large-Motion Frame Interpolation 22 Dec 2024 · 0 repositories · arXiv:2412.17042
-
An OpenMind for 3D medical vision self-supervised learning 22 Dec 2024 · 1 repository · arXiv:2412.17041
-
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models 22 Dec 2024 · 0 repositories · arXiv:2501.03246
-
CoF: Coarse to Fine-Grained Image Understanding for Multi-modal Large Language Models 22 Dec 2024 · 1 repository · arXiv:2412.16869
-
Differentially Private Random Block Coordinate Descent 22 Dec 2024 · 0 repositories · arXiv:2412.17054
-
DR-Encoder: Encode Low-rank Gradients with Random Prior for Large Language Models Differentially Privately 22 Dec 2024 · 0 repositories · arXiv:2412.17053
-
Enhancing Supply Chain Transparency in Emerging Economies Using Online Contents and LLMs 22 Dec 2024 · 0 repositories · arXiv:2412.16922
-
FADA: Fast Diffusion Avatar Synthesis with Mixed-Supervised Multi-CFG Distillation 22 Dec 2024 · 0 repositories · arXiv:2412.16915
-
Interactive Classification Metrics: A graphical application to build robust intuition for classification model evaluation 22 Dec 2024 · 1 repository · arXiv:2412.17066
-
Multi-classification of High-Frequency Oscillations Using iEEG Signals and Deep Learning Models 22 Dec 2024 · 0 repositories · arXiv:2412.17145
-
Multifaceted User Modeling in Recommendation: A Federated Foundation Models Approach 22 Dec 2024 · 1 repository · arXiv:2412.16969
-
NumbOD: A Spatial-Frequency Fusion Attack Against Object Detectors 22 Dec 2024 · 1 repository · arXiv:2412.16955Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora 22 Dec 2024 · 0 repositories · arXiv:2412.16976
-
PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental Health 22 Dec 2024 · 1 repository · arXiv:2412.16882
-
Reconsidering SMT Over NMT for Closely Related Languages: A Case Study of Persian-Hindi Pair 22 Dec 2024 · 0 repositories · arXiv:2412.16877
-
Reversed Attention: On The Gradient Descent Of Attention Layers In GPT 22 Dec 2024 · 1 repository · arXiv:2412.17019
-
Robustness of Large Language Models Against Adversarial Attacks 22 Dec 2024 · 0 repositories · arXiv:2412.17011
-
Semantic Hierarchical Prompt Tuning for Parameter-Efficient Fine-Tuning 22 Dec 2024 · 1 repository · arXiv:2412.16956
-
SubstationAI: Multimodal Large Model-Based Approaches for Analyzing Substation Equipment Faults 22 Dec 2024 · 0 repositories · arXiv:2412.17077
-
Survey on Abstractive Text Summarization: Dataset, Models, and Metrics 22 Dec 2024 · 2 repositories · arXiv:2412.17165
-
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction 22 Dec 2024 · 0 repositories · arXiv:2412.16919
-
AlzheimerRAG: Multimodal Retrieval Augmented Generation for PubMed articles 21 Dec 2024 · 0 repositories · arXiv:2412.16701
-
Anchor Learning with Potential Cluster Constraints for Multi-view Clustering 21 Dec 2024 · 1 repository · arXiv:2412.16519
-
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans? 21 Dec 2024 · 0 repositories · arXiv:2412.16772
-
Attention Entropy is a Key Factor: An Analysis of Parallel Context Encoding with Full-attention-based Pre-trained Language Models 21 Dec 2024 · 0 repositories · arXiv:2412.16545
-
Distilling Large Language Models for Efficient Clinical Information Extraction 21 Dec 2024 · 0 repositories · arXiv:2501.00031
-
Effective and Efficient Representation Learning for Flight Trajectories 21 Dec 2024 · 1 repository · arXiv:2412.16581
-
Effective Context Modeling Framework for Emotion Recognition in Conversations 21 Dec 2024 · 0 repositories · arXiv:2412.16444
-
Enhancing Contrastive Learning Inspired by the Philosophy of "The Blind Men and the Elephant" 21 Dec 2024 · 1 repository · arXiv:2412.16522
-
Enhancing Nighttime Vehicle Detection with Day-to-Night Style Transfer and Labeling-Free Augmentation 21 Dec 2024 · 0 repositories · arXiv:2412.16478
-
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16486