Methods › General › Output Functions › Softmax › Papers, page 21
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 21 of 375: papers 2,001 to 2,100 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Lessons from Deploying Learning-based CSI Localization on a Large-Scale ISAC Platform 24 Apr 2025 · 0 repositories · arXiv:2504.17173
-
Masked strategies for images with small objects 24 Apr 2025 · 0 repositories · arXiv:2504.17935
-
Optimism, Expectation, or Sarcasm? Multi-Class Hope Speech Detection in Spanish and English 24 Apr 2025 · 0 repositories · arXiv:2504.17974
-
polyGen: A Learning Framework for Atomic-level Polymer Structure Generation 24 Apr 2025 · 0 repositories · arXiv:2504.17656
-
Quadratic Interest Network for Multimodal Click-Through Rate Prediction 24 Apr 2025 · 1 repository · arXiv:2504.17699
-
The Sparse Frontier: Sparse Attention Trade-offs in Transformer LLMs 24 Apr 2025 · 0 repositories · arXiv:2504.17768
-
Token Sequence Compression for Efficient Multimodal Computing 24 Apr 2025 · 0 repositories · arXiv:2504.17892
-
Token-Shuffle: Towards High-Resolution Image Generation with Autoregressive Models 24 Apr 2025 · 0 repositories · arXiv:2504.17789
-
Towards Generalizable Deepfake Detection with Spatial-Frequency Collaborative Learning and Hierarchical Cross-Modal Fusion 24 Apr 2025 · 0 repositories · arXiv:2504.17223
-
Unsupervised EEG-based decoding of absolute auditory attention with canonical correlation analysis 24 Apr 2025 · 0 repositories · arXiv:2504.17724
-
Unveiling the Hidden: Movie Genre and User Bias in Spoiler Detection 24 Apr 2025 · 1 repository · arXiv:2504.17834
-
A Few-Shot Metric Learning Method with Dual-Channel Attention for Cross-Modal Same-Neuron Identification 23 Apr 2025 · 0 repositories · arXiv:2504.16520
-
A Novel Graph Transformer Framework for Gene Regulatory Network Inference 23 Apr 2025 · 0 repositories · arXiv:2504.16961
-
A Novel Hybrid Approach Using an Attention-Based Transformer + GRU Model for Predicting Cryptocurrency Prices 23 Apr 2025 · 0 repositories · arXiv:2504.17079
-
A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms 23 Apr 2025 · 0 repositories · arXiv:2504.16420
-
Advanced Chest X-Ray Analysis via Transformer-Based Image Descriptors and Cross-Model Attention Mechanism 23 Apr 2025 · 0 repositories · arXiv:2504.16774
-
Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate 23 Apr 2025 · 0 repositories · arXiv:2504.16489
-
Do Large Language Models know who did what to whom? 23 Apr 2025 · 0 repositories · arXiv:2504.16884
-
DP2FL: Dual Prompt Personalized Federated Learning in Foundation Models 23 Apr 2025 · 0 repositories · arXiv:2504.16357
-
DyMU: Dynamic Merging and Virtual Unmerging for Efficient VLMs 23 Apr 2025 · 0 repositories · arXiv:2504.17040
-
ECGDeDRDNet: A deep learning-based method for Electrocardiogram noise removal using a double recurrent dense network 23 Apr 2025 · 0 repositories · arXiv:2505.05477
-
Emo Pillars: Knowledge Distillation to Support Fine-Grained Context-Aware and Context-Less Emotion Classification 23 Apr 2025 · 0 repositories · arXiv:2504.16856
-
Frequency-Compensated Network for Daily Arctic Sea Ice Concentration Prediction 23 Apr 2025 · 1 repository · arXiv:2504.16745
-
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing 23 Apr 2025 · 0 repositories · arXiv:2504.16433
-
From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories 23 Apr 2025 · 0 repositories · arXiv:2504.16449
-
Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light 23 Apr 2025 · 1 repository · arXiv:2504.16922
-
How Effective are Generative Large Language Models in Performing Requirements Classification? 23 Apr 2025 · 0 repositories · arXiv:2504.16768
-
Learning Underwater Active Perception in Simulation 23 Apr 2025 · 0 repositories · arXiv:2504.17817
-
Leveraging LLMs as Meta-Judges: A Multi-Agent Framework for Evaluating LLM Judgments 23 Apr 2025 · 0 repositories · arXiv:2504.17087
-
Simplified Swarm Learning Framework for Robust and Scalable Diagnostic Services in Cancer Histopathology 23 Apr 2025 · 0 repositories · arXiv:2504.16732
-
Transformers for Complex Query Answering over Knowledge Hypergraphs 23 Apr 2025 · 0 repositories · arXiv:2504.16537
-
Synergistic Benefits of Joint Molecule Generation and Property Prediction 23 Apr 2025 · 0 repositories · arXiv:2504.16559
-
ZipR1: Reinforcing Token Sparsity in MLLMs 23 Apr 2025 · 0 repositories · arXiv:2504.18579
-
A Large-scale Class-level Benchmark Dataset for Code Generation with LLMs 22 Apr 2025 · 0 repositories · arXiv:2504.15564
-
A Non-Invasive Load Monitoring Method for Edge Computing Based on MobileNetV3 and Dynamic Time Regulation 22 Apr 2025 · 0 repositories · arXiv:2504.16142
-
Advancing Embodied Agent Security: From Safety Benchmarks to Input Moderation 22 Apr 2025 · 0 repositories · arXiv:2504.15699
-
Automated Bug Report Prioritization in Large Open-Source Projects 22 Apr 2025 · 1 repository · arXiv:2504.15912
-
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 22 Apr 2025 · 0 repositories · arXiv:2504.16027
-
CiteFix: Enhancing RAG Accuracy Through Post-Processing Citation Correction 22 Apr 2025 · 0 repositories · arXiv:2504.15629
-
COBRA: Algorithm-Architecture Co-optimized Binary Transformer Accelerator for Edge Inference 22 Apr 2025 · 0 repositories · arXiv:2504.16269
-
Comparative Analysis of Evolutionary Algorithms for Energy-Aware Production Scheduling 22 Apr 2025 · 0 repositories · arXiv:2504.15672
-
Comprehensive Evaluation of Quantitative Measurements from Automated Deep Segmentations of PSMA PET/CT Images 22 Apr 2025 · 0 repositories · arXiv:2504.16237
-
DINOv2-powered Few-Shot Semantic Segmentation: A Unified Framework via Cross-Model Distillation and 4D Correlation Mining 22 Apr 2025 · 0 repositories · arXiv:2504.15669
-
DiTPainter: Efficient Video Inpainting with Diffusion Transformers 22 Apr 2025 · 0 repositories · arXiv:2504.15661
-
DSDNet: Raw Domain Demoiréing via Dual Color-Space Synergy 22 Apr 2025 · 0 repositories · arXiv:2504.15756
-
FADEL: Uncertainty-aware Fake Audio Detection with Evidential Deep Learning 22 Apr 2025 · 0 repositories · arXiv:2504.15663
-
Few-shot Hate Speech Detection Based on the MindSpore Framework 22 Apr 2025 · 0 repositories · arXiv:2504.15987
-
FinDER: Financial Dataset for Question Answering and Evaluating Retrieval-Augmented Generation 22 Apr 2025 · 0 repositories · arXiv:2504.15800
-
FreeGraftor: Training-Free Cross-Image Feature Grafting for Subject-Driven Text-to-Image Generation 22 Apr 2025 · 1 repository · arXiv:2504.15958
-
Grounded in Context: Retrieval-Based Method for Hallucination Detection 22 Apr 2025 · 0 repositories · arXiv:2504.15771
-
How Private is Your Attention? Bridging Privacy with In-Context Learning 22 Apr 2025 · 0 repositories · arXiv:2504.16000
-
Last-layer committee machines for uncertainty estimations of benthic imagery 22 Apr 2025 · 0 repositories · arXiv:2504.16952
-
LongMamba: Enhancing Mamba's Long Context Capabilities via Training-Free Receptive Field Enlargement 22 Apr 2025 · 1 repository · arXiv:2504.16053
-
MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention 22 Apr 2025 · 1 repository · arXiv:2504.16083
-
Over-the-Air Transmission of Zak-OTFS with Spread Pilots on Sub-THz Communications Testbed 22 Apr 2025 · 0 repositories · arXiv:2504.15947
-
Quantum Doubly Stochastic Transformers 22 Apr 2025 · 0 repositories · arXiv:2504.16275Syntology 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Recent Advances and Future Directions in Extended Reality (XR): Exploring AI-Powered Spatial Intelligence 22 Apr 2025 · 0 repositories · arXiv:2504.15970
-
Research on Cloud Platform Network Traffic Monitoring and Anomaly Detection System based on Large Language Models 22 Apr 2025 · 0 repositories · arXiv:2504.17807
-
Sentiment Analysis in Software Engineering: Evaluating Generative Pre-trained Transformers 22 Apr 2025 · 0 repositories · arXiv:2505.14692
-
SUPRA: Subspace Parameterized Attention for Neural Operator on General Domains 22 Apr 2025 · 0 repositories · arXiv:2504.15897
-
Synergizing RAG and Reasoning: A Systematic Review 22 Apr 2025 · 0 repositories · arXiv:2504.15909
-
TeLLMe: An Energy-Efficient Ternary LLM Accelerator for Prefilling and Decoding on Edge FPGAs 22 Apr 2025 · 0 repositories · arXiv:2504.16266
-
The Viability of Crowdsourcing for RAG Evaluation 22 Apr 2025 · 1 repository · arXiv:2504.15689
-
Universal Approximation with Softmax Attention 22 Apr 2025 · 1 repository · arXiv:2504.15956Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Call for New Recipes to Enhance Spatial Reasoning in MLLMs 21 Apr 2025 · 0 repositories · arXiv:2504.15037
-
A Deep Learning Framework for Sequence Mining with Bidirectional LSTM and Multi-Scale Attention 21 Apr 2025 · 0 repositories · arXiv:2504.15223
-
Acquire and then Adapt: Squeezing out Text-to-Image Model for Image Restoration 21 Apr 2025 · 0 repositories · arXiv:2504.15159
-
AlignRAG: Leveraging Critique Learning for Evidence-Sensitive Retrieval-Augmented Reasoning 21 Apr 2025 · 1 repository · arXiv:2504.14858Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
An Efficient Aerial Image Detection with Variable Receptive Fields 21 Apr 2025 · 0 repositories · arXiv:2504.15165
-
An LMM for Efficient Video Understanding via Reinforced Compression of Video Cubes 21 Apr 2025 · 0 repositories · arXiv:2504.15270
-
Automated Measurement of Eczema Severity with Self-Supervised Learning 21 Apr 2025 · 0 repositories · arXiv:2504.15193
-
Backdoor Defense in Diffusion Models via Spatial Attention Unlearning 21 Apr 2025 · 0 repositories · arXiv:2504.18563
-
Distribution-aware Dataset Distillation for Efficient Image Restoration 21 Apr 2025 · 0 repositories · arXiv:2504.14826
-
DyST-XL: Dynamic Layout Planning and Content Control for Compositional Text-to-Video Generation 21 Apr 2025 · 0 repositories · arXiv:2504.15032
-
ECViT: Efficient Convolutional Vision Transformer with Local-Attention and Multi-scale Stages 21 Apr 2025 · 1 repository · arXiv:2504.14825
-
Efficient Document Retrieval with G-Retriever 21 Apr 2025 · 1 repository · arXiv:2504.14955
-
Efficient Pretraining Length Scaling 21 Apr 2025 · 0 repositories · arXiv:2504.14992
-
Hierarchical Attention Fusion of Visual and Textual Representations for Cross-Domain Sequential Recommendation 21 Apr 2025 · 0 repositories · arXiv:2504.15085
-
Impact of Latent Space Dimension on IoT Botnet Detection Performance: VAE-Encoder Versus ViT-Encoder 21 Apr 2025 · 0 repositories · arXiv:2504.14879
-
Improving Sound Source Localization with Joint Slot Attention on Image and Audio 21 Apr 2025 · 0 repositories · arXiv:2504.15118
-
Insert Anything: Image Insertion via In-Context Editing in DiT 21 Apr 2025 · 0 repositories · arXiv:2504.15009
-
Integrating Response Time and Attention Duration in Bayesian Preference Learning for Multiple Criteria Decision Aiding 21 Apr 2025 · 0 repositories · arXiv:2504.14938
-
KeyDiff: Key Similarity-Based KV Cache Eviction for Long-Context LLM Inference in Resource-Constrained Environments 21 Apr 2025 · 0 repositories · arXiv:2504.15364
-
Leveraging Language Models for Automated Patient Record Linkage 21 Apr 2025 · 0 repositories · arXiv:2504.15261
-
LLMs as Data Annotators: How Close Are We to Human Performance 21 Apr 2025 · 0 repositories · arXiv:2504.15022
-
Mitigating Degree Bias in Graph Representation Learning with Learnable Structural Augmentation and Structural Self-Attention 21 Apr 2025 · 1 repository · arXiv:2504.15075
-
MoE Parallel Folding: Heterogeneous Parallelism Mappings for Efficient Large-Scale MoE Model Training with Megatron Core 21 Apr 2025 · 0 repositories · arXiv:2504.14960
-
POLYRAG: Integrating Polyviews into Retrieval-Augmented Generation for Medical Applications 21 Apr 2025 · 0 repositories · arXiv:2504.14917
-
Retrieval Augmented Generation Evaluation in the Era of Large Language Models: A Comprehensive Survey 21 Apr 2025 · 1 repository · arXiv:2504.14891
-
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction 21 Apr 2025 · 1 repository · arXiv:2504.15266
-
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images 21 Apr 2025 · 0 repositories · arXiv:2504.15007
-
Structure-guided Diffusion Transformer for Low-Light Image Enhancement 21 Apr 2025 · 0 repositories · arXiv:2504.15054
-
Support Evaluation for the TREC 2024 RAG Track: Comparing Human versus LLM Judges 21 Apr 2025 · 0 repositories · arXiv:2504.15205
-
The 1st EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval 21 Apr 2025 · 0 repositories · arXiv:2504.14788
-
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models 21 Apr 2025 · 0 repositories · arXiv:2504.15068
-
The Synthetic Imputation Approach: Generating Optimal Synthetic Texts For Underrepresented Categories In Supervised Classification Tasks 21 Apr 2025 · 0 repositories · arXiv:2504.15160
-
Vision6D: 3D-to-2D Interactive Visualization and Annotation Tool for 6D Pose Estimation 21 Apr 2025 · 2 repositories · arXiv:2504.15329
-
What Lurks Within? Concept Auditing for Shared Diffusion Models at Scale 21 Apr 2025 · 0 repositories · arXiv:2504.14815
-
WMKA-Net: A Weighted Multi-Kernel Attention NetworkMethod for Retinal Vessel Segmentation 21 Apr 2025 · 0 repositories · arXiv:2504.14888
-
A Framework for Benchmarking and Aligning Task-Planning Safety in LLM-Based Embodied Agents 20 Apr 2025 · 0 repositories · arXiv:2504.14650