Methods › General › Output Functions › Softmax › Papers, page 71
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 71 of 375: papers 7,001 to 7,100 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RAP-SR: RestorAtion Prior Enhancement in Diffusion Models for Realistic Image Super-Resolution 10 Dec 2024 · 1 repository · arXiv:2412.07149
-
Retaining and Enhancing Pre-trained Knowledge in Vision-Language Models with Prompt Ensembling 10 Dec 2024 · 0 repositories · arXiv:2412.07077
-
Rethinking Emotion Annotations in the Era of Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07906
-
STIV: Scalable Text and Image Conditioned Video Generation 10 Dec 2024 · 0 repositories · arXiv:2412.07730
-
StoryWeaver: A Unified World Model for Knowledge-Enhanced Story Character Customization 10 Dec 2024 · 1 repository · arXiv:2412.07375
-
Streaming Private Continual Counting via Binning 10 Dec 2024 · 1 repository · arXiv:2412.07093Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Superficial Consciousness Hypothesis for Autoregressive Transformers 10 Dec 2024 · 1 repository · arXiv:2412.07278
-
SurvBETA: Ensemble-Based Survival Models Using Beran Estimators and Several Attention Mechanisms 10 Dec 2024 · 1 repository · arXiv:2412.07638
-
Towards Automated Cross-domain Exploratory Data Analysis through Large Language Models 10 Dec 2024 · 2 repositories · arXiv:2412.07214
-
Towards Predictive Communication with Brain-Computer Interfaces integrating Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07355
-
TTVD: Towards a Geometric Framework for Test-Time Adaptation Based on Voronoi Diagram 10 Dec 2024 · 0 repositories · arXiv:2412.07980
-
Untapped Potential in Self-Optimization of Hopfield Networks: The Creativity of Unsupervised Learning 10 Dec 2024 · 0 repositories · arXiv:2501.04007
-
Video Motion Transfer with Diffusion Transformers 10 Dec 2024 · 1 repository · arXiv:2412.07776
-
3D Graph Attention Networks for High Fidelity Pediatric Glioma Segmentation 9 Dec 2024 · 0 repositories · arXiv:2412.06743
-
Advancing Extended Reality with 3D Gaussian Splatting: Innovations and Prospects 9 Dec 2024 · 0 repositories · arXiv:2412.06257
-
AgentAlign: Misalignment-Adapted Multi-Agent Perception for Resilient Inter-Agent Sensor Correlations 9 Dec 2024 · 0 repositories · arXiv:2412.06142
-
Analysing Public Transport User Sentiment on Low Resource Multilingual Data 9 Dec 2024 · 0 repositories · arXiv:2412.06951
-
Anchoring Bias in Large Language Models: An Experimental Study 9 Dec 2024 · 0 repositories · arXiv:2412.06593
-
AnomalyControl: Learning Cross-modal Semantic Features for Controllable Anomaly Synthesis 9 Dec 2024 · 0 repositories · arXiv:2412.06510
-
ASGDiffusion: Parallel High-Resolution Generation with Asynchronous Structure Guidance 9 Dec 2024 · 0 repositories · arXiv:2412.06163
-
Attention-Enhanced Lightweight Hourglass Network for Human Pose Estimation 9 Dec 2024 · 1 repository · arXiv:2412.06227
-
BatchTopK Sparse Autoencoders 9 Dec 2024 · 2 repositories · arXiv:2412.06410Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Bridging the Divide: Reconsidering Softmax and Linear Attention 9 Dec 2024 · 1 repository · arXiv:2412.06590Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 22 pointer-only (licence)
-
Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness 9 Dec 2024 · 0 repositories · arXiv:2412.07022
-
Edge Delayed Deep Deterministic Policy Gradient: efficient continuous control for edge scenarios 9 Dec 2024 · 0 repositories · arXiv:2412.06390
-
Efficient user history modeling with amortized inference for deep learning recommendation models 9 Dec 2024 · 0 repositories · arXiv:2412.06924
-
EMOv2: Pushing 5M Vision Model Frontier 9 Dec 2024 · 1 repository · arXiv:2412.06674
-
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit 9 Dec 2024 · 0 repositories · arXiv:2412.06370
-
Flexible and Scalable Deep Dendritic Spiking Neural Networks with Multiple Nonlinear Branching 9 Dec 2024 · 0 repositories · arXiv:2412.06355
-
Gated Delta Networks: Improving Mamba2 with Delta Rule 9 Dec 2024 · 4 repositories · arXiv:2412.06464Syntology official (archive's flag): 5 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 7 pointer-only (licence)
-
HES-UNet: A U-Net for Hepatic Echinococcosis Lesion Segmentation 9 Dec 2024 · 0 repositories · arXiv:2412.06530
-
HSDA: High-frequency Shuffle Data Augmentation for Bird's-Eye-View Map Segmentation 9 Dec 2024 · 1 repository · arXiv:2412.06127
-
HYATT-Net is Grand: A Hybrid Attention Network for Performant Anatomical Landmark Detection 9 Dec 2024 · 1 repository · arXiv:2412.06499
-
InstantRestore: Single-Step Personalized Face Restoration with Shared-Image Attention 9 Dec 2024 · 0 repositories · arXiv:2412.06753
-
Inverting Transformer-based Vision Models 9 Dec 2024 · 2 repositories · arXiv:2412.06534
-
Knowledge Transfer and Domain Adaptation for Fine-Grained Remote Sensing Image Segmentation 9 Dec 2024 · 1 repository · arXiv:2412.06664
-
LLM as HPC Expert: Extending RAG Architecture for HPC Data 9 Dec 2024 · 0 repositories · arXiv:2501.14733
-
LLM-BIP: Structured Pruning for Large Language Models with Block-Wise Forward Importance Propagation 9 Dec 2024 · 0 repositories · arXiv:2412.06419
-
Local Attention Transformers for High-Detail Optical Flow Upsampling 9 Dec 2024 · 0 repositories · arXiv:2412.06439
-
Normalizing Flows are Capable Generative Models 9 Dec 2024 · 3 repositories · arXiv:2412.06329Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Open-Vocabulary High-Resolution 3D (OVHR3D) Data Segmentation and Annotation Framework 9 Dec 2024 · 0 repositories · arXiv:2412.06268
-
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06249
-
Ranked from Within: Ranking Large Multimodal Models for Visual Question Answering Without Labels 9 Dec 2024 · 0 repositories · arXiv:2412.06461
-
Reconfigurable Holographic Surface-aided Distributed MIMO Radar Systems 9 Dec 2024 · 0 repositories · arXiv:2412.06279
-
S²FT: Efficient, Scalable and Generalizable LLM Fine-tuning by Structured Sparsity 9 Dec 2024 · 0 repositories · arXiv:2412.06289
-
SiReRAG: Indexing Similar and Related Information for Multihop Reasoning 9 Dec 2024 · 0 repositories · arXiv:2412.06206Syntology 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
SparseAccelerate: Efficient Long-Context Inference for Mid-Range GPUs 9 Dec 2024 · 0 repositories · arXiv:2412.06198
-
Splatter-360: Generalizable 360° Gaussian Splatting for Wide-baseline Panoramic Images 9 Dec 2024 · 1 repository · arXiv:2412.06250
-
Static Key Attention in Vision 9 Dec 2024 · 0 repositories · arXiv:2412.07049
-
Stock Type Prediction Model Based on Hierarchical Graph Neural Network 9 Dec 2024 · 0 repositories · arXiv:2412.06862
-
The Computational Limits of State-Space Models and Mamba via the Lens of Circuit Complexity 9 Dec 2024 · 0 repositories · arXiv:2412.06148
-
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.17821
-
Toward Non-Invasive Diagnosis of Bankart Lesions with Deep Learning 9 Dec 2024 · 1 repository · arXiv:2412.06717
-
UAV Virtual Antenna Array Deployment for Uplink Interference Mitigation in Data Collection Networks 9 Dec 2024 · 0 repositories · arXiv:2412.06456
-
Understanding Factual Recall in Transformers via Associative Memories 9 Dec 2024 · 0 repositories · arXiv:2412.06538
-
UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts 9 Dec 2024 · 0 repositories · arXiv:2412.06340
-
Unseen Attack Detection in Software-Defined Networking Using a BERT-Based Large Language Model 9 Dec 2024 · 0 repositories · arXiv:2412.06239
-
VP-MEL: Visual Prompts Guided Multimodal Entity Linking 9 Dec 2024 · 0 repositories · arXiv:2412.06720
-
VQ4ALL: Efficient Neural Network Representation via a Universal Codebook 9 Dec 2024 · 0 repositories · arXiv:2412.06875
-
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06292
-
A Collaborative Multi-Agent Approach to Retrieval-Augmented Generation Across Diverse Data 8 Dec 2024 · 0 repositories · arXiv:2412.05838
-
A Two-Stage AI-Powered Motif Mining Method for Efficient Power System Topological Analysis 8 Dec 2024 · 0 repositories · arXiv:2412.05957
-
A4-Unet: Deformable Multi-Scale Attention Network for Brain Tumor Segmentation 8 Dec 2024 · 1 repository · arXiv:2412.06088
-
Are Clinical T5 Models Better for Clinical Text? 8 Dec 2024 · 1 repository · arXiv:2412.05845
-
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs 8 Dec 2024 · 1 repository · arXiv:2412.05819Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Curse of Attention: A Kernel-Based Perspective for Why Transformers Fail to Generalize on Time Series Forecasting and Beyond 8 Dec 2024 · 0 repositories · arXiv:2412.06061
-
Enhanced Computationally Efficient Long LoRA Inspired Perceiver Architectures for Auto-Regressive Language Modeling 8 Dec 2024 · 0 repositories · arXiv:2412.06106
-
Enhancing Content Representation for AR Image Quality Assessment Using Knowledge Distillation 8 Dec 2024 · 0 repositories · arXiv:2412.06003
-
Evaluating Robustness of LLMs on Crisis-Related Microblogs across Events, Information Types, and Linguistic Features 8 Dec 2024 · 0 repositories · arXiv:2412.10413
-
Fully Open Source Moxin-7B Technical Report 8 Dec 2024 · 1 repository · arXiv:2412.06845
-
GBR: Generative Bundle Refinement for High-fidelity Gaussian Splatting and Meshing 8 Dec 2024 · 0 repositories · arXiv:2412.05908
-
Heuristic-Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models 8 Dec 2024 · 1 repository · arXiv:2412.05934
-
Imputation Matters: A Deeper Look into an Overlooked Step in Longitudinal Health and Behavior Sensing Research 8 Dec 2024 · 0 repositories · arXiv:2412.06018
-
KITE-DDI: A Knowledge graph Integrated Transformer Model for accurately predicting Drug-Drug Interaction Events from Drug SMILES and Biomedical Knowledge Graph 8 Dec 2024 · 0 repositories · arXiv:2412.05770
-
Language-Guided Image Tokenization for Generation 8 Dec 2024 · 0 repositories · arXiv:2412.05796
-
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor 8 Dec 2024 · 1 repository · arXiv:2412.07801
-
Lightweight Spatial Embedding for Vision-based 3D Occupancy Prediction 8 Dec 2024 · 0 repositories · arXiv:2412.05976
-
LVS-Net: A Lightweight Vessels Segmentation Network for Retinal Image Analysis 8 Dec 2024 · 0 repositories · arXiv:2412.05968
-
M³-20M: A Large-Scale Multi-Modal Molecule Dataset for AI-driven Drug Design and Discovery 8 Dec 2024 · 1 repository · arXiv:2412.06847
-
Mixture-of-PageRanks: Replacing Long-Context with Real-Time, Sparse GraphRAG 8 Dec 2024 · 0 repositories · arXiv:2412.06078
-
Paddy Disease Detection and Classification Using Computer Vision Techniques: A Mobile Application to Detect Paddy Disease 8 Dec 2024 · 0 repositories · arXiv:2412.05996
-
Vision Transformer-based Semantic Communications With Importance-Aware Quantization 8 Dec 2024 · 0 repositories · arXiv:2412.06038
-
A Comparative Study on Code Generation with Transformers 7 Dec 2024 · 0 repositories · arXiv:2412.05749
-
BERTCaps: BERT Capsule for Persian Multi-Domain Sentiment Analysis 7 Dec 2024 · 0 repositories · arXiv:2412.05591
-
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds 7 Dec 2024 · 1 repository · arXiv:2412.05631Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
Flex Attention: A Programming Model for Generating Optimized Attention Kernels 7 Dec 2024 · 0 repositories · arXiv:2412.05496
-
GAF-FusionNet: Multimodal ECG Analysis via Gramian Angular Fields and Split Attention 7 Dec 2024 · 1 repository · arXiv:2501.01960
-
Innovative Sentiment Analysis and Prediction of Stock Price Using FinBERT, GPT-4 and Logistic Regression: A Data-Driven Approach 7 Dec 2024 · 0 repositories · arXiv:2412.06837
-
Integrating YOLO11 and Convolution Block Attention Module for Multi-Season Segmentation of Tree Trunks and Branches in Commercial Apple Orchards 7 Dec 2024 · 0 repositories · arXiv:2412.05728
-
Jointly RS Image Deblurring and Super-Resolution with Adjustable-Kernel and Multi-Domain Attention 7 Dec 2024 · 1 repository · arXiv:2412.05696
-
KG-Retriever: Efficient Knowledge Indexing for Retrieval-Augmented Large Language Models 7 Dec 2024 · 1 repository · arXiv:2412.05547Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Learning Soft Driving Constraints from Vectorized Scene Embeddings while Imitating Expert Trajectories 7 Dec 2024 · 0 repositories · arXiv:2412.05717
-
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods 7 Dec 2024 · 1 repository · arXiv:2412.05579
-
M³PC: Test-time Model Predictive Control for Pretrained Masked Trajectory Model 7 Dec 2024 · 1 repository · arXiv:2412.05675Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
Multimodal Biometric Authentication Using Camera-Based PPG and Fingerprint Fusion 7 Dec 2024 · 0 repositories · arXiv:2412.05660
-
On the Expressive Power of Modern Hopfield Networks 7 Dec 2024 · 0 repositories · arXiv:2412.05562
-
PrivAgent: Agentic-based Red-teaming for LLM Privacy Leakage 7 Dec 2024 · 1 repository · arXiv:2412.05734
-
RefSAM3D: Adapting SAM with Cross-modal Reference for 3D Medical Image Segmentation 7 Dec 2024 · 0 repositories · arXiv:2412.05605
-
Shifting NER into High Gear: The Auto-AdvER Approach 7 Dec 2024 · 0 repositories · arXiv:2412.05655