Methods › General › Output Functions › Softmax › Papers, page 8
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 8 of 375: papers 701 to 800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
NGPU-LM: GPU-Accelerated N-Gram Language Model for Context-Biasing in Greedy ASR Decoding 28 May 2025 · 0 repositories · arXiv:2505.22857
-
Non-convex entropic mean-field optimization via Best Response flow 28 May 2025 · 0 repositories · arXiv:2505.22760
-
Online Fair Division for Personalized 2-Value Instances 28 May 2025 · 0 repositories · arXiv:2505.22174
-
Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation 28 May 2025 · 0 repositories · arXiv:2505.22391
-
PS4PRO: Pixel-to-pixel Supervision for Photorealistic Rendering and Optimization 28 May 2025 · 0 repositories · arXiv:2505.22616
-
RAGPPI: RAG Benchmark for Protein-Protein Interactions in Drug Discovery 28 May 2025 · 1 repository · arXiv:2505.23823
-
Re-ttention: Ultra Sparse Visual Generation via Attention Statistical Reshape 28 May 2025 · 1 repository · arXiv:2505.22918
-
Say What You Mean: Natural Language Access Control with Large Language Models for Internet of Things 28 May 2025 · 0 repositories · arXiv:2505.23835
-
SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels 28 May 2025 · 0 repositories · arXiv:2505.22461
-
SkewRoute: Training-Free LLM Routing for Knowledge Graph Retrieval-Augmented Generation via Score Skewness of Retrieved Context 28 May 2025 · 0 repositories · arXiv:2505.23841
-
SlimLLM: Accurate Structured Pruning for Large Language Models 28 May 2025 · 0 repositories · arXiv:2505.22689
-
SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space Splitting 28 May 2025 · 0 repositories · arXiv:2505.22370
-
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models 28 May 2025 · 0 repositories · arXiv:2505.22271
-
Triple Attention Transformer Architecture for Time-Dependent Concrete Creep Prediction 28 May 2025 · 0 repositories · arXiv:2506.04243
-
UP-SLAM: Adaptively Structured Gaussian SLAM with Uncertainty Prediction in Dynamic Environments 28 May 2025 · 0 repositories · arXiv:2505.22335
-
Update Your Transformer to the Latest Release: Re-Basin of Task Vectors 28 May 2025 · 1 repository · arXiv:2505.22697Syntology official (archive's flag): 1 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 3 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning 28 May 2025 · 1 repository · arXiv:2505.22019Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
When Does Neuroevolution Outcompete Reinforcement Learning in Transfer Learning Tasks? 28 May 2025 · 1 repository · arXiv:2505.22696
-
A domain adaptation neural network for digital twin-supported fault diagnosis 27 May 2025 · 1 repository · arXiv:2505.21046
-
AgriFM: A Multi-source Temporal Remote Sensing Foundation Model for Crop Mapping 27 May 2025 · 1 repository · arXiv:2505.21357
-
BacktrackAgent: Enhancing GUI Agent with Error Detection and Backtracking Mechanism 27 May 2025 · 0 repositories · arXiv:2505.20660
-
Beyond 1D: Vision Transformers and Multichannel Signal Images for PPG-to-ECG Reconstruction 27 May 2025 · 0 repositories · arXiv:2505.21767
-
CNN-Based Channel Map Estimation for Movable Antenna Systems 27 May 2025 · 0 repositories · arXiv:2505.21001
-
Continuous-Time Attention: PDE-Guided Mechanisms for Long-Sequence Transformers 27 May 2025 · 0 repositories · arXiv:2505.20666
-
DeepConvContext: A Multi-Scale Approach to Timeseries Classification in Human Activity Recognition 27 May 2025 · 1 repository · arXiv:2505.20894
-
Diagnosing and Resolving Cloud Platform Instability with Multi-modal RAG LLMs 27 May 2025 · 0 repositories · arXiv:2505.21419
-
DiMoSR: Feature Modulation via Multi-Branch Dilated Convolutions for Efficient Image Super-Resolution 27 May 2025 · 1 repository · arXiv:2505.21262
-
Don't Think Longer, Think Wisely: Optimizing Thinking Dynamics for Large Reasoning Models 27 May 2025 · 0 repositories · arXiv:2505.21765
-
Efficient Leaf Disease Classification and Segmentation using Midpoint Normalization Technique and Attention Mechanism 27 May 2025 · 0 repositories · arXiv:2505.21316
-
Emotion-aware Dual Cross-Attentive Neural Network with Label Fusion for Stance Detection in Misinformative Social Media Content 27 May 2025 · 1 repository · arXiv:2505.23812
-
Explainability of Large Language Models using SMILE: Statistical Model-agnostic Interpretability with Local Explanations 27 May 2025 · 1 repository · arXiv:2505.21657
-
FastFace: Tuning Identity Preservation in Distilled Diffusion via Guidance and Attention 27 May 2025 · 1 repository · arXiv:2505.21144
-
From prosthetic memory to prosthetic denial: Auditing whether large language models are prone to mass atrocity denialism 27 May 2025 · 0 repositories · arXiv:2505.21753
-
HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling 27 May 2025 · 0 repositories · arXiv:2505.20836
-
Hardware-Efficient Attention for Fast Decoding 27 May 2025 · 2 repositories · arXiv:2505.21487
-
HTMNet: A Hybrid Network with Transformer-Mamba Bottleneck Multimodal Fusion for Transparent and Reflective Objects Depth Completion 27 May 2025 · 0 repositories · arXiv:2505.20904
-
Long Context Scaling: Divide and Conquer via Multi-Agent Question-driven Collaboration 27 May 2025 · 0 repositories · arXiv:2505.20625
-
MARS-Bench: A Multi-turn Athletic Real-world Scenario Benchmark for Dialogue Evaluation 27 May 2025 · 0 repositories · arXiv:2505.23810
-
Minute-Long Videos with Dual Parallelisms 27 May 2025 · 1 repository · arXiv:2505.21070
-
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration 27 May 2025 · 0 repositories · arXiv:2505.21472
-
MoE-Gyro: Self-Supervised Over-Range Reconstruction and Denoising for MEMS Gyroscopes 27 May 2025 · 0 repositories · arXiv:2506.06318
-
MoPFormer: Motion-Primitive Transformer for Wearable-Sensor Activity Recognition 27 May 2025 · 0 repositories · arXiv:2505.20744
-
Music's Multimodal Complexity in AVQA: Why We Need More than General Multimodal LLMs 27 May 2025 · 0 repositories · arXiv:2505.20638
-
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning 27 May 2025 · 0 repositories · arXiv:2505.20962
-
Pause Tokens Strictly Increase the Expressivity of Constant-Depth Transformers 27 May 2025 · 0 repositories · arXiv:2505.21024
-
Plug-and-Play Co-Occurring Face Attention for Robust Audio-Visual Speaker Extraction 27 May 2025 · 0 repositories · arXiv:2505.20635
-
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs) 27 May 2025 · 0 repositories · arXiv:2505.21091
-
Privacy-Preserving Chest X-ray Report Generation via Multimodal Federated Learning with ViT and GPT-2 27 May 2025 · 0 repositories · arXiv:2505.21715
-
SageAttention2++: A More Efficient Implementation of SageAttention2 27 May 2025 · 2 repositories · arXiv:2505.21136Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
SOSBENCH: Benchmarking Safety Alignment on Scientific Knowledge 27 May 2025 · 0 repositories · arXiv:2505.21605
-
SpecExtend: A Drop-in Enhancement for Speculative Decoding of Long Sequences 27 May 2025 · 1 repository · arXiv:2505.20776
-
tenSVD algorithm for compression 27 May 2025 · 0 repositories · arXiv:2505.21686
-
Time-Series Learning for Proactive Fault Prediction in Distributed Systems with Deep Neural Structures 27 May 2025 · 0 repositories · arXiv:2505.20705
-
Towards Robust Assessment of Pathological Voices via Combined Low-Level Descriptors and Foundation Model Representations 27 May 2025 · 0 repositories · arXiv:2505.21356
-
Unpaired Image-to-Image Translation for Segmentation and Signal Unmixing 27 May 2025 · 0 repositories · arXiv:2505.20746
-
Absolute Coordinates Make Motion Generation Easy 26 May 2025 · 0 repositories · arXiv:2505.19377
-
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling 26 May 2025 · 0 repositories · arXiv:2505.19931
-
Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing 26 May 2025 · 0 repositories · arXiv:2505.19578
-
AdaTP: Attention-Debiased Token Pruning for Video Large Language Models 26 May 2025 · 0 repositories · arXiv:2505.20100
-
Aggregated Structural Representation with Large Language Models for Human-Centric Layout Generation 26 May 2025 · 0 repositories · arXiv:2505.19554
-
Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning 26 May 2025 · 1 repository · arXiv:2505.19611
-
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection 26 May 2025 · 1 repository · arXiv:2505.19528
-
AMQA: An Adversarial Dataset for Benchmarking Bias of LLMs in Medicine and Healthcare 26 May 2025 · 1 repository · arXiv:2505.19562
-
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents 26 May 2025 · 0 repositories · arXiv:2505.19494
-
Automated evaluation of children's speech fluency for low-resource languages 26 May 2025 · 0 repositories · arXiv:2505.19671
-
Balancing Computation Load and Representation Expressivity in Parallel Hybrid Neural Networks 26 May 2025 · 0 repositories · arXiv:2505.19472
-
Benchmarking Multimodal Knowledge Conflict for Large Multimodal Models 26 May 2025 · 1 repository · arXiv:2505.19509
-
Beyond Simple Concatenation: Fairly Assessing PLM Architectures for Multi-Chain Protein-Protein Interactions Prediction 26 May 2025 · 0 repositories · arXiv:2505.20036
-
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages 26 May 2025 · 0 repositories · arXiv:2505.19851
-
Burst Image Super-Resolution via Multi-Cross Attention Encoding and Multi-Scan State-Space Decoding 26 May 2025 · 0 repositories · arXiv:2505.19668
-
CA3D: Convolutional-Attentional 3D Nets for Efficient Video Activity Recognition on the Edge 26 May 2025 · 0 repositories · arXiv:2505.19928
-
Calibrating Pre-trained Language Classifiers on LLM-generated Noisy Labels via Iterative Refinement 26 May 2025 · 1 repository · arXiv:2505.19675
-
CardioPatternFormer: Pattern-Guided Attention for Interpretable ECG Classification with Transformer Architecture 26 May 2025 · 0 repositories · arXiv:2505.20481
-
Compliance-to-Code: Enhancing Financial Compliance Checking via Code Generation 26 May 2025 · 1 repository · arXiv:2505.19804
-
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language 26 May 2025 · 0 repositories · arXiv:2505.19971
-
Dependency Parsing is More Parameter-Efficient with Normalization 26 May 2025 · 0 repositories · arXiv:2505.20215
-
Detection of Suicidal Risk on Social Media: A Hybrid Model 26 May 2025 · 0 repositories · arXiv:2505.23797
-
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems 26 May 2025 · 0 repositories · arXiv:2505.19847
-
DoctorRAG: Medical RAG Fusing Knowledge with Patient Analogy through Textual Gradients 26 May 2025 · 0 repositories · arXiv:2505.19538
-
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation 26 May 2025 · 0 repositories · arXiv:2505.19774
-
Electrolyzers-HSI: Close-Range Multi-Scene Hyperspectral Imaging Benchmark Dataset 26 May 2025 · 0 repositories · arXiv:2505.20507
-
Emotion Classification In-Context in Spanish 26 May 2025 · 0 repositories · arXiv:2505.20571
-
Equivariant Representation Learning for Symmetry-Aware Inference with Guarantees 26 May 2025 · 0 repositories · arXiv:2505.19809
-
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining 26 May 2025 · 0 repositories · arXiv:2505.19893
-
FlowCut: Rethinking Redundancy via Information Flow for Efficient Vision-Language Models 26 May 2025 · 1 repository · arXiv:2505.19536Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Force Prompting: Video Generation Models Can Learn and Generalize Physics-based Control Signals 26 May 2025 · 0 repositories · arXiv:2505.19386
-
GoLF-NRT: Integrating Global Context and Local Geometry for Few-Shot View Synthesis 26 May 2025 · 1 repository · arXiv:2505.19813Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Grokking ExPLAIND: Unifying Model, Data, and Training Attribution to Study Model Behavior 26 May 2025 · 1 repository · arXiv:2505.20076Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots 26 May 2025 · 1 repository · arXiv:2505.20288
-
Hierarchical Tree Search-based User Lifelong Behavior Modeling on Large Language Model 26 May 2025 · 0 repositories · arXiv:2505.19505
-
How Syntax Specialization Emerges in Language Models 26 May 2025 · 0 repositories · arXiv:2505.19548
-
Improvement Strategies for Few-Shot Learning in OCT Image Classification of Rare Retinal Diseases 26 May 2025 · 0 repositories · arXiv:2505.20149
-
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model 26 May 2025 · 1 repository · arXiv:2505.20007
-
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation 26 May 2025 · 0 repositories · arXiv:2505.20271
-
Inference-time Alignment in Continuous Space 26 May 2025 · 1 repository · arXiv:2505.20081
-
KnowTrace: Bootstrapping Iterative Retrieval-Augmented Generation with Structured Knowledge Tracing 26 May 2025 · 1 repository · arXiv:2505.20245
-
Large Language Models' Reasoning Stalls: An Investigation into the Capabilities of Frontier Models 26 May 2025 · 0 repositories · arXiv:2505.19676
-
LeCoDe: A Benchmark Dataset for Interactive Legal Consultation Dialogue Evaluation 26 May 2025 · 0 repositories · arXiv:2505.19667
-
LlamaSeg: Image Segmentation via Autoregressive Mask Generation 26 May 2025 · 0 repositories · arXiv:2505.19422
-
Long-Context State-Space Video World Models 26 May 2025 · 0 repositories · arXiv:2505.20171