Methods › General › Output Functions › Softmax › Papers, page 51
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 51 of 375: papers 5,001 to 5,100 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TeLL-Drive: Enhancing Autonomous Driving with Teacher LLM-Guided Deep Reinforcement Learning 3 Feb 2025 · 0 repositories · arXiv:2502.01387
-
TFBS-Finder: Deep Learning-based Model with DNABERT and Convolutional Networks to Predict Transcription Factor Binding Sites 3 Feb 2025 · 1 repository · arXiv:2502.01311
-
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models 3 Feb 2025 · 0 repositories · arXiv:2502.01386
-
Toward Neurosymbolic Program Comprehension 3 Feb 2025 · 0 repositories · arXiv:2502.01806
-
Transformers trained on proteins can learn to attend to Euclidean distance 3 Feb 2025 · 1 repository · arXiv:2502.01533
-
VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos 3 Feb 2025 · 1 repository · arXiv:2502.01549
-
VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control 3 Feb 2025 · 1 repository · arXiv:2502.01101Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A method for estimating forest carbon storage distribution density via artificial intelligence generated content model 2 Feb 2025 · 0 repositories · arXiv:2502.00783
-
Attention Sinks and Outlier Features: A 'Catch, Tag, and Release' Mechanism for Embeddings 2 Feb 2025 · 0 repositories · arXiv:2502.00919
-
Decision-informed Neural Networks with Large Language Model Integration for Portfolio Optimization 2 Feb 2025 · 0 repositories · arXiv:2502.00828
-
DeepGate4: Efficient and Effective Representation Learning for Circuit Design at Scale 2 Feb 2025 · 1 repository · arXiv:2502.01681Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis 2 Feb 2025 · 0 repositories · arXiv:2502.00654
-
Estimating forest carbon stocks from high-resolution remote sensing imagery by reducing domain shift with style transfer 2 Feb 2025 · 0 repositories · arXiv:2502.00784
-
Explainability in Practice: A Survey of Explainable NLP Across Various Domains 2 Feb 2025 · 0 repositories · arXiv:2502.00837
-
Fundamental limits of learning in sequence multi-index models and deep attention networks: High-dimensional asymptotics and sharp thresholds 2 Feb 2025 · 1 repository · arXiv:2502.00901
-
HuViDPO:Enhancing Video Generation through Direct Preference Optimization for Human-Centric Alignment 2 Feb 2025 · 0 repositories · arXiv:2502.01690
-
Language Models Use Trigonometry to Do Addition 2 Feb 2025 · 0 repositories · arXiv:2502.00873
-
LIBRA: Measuring Bias of Large Language Model from a Local Context 2 Feb 2025 · 0 repositories · arXiv:2502.01679
-
MINT: Mitigating Hallucinations in Large Vision-Language Models via Token Reduction 2 Feb 2025 · 0 repositories · arXiv:2502.00717
-
scGSDR: Harnessing Gene Semantics for Single-Cell Pharmacological Profiling 2 Feb 2025 · 1 repository · arXiv:2502.01689
-
A framework for river connectivity classification using temporal image processing and attention based neural networks 1 Feb 2025 · 0 repositories · arXiv:2502.00474
-
A Study on the Performance of U-Net Modifications in Retroperitoneal Tumor Segmentation 1 Feb 2025 · 1 repository · arXiv:2502.00314
-
Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset 1 Feb 2025 · 0 repositories · arXiv:2502.01676
-
CoddLLM: Empowering Large Language Models for Data Analytics 1 Feb 2025 · 0 repositories · arXiv:2502.00329
-
Complex Wavelet Mutual Information Loss: A Multi-Scale Loss Function for Semantic Segmentation 1 Feb 2025 · 1 repository · arXiv:2502.00563Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Contrastive Forward-Forward: A Training Algorithm of Vision Transformer 1 Feb 2025 · 0 repositories · arXiv:2502.00571
-
Converting Transformers into DGNNs Form 1 Feb 2025 · 1 repository · arXiv:2502.00585
-
Data-Driven Mispronunciation Pattern Discovery for Robust Speech Recognition 1 Feb 2025 · 0 repositories · arXiv:2502.00583
-
Dominated Novelty Search: Rethinking Local Competition in Quality-Diversity 1 Feb 2025 · 1 repository · arXiv:2502.00593Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples)
-
Evaluation of End-to-End Continuous Spanish Lipreading in Different Data Conditions 1 Feb 2025 · 1 repository · arXiv:2502.00464
-
Explainable AI for Sentiment Analysis of Human Metapneumovirus (HMPV) Using XLNet 1 Feb 2025 · 0 repositories · arXiv:2502.01663
-
Fast Solvers for Discrete Diffusion Models: Theory and Applications of High-Order Algorithms 1 Feb 2025 · 0 repositories · arXiv:2502.00234
-
Generating crossmodal gene expression from cancer histopathology improves multimodal AI predictions 1 Feb 2025 · 1 repository · arXiv:2502.00568
-
M+: Extending MemoryLLM with Scalable Long-Term Memory 1 Feb 2025 · 1 repository · arXiv:2502.00592Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
MambaGlue: Fast and Robust Local Feature Matching With Mamba 1 Feb 2025 · 1 repository · arXiv:2502.00462
-
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition 1 Feb 2025 · 1 repository · arXiv:2502.00547
-
Multi-Order Hyperbolic Graph Convolution and Aggregated Attention for Social Event Detection 1 Feb 2025 · 0 repositories · arXiv:2502.00351
-
Pause-Tuning for Long-Context Comprehension: A Lightweight Approach to LLM Attention Recalibration 1 Feb 2025 · 0 repositories · arXiv:2502.20405
-
PM-MOE: Mixture of Experts on Private Model Parameters for Personalized Federated Learning 1 Feb 2025 · 1 repository · arXiv:2502.00354Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Provably-Stable Neural Network-Based Control of Nonlinear Systems 1 Feb 2025 · 1 repository · arXiv:2502.00248
-
Riddle Me This! Stealthy Membership Inference for Retrieval-Augmented Generation 1 Feb 2025 · 1 repository · arXiv:2502.00306
-
Sigmoid Self-Attention has Lower Sample Complexity than Softmax Self-Attention: A Mixture-of-Experts Perspective 1 Feb 2025 · 0 repositories · arXiv:2502.00281
-
SigWavNet: Learning Multiresolution Signal Wavelet Network for Speech Emotion Recognition 1 Feb 2025 · 1 repository · arXiv:2502.00310
-
Spectro-Riemannian Graph Neural Networks 1 Feb 2025 · 0 repositories · arXiv:2502.00401Syntology 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
SSRepL-ADHD: Adaptive Complex Representation Learning Framework for ADHD Detection from Visual Attention Tasks 1 Feb 2025 · 0 repositories · arXiv:2502.00376
-
VertiFormer: A Data-Efficient Multi-Task Transformer for Off-Road Robot Mobility 1 Feb 2025 · 1 repository · arXiv:2502.00543
-
Accelerating Diffusion Transformer via Error-Optimized Cache 31 Jan 2025 · 0 repositories · arXiv:2501.19243
-
Beyond Token Compression: A Training-Free Reduction Framework for Efficient Visual Processing in MLLMs 31 Jan 2025 · 0 repositories · arXiv:2501.19036
-
Can AI Solve the Peer Review Crisis? A Large Scale Cross Model Experiment of LLMs' Performance and Biases in Evaluating over 1000 Economics Papers 31 Jan 2025 · 0 repositories · arXiv:2502.00070
-
CerraData-4MM: A multimodal benchmark dataset on Cerrado for land use and land cover classification 31 Jan 2025 · 1 repository · arXiv:2502.00083
-
Clustering in hyperbolic balls 31 Jan 2025 · 0 repositories · arXiv:2501.19247
-
Collaborative Diffusion Model for Recommender System 31 Jan 2025 · 0 repositories · arXiv:2501.18997
-
Context Matters: Query-aware Dynamic Long Sequence Modeling of Gigapixel Images 31 Jan 2025 · 1 repository · arXiv:2501.18984
-
ContextFormer: Redefining Efficiency in Semantic Segmentation 31 Jan 2025 · 0 repositories · arXiv:2501.19255
-
Do LLMs Strategically Reveal, Conceal, and Infer Information? A Theoretical and Empirical Analysis in The Chameleon Game 31 Jan 2025 · 1 repository · arXiv:2501.19398
-
Efficient Supernet Training with Orthogonal Softmax for Scalable ASR Model Compression 31 Jan 2025 · 0 repositories · arXiv:2501.18895
-
Employee Turnover Prediction: A Cross-component Attention Transformer with Consideration of Competitor Influence and Contagious Effect 31 Jan 2025 · 0 repositories · arXiv:2502.01660
-
FlexiCrackNet: A Flexible Pipeline for Enhanced Crack Segmentation with General Features Transfered from SAM 31 Jan 2025 · 0 repositories · arXiv:2501.18855
-
From Semantic Segmentation of Natural Images to Medical Image Segmentation Using ViT-Based Architectures 31 Jan 2025 · 0 repositories
-
Full-scale Representation Guided Network for Retinal Vessel Segmentation 31 Jan 2025 · 1 repository · arXiv:2501.18921
-
Homogeneity Bias as Differential Sampling Uncertainty in Language Models 31 Jan 2025 · 0 repositories · arXiv:2501.19337
-
Improving vision-language alignment with graph spiking hybrid Networks 31 Jan 2025 · 0 repositories · arXiv:2501.19069
-
Intrinsic Tensor Field Propagation in Large Language Models: A Novel Approach to Contextual Information Flow 31 Jan 2025 · 0 repositories · arXiv:2501.18957
-
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search 31 Jan 2025 · 1 repository · arXiv:2501.18922Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
Large Language Models' Accuracy in Emulating Human Experts' Evaluation of Public Sentiments about Heated Tobacco Products on Social Media 31 Jan 2025 · 0 repositories · arXiv:2502.01658
-
Laser: Efficient Language-Guided Segmentation in Neural Radiance Fields 31 Jan 2025 · 1 repository · arXiv:2501.19084
-
LiDAR Loop Closure Detection using Semantic Graphs with Graph Attention Networks 31 Jan 2025 · 1 repository · arXiv:2501.19382
-
Longer Attention Span: Increasing Transformer Context Length with Sparse Graph Processing Techniques 31 Jan 2025 · 1 repository · arXiv:2502.01659
-
Multi-Frame Blind Manifold Deconvolution for Rotating Synthetic Aperture Imaging 31 Jan 2025 · 0 repositories · arXiv:2501.19386
-
Neuro-LIFT: A Neuromorphic, LLM-based Interactive Framework for Autonomous Drone FlighT at the Edge 31 Jan 2025 · 0 repositories · arXiv:2501.19259
-
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models 31 Jan 2025 · 0 repositories · arXiv:2501.19090
-
PixelWorld: Towards Perceiving Everything as Pixels 31 Jan 2025 · 0 repositories · arXiv:2501.19339
-
Privacy Preserving Charge Location Prediction for Electric Vehicles 31 Jan 2025 · 0 repositories · arXiv:2502.00068
-
RGB-Event ISP: The Dataset and Benchmark 31 Jan 2025 · 1 repository · arXiv:2501.19129
-
Scalable-Softmax Is Superior for Attention 31 Jan 2025 · 1 repository · arXiv:2501.19399
-
Self-Supervised Cross-Modal Text-Image Time Series Retrieval in Remote Sensing 31 Jan 2025 · 0 repositories · arXiv:2501.19043
-
Self-Supervised Learning Using Nonlinear Dependence 31 Jan 2025 · 0 repositories · arXiv:2501.18875
-
Strassen Attention: Unlocking Compositional Abilities in Transformers Based on a New Lower Bound Method 31 Jan 2025 · 0 repositories · arXiv:2501.19215
-
Through the Looking Glass: LLM-Based Analysis of AR/VR Android Applications Privacy Policies 31 Jan 2025 · 0 repositories · arXiv:2501.19223
-
Understanding Generalization in Physics Informed Models through Affine Variety Dimensions 31 Jan 2025 · 0 repositories · arXiv:2501.18879
-
A Learnable Multi-views Contrastive Framework with Reconstruction Discrepancy for Medical Time-Series 30 Jan 2025 · 0 repositories · arXiv:2501.18367
-
A Unified Perspective on the Dynamics of Deep Transformers 30 Jan 2025 · 0 repositories · arXiv:2501.18322
-
Adaptive Object Detection for Indoor Navigation Assistance: A Performance Evaluation of Real-Time Algorithms 30 Jan 2025 · 0 repositories · arXiv:2501.18444
-
AlphaAdam:Asynchronous Masked Optimization with Dynamic Alpha for Selective Updates 30 Jan 2025 · 0 repositories · arXiv:2501.18094
-
Arbitrary Data as Images: Fusion of Patient Data Across Modalities and Irregular Intervals with Vision Transformers 30 Jan 2025 · 0 repositories · arXiv:2501.18237
-
Can we Retrieve Everything All at Once? ARM: An Alignment-Oriented LLM-based Retrieval Method 30 Jan 2025 · 0 repositories · arXiv:2501.18539
-
Contextually Structured Token Dependency Encoding for Large Language Models 30 Jan 2025 · 0 repositories · arXiv:2501.18205
-
DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights 30 Jan 2025 · 0 repositories · arXiv:2501.18596
-
Economic Rationality under Specialization: Evidence of Decision Bias in AI Agents 30 Jan 2025 · 0 repositories · arXiv:2501.18190
-
Efficient Transformer for High Resolution Image Motion Deblurring 30 Jan 2025 · 1 repository · arXiv:2501.18403
-
Entropy-Synchronized Neural Hashing for Unsupervised Ransomware Detection 30 Jan 2025 · 0 repositories · arXiv:2501.18131
-
Evaluating Large Language Models in Vulnerability Detection Under Variable Context Windows 30 Jan 2025 · 0 repositories · arXiv:2502.00064
-
GDformer: Going Beyond Subsequence Isolation for Multivariate Time Series Anomaly Detection 30 Jan 2025 · 1 repository · arXiv:2501.18196
-
General Embedding vs. Task-Specific Embedding: A Comparative Approach to Enhancing NLP Performance 30 Jan 2025 · 0 repositories
-
GENIE: Generative Note Information Extraction model for structuring EHR data 30 Jan 2025 · 0 repositories · arXiv:2501.18435
-
Hierarchical Multi-field Representations for Two-Stage E-commerce Retrieval 30 Jan 2025 · 0 repositories · arXiv:2501.18707
-
On the Role of Transformer Feed-Forward Layers in Nonlinear In-Context Learning 30 Jan 2025 · 0 repositories · arXiv:2501.18187
-
Israel-Hamas war through Telegram, Reddit and Twitter 30 Jan 2025 · 0 repositories · arXiv:2502.00060
-
Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models 30 Jan 2025 · 0 repositories · arXiv:2501.18280
-
Leveraging LLM Agents for Automated Optimization Modeling for SASP Problems: A Graph-RAG based Approach 30 Jan 2025 · 0 repositories · arXiv:2501.18320