Methods › General › Output Functions › Softmax › Papers, page 80
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 80 of 375: papers 7,901 to 8,000 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Squeezed Attention: Accelerating Long Context Length LLM Inference 14 Nov 2024 · 1 repository · arXiv:2411.09688Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Stability and Generalization for Distributed SGDA 14 Nov 2024 · 0 repositories · arXiv:2411.09365
-
The Future of Skill: What Is It to Be Skilled at Work? 14 Nov 2024 · 0 repositories · arXiv:2411.10488
-
Towards a Classification of Open-Source ML Models and Datasets for Software Engineering 14 Nov 2024 · 0 repositories · arXiv:2411.09683
-
A Large-Scale Study of Relevance Assessments with Large Language Models: An Initial Look 13 Nov 2024 · 1 repository · arXiv:2411.08275
-
A Transformer-Based Visual Piano Transcription Algorithm 13 Nov 2024 · 0 repositories · arXiv:2411.09037
-
AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding 13 Nov 2024 · 0 repositories · arXiv:2411.08451
-
Advanced Nonlinear SCMA Codebook Design Based on Lattice Constellations 13 Nov 2024 · 0 repositories · arXiv:2411.08493
-
Analyst Reports and Stock Performance: Evidence from the Chinese Market 13 Nov 2024 · 0 repositories · arXiv:2411.08726
-
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection 13 Nov 2024 · 0 repositories · arXiv:2411.08868
-
Continuous GNN-based Anomaly Detection on Edge using Efficient Adaptive Knowledge Graph Learning 13 Nov 2024 · 0 repositories · arXiv:2411.09072
-
Cut Your Losses in Large-Vocabulary Language Models 13 Nov 2024 · 2 repositories · arXiv:2411.09009Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
FinRobot: AI Agent for Equity Research and Valuation with Large Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08804Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples)
-
Flow reconstruction in time-varying geometries using graph neural networks 13 Nov 2024 · 0 repositories · arXiv:2411.08764
-
Fluoroformer: Scaling multiple instance learning to multiplexed images via attention-based channel fusion 13 Nov 2024 · 1 repository · arXiv:2411.08975
-
LLMStinger: Jailbreaking LLMs using RL fine-tuned LLMs 13 Nov 2024 · 0 repositories · arXiv:2411.08862
-
LogLLM: Log-based Anomaly Detection Using Large Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08561
-
Multimodal Object Detection using Depth and Image Data for Manufacturing Parts 13 Nov 2024 · 0 repositories · arXiv:2411.09062
-
Oblique Bayesian additive regression trees 13 Nov 2024 · 0 repositories · arXiv:2411.08849
-
PerceiverS: A Multi-Scale Perceiver with Effective Segmentation for Long-Term Expressive Symbolic Music Generation 13 Nov 2024 · 0 repositories · arXiv:2411.08307
-
Quantity versus Diversity: Influence of Data on Detecting EEG Pathology with Advanced ML Models 13 Nov 2024 · 0 repositories · arXiv:2411.17709
-
ReMP: Reusable Motion Prior for Multi-domain 3D Human Pose Estimation and Motion Inbetweening 13 Nov 2024 · 0 repositories · arXiv:2411.09435
-
RESOLVE: Relational Reasoning with Symbolic and Object-Level Features Using Vector Symbolic Processing 13 Nov 2024 · 1 repository · arXiv:2411.08290
-
Responsible AI in Construction Safety: Systematic Evaluation of Large Language Models and Prompt Engineering 13 Nov 2024 · 0 repositories · arXiv:2411.08320
-
Retrieval Augmented Recipe Generation 13 Nov 2024 · 0 repositories · arXiv:2411.08715
-
SAD-TIME: a Spatiotemporal-fused network for depression detection with Automated multi-scale Depth-wise and TIME-interval-related common feature extractor 13 Nov 2024 · 0 repositories · arXiv:2411.08521
-
SAM-I2I: Unleash the Power of Segment Anything Model for Medical Image Translation 13 Nov 2024 · 0 repositories · arXiv:2411.12755
-
SASE: A Searching Architecture for Squeeze and Excitation Operations 13 Nov 2024 · 0 repositories · arXiv:2411.08333
-
Scale Contrastive Learning with Selective Attentions for Blind Image Quality Assessment 13 Nov 2024 · 0 repositories · arXiv:2411.09007
-
ScaleNet: Scale Invariance Learning in Directed Graphs 13 Nov 2024 · 1 repository · arXiv:2411.08758
-
Towards Objective and Unbiased Decision Assessments with LLM-Enhanced Hierarchical Attention Networks 13 Nov 2024 · 1 repository · arXiv:2411.08504
-
Towards Optimizing a Retrieval Augmented Generation using Large Language Model on Academic Data 13 Nov 2024 · 0 repositories · arXiv:2411.08438
-
TRACE: Transformer-based Risk Assessment for Clinical Evaluation 13 Nov 2024 · 1 repository · arXiv:2411.08701
-
UIFormer: A Unified Transformer-based Framework for Incremental Few-Shot Object Detection and Instance Segmentation 13 Nov 2024 · 0 repositories · arXiv:2411.08569
-
VALTEST: Automated Validation of Language Model Generated Test Cases 13 Nov 2024 · 0 repositories · arXiv:2411.08254
-
A Preview of XiYan-SQL: A Multi-Generator Ensemble Framework for Text-to-SQL 13 Nov 2024 · 5 repositories · arXiv:2411.08599
-
Breaking the Low-Rank Dilemma of Linear Attention 12 Nov 2024 · 1 repository · arXiv:2411.07635Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
BudgetMLAgent: A Cost-Effective LLM Multi-Agent system for Automating Machine Learning Tasks 12 Nov 2024 · 0 repositories · arXiv:2411.07464
-
Can adversarial attacks by large language models be attributed? 12 Nov 2024 · 0 repositories · arXiv:2411.08003
-
Circuit Complexity Bounds for RoPE-based Transformer Architecture 12 Nov 2024 · 0 repositories · arXiv:2411.07602
-
Contrastive Language Prompting to Ease False Positives in Medical Anomaly Detection 12 Nov 2024 · 1 repository · arXiv:2411.07546
-
Controlled Evaluation of Syntactic Knowledge in Multilingual Language Models 12 Nov 2024 · 1 repository · arXiv:2411.07474
-
Deceiving Question-Answering Models: A Hybrid Word-Level Adversarial Approach 12 Nov 2024 · 1 repository · arXiv:2411.08248
-
Deep Learning 2.0: Artificial Neurons That Matter -- Reject Correlation, Embrace Orthogonality 12 Nov 2024 · 0 repositories · arXiv:2411.08085
-
Depthwise Separable Convolutions with Deep Residual Convolutions 12 Nov 2024 · 0 repositories · arXiv:2411.07544
-
DINO-LG: A Task-Specific DINO Model for Coronary Calcium Scoring 12 Nov 2024 · 0 repositories · arXiv:2411.07976
-
Efficient Federated Finetuning of Tiny Transformers with Resource-Constrained Devices 12 Nov 2024 · 0 repositories · arXiv:2411.07826
-
Emotion Classification of Children Expressions 12 Nov 2024 · 0 repositories · arXiv:2411.07708
-
Enhancing Link Prediction with Fuzzy Graph Attention Networks and Dynamic Negative Sampling 12 Nov 2024 · 0 repositories · arXiv:2411.07482
-
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis 12 Nov 2024 · 1 repository · arXiv:2411.07529
-
Fair Summarization: Bridging Quality and Diversity in Extractive Summaries 12 Nov 2024 · 1 repository · arXiv:2411.07521Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Fast Disentangled Slim Tensor Learning for Multi-view Clustering 12 Nov 2024 · 1 repository · arXiv:2411.07685
-
FM-TS: Flow Matching for Time Series Generation 12 Nov 2024 · 1 repository · arXiv:2411.07506Syntology official (archive's flag): 20 ran · 20 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 2 honoured, 3 violated, 11 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
HMIL: Hierarchical Multi-Instance Learning for Fine-Grained Whole Slide Image Classification 12 Nov 2024 · 1 repository · arXiv:2411.07660
-
Improving Grapheme-to-Phoneme Conversion through In-Context Knowledge Retrieval with Large Language Models 12 Nov 2024 · 0 repositories · arXiv:2411.07563
-
Interaction Asymmetry: A General Principle for Learning Composable Abstractions 12 Nov 2024 · 1 repository · arXiv:2411.07784Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Joint multi-dimensional dynamic attention and transformer for general image restoration 12 Nov 2024 · 1 repository · arXiv:2411.07893
-
Large Language Models Can Self-Improve in Long-context Reasoning 12 Nov 2024 · 1 repository · arXiv:2411.08147Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Leveraging Multimodal Models for Enhanced Neuroimaging Diagnostics in Alzheimer's Disease 12 Nov 2024 · 0 repositories · arXiv:2411.07871
-
LLM App Squatting and Cloning 12 Nov 2024 · 0 repositories · arXiv:2411.07518
-
Multi-task Feature Enhancement Network for No-Reference Image Quality Assessment 12 Nov 2024 · 0 repositories · arXiv:2411.07556
-
Multimodal Clinical Reasoning through Knowledge-augmented Rationale Generation 12 Nov 2024 · 0 repositories · arXiv:2411.07611
-
New Emerged Security and Privacy of Pre-trained Model: a Survey and Outlook 12 Nov 2024 · 0 repositories · arXiv:2411.07691
-
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models 12 Nov 2024 · 0 repositories · arXiv:2411.07820
-
Rendering-Oriented 3D Point Cloud Attribute Compression using Sparse Tensor-based Transformer 12 Nov 2024 · 0 repositories · arXiv:2411.07899
-
Retrieval Augmented Time Series Forecasting 12 Nov 2024 · 1 repository · arXiv:2411.08249Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Spatially Regularized Graph Attention Autoencoder Framework for Detecting Rainfall Extremes 12 Nov 2024 · 0 repositories · arXiv:2411.07753
-
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders 12 Nov 2024 · 0 repositories · arXiv:2411.07870
-
Two-Layer Attention Optimization for Bimanual Coordination 12 Nov 2024 · 0 repositories · arXiv:2411.07470
-
Unraveling the Gradient Descent Dynamics of Transformers 12 Nov 2024 · 0 repositories · arXiv:2411.07538
-
Verbosity ≠ Veracity: Demystify Verbosity Compensation Behavior of Large Language Models 12 Nov 2024 · 1 repository · arXiv:2411.07858
-
World Models: The Safety Perspective 12 Nov 2024 · 0 repositories · arXiv:2411.07690
-
A Unified Multi-Task Learning Architecture for Hate Detection Leveraging User-Based Information 11 Nov 2024 · 0 repositories · arXiv:2411.06855
-
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models 11 Nov 2024 · 1 repository · arXiv:2411.07232
-
AEROMamba: An efficient architecture for audio super-resolution using generative adversarial networks and state space models 11 Nov 2024 · 1 repository · arXiv:2411.07364
-
Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models 11 Nov 2024 · 0 repositories · arXiv:2411.06713
-
An Efficient Memory Module for Graph Few-Shot Class-Incremental Learning 11 Nov 2024 · 1 repository · arXiv:2411.06659
-
AssistRAG: Boosting the Potential of Large Language Models with an Intelligent Information Assistant 11 Nov 2024 · 1 repository · arXiv:2411.06805Syntology official (archive's flag): 10 ran · 10 ran (of which 1 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Autonomous Droplet Microfluidic Design Framework with Large Language Models 11 Nov 2024 · 1 repository · arXiv:2411.06691
-
Can KAN Work? Exploring the Potential of Kolmogorov-Arnold Networks in Computer Vision 11 Nov 2024 · 0 repositories · arXiv:2411.06727
-
Cancer-Answer: Empowering Cancer Care with Advanced Large Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.06946
-
ConvMixFormer- A Resource-efficient Convolution Mixer for Transformer-based Dynamic Hand Gesture Recognition 11 Nov 2024 · 1 repository · arXiv:2411.07118
-
Data-Driven Analysis of AI in Medical Device Software in China: Deep Learning and General AI Trends Based on Regulatory Data 11 Nov 2024 · 0 repositories · arXiv:2411.07378
-
Evaluating Large Language Models on Financial Report Summarization: An Empirical Study 11 Nov 2024 · 0 repositories · arXiv:2411.06852
-
Explore the Reasoning Capability of LLMs in the Chess Testbed 11 Nov 2024 · 0 repositories · arXiv:2411.06655
-
Fast and Robust Contextual Node Representation Learning over Dynamic Graphs 11 Nov 2024 · 0 repositories · arXiv:2411.07123
-
GTA-Net: An IoT-Integrated 3D Human Pose Estimation System for Real-Time Adolescent Sports Posture Correction 11 Nov 2024 · 0 repositories · arXiv:2411.06725
-
SynCL: A Synergistic Training Strategy with Instance-Aware Contrastive Learning for End-to-End Multi-Camera 3D Tracking 11 Nov 2024 · 0 repositories · arXiv:2411.06780
-
Invar-RAG: Invariant LLM-aligned Retrieval for Better Generation 11 Nov 2024 · 0 repositories · arXiv:2411.07021
-
Isochrony-Controlled Speech-to-Text Translation: A study on translating from Sino-Tibetan to Indo-European Languages 11 Nov 2024 · 0 repositories · arXiv:2411.07387
-
LA4SR: illuminating the dark proteome with generative AI 11 Nov 2024 · 0 repositories · arXiv:2411.06798
-
Layout Control and Semantic Guidance with Attention Loss Backward for T2I Diffusion Model 11 Nov 2024 · 0 repositories · arXiv:2411.06692
-
LongSafetyBench: Long-Context LLMs Struggle with Safety Issues 11 Nov 2024 · 1 repository · arXiv:2411.06899Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MapSAM: Adapting Segment Anything Model for Automated Feature Detection in Historical Maps 11 Nov 2024 · 1 repository · arXiv:2411.06971
-
Modeling variable guide efficiency in pooled CRISPR screens with ContrastiveVI+ 11 Nov 2024 · 0 repositories · arXiv:2411.08072
-
More Expressive Attention with Negative Weights 11 Nov 2024 · 1 repository · arXiv:2411.07176Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Multi-Modal interpretable automatic video captioning 11 Nov 2024 · 0 repositories · arXiv:2411.06872
-
On Active Privacy Auditing in Supervised Fine-tuning for White-Box Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.07070
-
PCNet: a human pose compensation network based on incremental learning for sports actions estimation 11 Nov 2024 · 0 repositories
-
ScaleKD: Strong Vision Transformers Could Be Excellent Teachers 11 Nov 2024 · 1 repository · arXiv:2411.06786