Methods › General › Output Functions › Softmax › Papers, page 98
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 98 of 375: papers 9,701 to 9,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards Understanding the Feasibility of Machine Unlearning 3 Oct 2024 · 0 repositories · arXiv:2410.03043
-
Towards Understanding the Universality of Transformers for Next-Token Prediction 3 Oct 2024 · 0 repositories · arXiv:2410.03011
-
Training Language Models on Synthetic Edit Sequences Improves Code Synthesis 3 Oct 2024 · 1 repository · arXiv:2410.02749Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Training Nonlinear Transformers for Chain-of-Thought Inference: A Theoretical Generalization Analysis 3 Oct 2024 · 0 repositories · arXiv:2410.02167
-
TrajGPT: Irregular Time-Series Representation Learning for Health Trajectory Analysis 3 Oct 2024 · 0 repositories · arXiv:2410.02133
-
TRIS-HAR: Transmissive Reconfigurable Intelligent Surfaces-assisted Cognitive Wireless Human Activity Recognition Using State Space Models 3 Oct 2024 · 0 repositories · arXiv:2410.02334
-
UncertaintyRAG: Span-Level Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation 3 Oct 2024 · 0 repositories · arXiv:2410.02719
-
Vinoground: Scrutinizing LMMs over Dense Temporal Reasoning with Short Videos 3 Oct 2024 · 1 repository · arXiv:2410.02763
-
Visual Editing with LLM-based Tool Chaining: An Efficient Distillation Approach for Real-Time Applications 3 Oct 2024 · 1 repository · arXiv:2410.02952
-
A Control Barrier Function Candidate for Quadrotors with Limited Field of View 2 Oct 2024 · 0 repositories · arXiv:2410.01277
-
A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts 2 Oct 2024 · 0 repositories · arXiv:2410.01485
-
A Spark of Vision-Language Intelligence: 2-Dimensional Autoregressive Transformer for Efficient Finegrained Image Generation 2 Oct 2024 · 1 repository · arXiv:2410.01912
-
A versatile machine learning workflow for high-throughput analysis of supported metal catalyst particles 2 Oct 2024 · 1 repository · arXiv:2410.01213
-
Addressing Data Heterogeneity in Federated Learning with Adaptive Normalization-Free Feature Recalibration 2 Oct 2024 · 0 repositories · arXiv:2410.02006
-
AHP-Powered LLM Reasoning for Multi-Criteria Evaluation of Open-Ended Responses 2 Oct 2024 · 0 repositories · arXiv:2410.01246
-
Attention layers provably solve single-location regression 2 Oct 2024 · 1 repository · arXiv:2410.01537Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Automated Red Teaming with GOAT: the Generative Offensive Agent Tester 2 Oct 2024 · 0 repositories · arXiv:2410.01606
-
Automatic deductive coding in discourse analysis: an application of large language models in learning analytics 2 Oct 2024 · 1 repository · arXiv:2410.01240
-
BordIRlines: A Dataset for Evaluating Cross-lingual Retrieval-Augmented Generation 2 Oct 2024 · 1 repository · arXiv:2410.01171Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
ComaDICE: Offline Cooperative Multi-Agent Reinforcement Learning with Stationary Distribution Shift Regularization 2 Oct 2024 · 0 repositories · arXiv:2410.01954
-
DeepProtein: Deep Learning Library and Benchmark for Protein Sequence Learning 2 Oct 2024 · 1 repository · arXiv:2410.02023
-
Depth Pro: Sharp Monocular Metric Depth in Less Than a Second 2 Oct 2024 · 1 repository · arXiv:2410.02073Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
DRUPI: Dataset Reduction Using Privileged Information 2 Oct 2024 · 0 repositories · arXiv:2410.01611
-
DynFrs: An Efficient Framework for Machine Unlearning in Random Forest 2 Oct 2024 · 1 repository · arXiv:2410.01588Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Efficient Length-Generalizable Attention via Causal Retrieval for Long-Context Language Modeling 2 Oct 2024 · 0 repositories · arXiv:2410.01651
-
Efficient Streaming LLM for Speech Recognition 2 Oct 2024 · 0 repositories · arXiv:2410.03752
-
Emotion-Aware Embedding Fusion in LLMs (Flan-T5, LLAMA 2, DeepSeek-R1, and ChatGPT 4) for Intelligent Response Generation 2 Oct 2024 · 0 repositories · arXiv:2410.01306
-
Enhancing LLM Fine-tuning for Text-to-SQLs by SQL Quality Measurement 2 Oct 2024 · 0 repositories · arXiv:2410.01869
-
Enhancing Retrieval in QA Systems with Derived Feature Association 2 Oct 2024 · 1 repository · arXiv:2410.03754
-
ENTP: Encoder-only Next Token Prediction 2 Oct 2024 · 0 repositories · arXiv:2410.01600
-
ET-Plan-Bench: Embodied Task-level Planning Benchmark Towards Spatial-Temporal Cognition with Foundation Models 2 Oct 2024 · 0 repositories · arXiv:2410.14682
-
EVA-Gaussian: 3D Gaussian-based Real-time Human Novel View Synthesis under Diverse Camera Settings 2 Oct 2024 · 0 repositories · arXiv:2410.01425
-
Facial Action Unit Detection by Adaptively Constraining Self-Attention and Causally Deconfounding Sample 2 Oct 2024 · 1 repository · arXiv:2410.01251
-
Financial Sentiment Analysis on News and Reports Using Large Language Models and FinBERT 2 Oct 2024 · 0 repositories · arXiv:2410.01987
-
FLAG: Financial Long Document Classification via AMR-based GNN 2 Oct 2024 · 1 repository · arXiv:2410.02024
-
FlashMask: Efficient and Rich Mask Extension of FlashAttention 2 Oct 2024 · 1 repository · arXiv:2410.01359Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
GCM-Net: Graph-enhanced Cross-Modal Infusion with a Metaheuristic-Driven Network for Video Sentiment and Emotion Analysis 2 Oct 2024 · 0 repositories · arXiv:2410.12828
-
Getting Free Bits Back from Rotational Symmetries in LLMs 2 Oct 2024 · 0 repositories · arXiv:2410.01309
-
HyperBrain: Anomaly Detection for Temporal Hypergraph Brain Networks 2 Oct 2024 · 1 repository · arXiv:2410.02087
-
LS-HAR: Language Supervised Human Action Recognition with Salient Fusion, Construction Sites as a Use-Case 2 Oct 2024 · 0 repositories · arXiv:2410.01962
-
Locret: Enhancing Eviction in Long-Context LLM Inference with Trained Retaining Heads on Consumer-Grade Devices 2 Oct 2024 · 1 repository · arXiv:2410.01805
-
MARPLE: A Benchmark for Long-Horizon Inference 2 Oct 2024 · 1 repository · arXiv:2410.01926Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Mind Scramble: Unveiling Large Language Model Psychology Via Typoglycemia 2 Oct 2024 · 1 repository · arXiv:2410.01677
-
Normalizing Flow-Based Metric for Image Generation 2 Oct 2024 · 1 repository · arXiv:2410.02004
-
OmniGenBench: Automating Large-scale in-silico Benchmarking for Genomic Foundation Models 2 Oct 2024 · 3 repositories · arXiv:2410.01784
-
OmniSR: Shadow Removal under Direct and Indirect Lighting 2 Oct 2024 · 1 repository · arXiv:2410.01719
-
On The Adaptation of Unlimiformer for Decoder-Only Transformers 2 Oct 2024 · 0 repositories · arXiv:2410.01637
-
Open-RAG: Enhanced Retrieval-Augmented Reasoning with Open-Source Large Language Models 2 Oct 2024 · 1 repository · arXiv:2410.01782
-
Perceptual Piercing: Human Visual Cue-based Object Detection in Low Visibility Conditions 2 Oct 2024 · 1 repository · arXiv:2410.01225
-
Positional Attention: Expressivity and Learnability of Algorithmic Computation 2 Oct 2024 · 1 repository · arXiv:2410.01686Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Price-guided user attention in large-scale E-commerce group recommendation 2 Oct 2024 · 0 repositories · arXiv:2410.02074
-
Quantifying Generalization Complexity for Large Language Models 2 Oct 2024 · 1 repository · arXiv:2410.01769Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
RADAR: Robust Two-stage Modality-incomplete Industrial Anomaly Detection 2 Oct 2024 · 0 repositories · arXiv:2410.01737
-
RS-FME-SwinT: A Novel Feature Map Enhancement Framework Integrating Customized SwinT with Residual and Spatial CNN for Monkeypox Diagnosis 2 Oct 2024 · 0 repositories · arXiv:2410.01216
-
Saliency-Guided DETR for Moment Retrieval and Highlight Detection 2 Oct 2024 · 1 repository · arXiv:2410.01615
-
Seeing Eye to AI: Human Alignment via Gaze-Based Response Rewards for Large Language Models 2 Oct 2024 · 1 repository · arXiv:2410.01532Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
SurgPointTransformer: Vertebrae Shape Completion with RGB-D Data 2 Oct 2024 · 0 repositories · arXiv:2410.01443
-
TAEGAN: Generating Synthetic Tabular Data For Data Augmentation 2 Oct 2024 · 0 repositories · arXiv:2410.01933
-
The potential of LLM-generated reports in DevSecOps 2 Oct 2024 · 0 repositories · arXiv:2410.01899
-
TIGER: Time-frequency Interleaved Gain Extraction and Reconstruction for Efficient Speech Separation 2 Oct 2024 · 0 repositories · arXiv:2410.01469
-
TiVaT: A Transformer with a Single Unified Mechanism for Capturing Asynchronous Dependencies in Multivariate Time Series Forecasting 2 Oct 2024 · 0 repositories · arXiv:2410.01531
-
Towards a Deeper Understanding of Transformer for Residential Non-intrusive Load Monitoring 2 Oct 2024 · 0 repositories · arXiv:2410.03758
-
Towards Dynamic Graph Neural Networks with Provably High-Order Expressive Power 2 Oct 2024 · 0 repositories · arXiv:2410.01367
-
Tracking objects that change in appearance with phase synchrony 2 Oct 2024 · 0 repositories · arXiv:2410.02094
-
UlcerGPT: A Multimodal Approach Leveraging Large Language and Vision Models for Diabetic Foot Ulcer Image Transcription 2 Oct 2024 · 0 repositories · arXiv:2410.01989
-
VectorGraphNET: Graph Attention Networks for Accurate Segmentation of Complex Technical Drawings 2 Oct 2024 · 0 repositories · arXiv:2410.01336
-
Why context matters in VQA and Reasoning: Semantic interventions for VLM input modalities 2 Oct 2024 · 0 repositories · arXiv:2410.01690
-
Addition is All You Need for Energy-efficient Language Models 1 Oct 2024 · 0 repositories · arXiv:2410.00907
-
Advanced Arabic Alphabet Sign Language Recognition Using Transfer Learning and Transformer Models 1 Oct 2024 · 0 repositories · arXiv:2410.00681
-
Advancing RVFL networks: Robust classification with the HawkEye loss function 1 Oct 2024 · 1 repository · arXiv:2410.00510
-
Unleashing the Unseen: Harnessing Benign Datasets for Jailbreaking Large Language Models 1 Oct 2024 · 1 repository · arXiv:2410.00451
-
AI Persuasion, Bayesian Attribution, and Career Concerns of Doctors 1 Oct 2024 · 0 repositories · arXiv:2410.01114
-
AlignSum: Data Pyramid Hierarchical Fine-tuning for Aligning with Human Summarization Preference 1 Oct 2024 · 1 repository · arXiv:2410.00409Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Explainable AI for Fraud Detection: An Attention-Based Ensemble of CNNs, GNNs, and A Confidence-Driven Gating Mechanism 1 Oct 2024 · 0 repositories · arXiv:2410.09069
-
BabelBench: An Omni Benchmark for Code-Driven Analysis of Multimodal and Multistructured Data 1 Oct 2024 · 1 repository · arXiv:2410.00773
-
Creative and Context-Aware Translation of East Asian Idioms with GPT-4 1 Oct 2024 · 1 repository · arXiv:2410.00988
-
Data-driven Framework for Forward and Inverse Problems in Guided Waves-Based Structural Health Monitoring Under Varying Environmental and Operating Conditions 1 Oct 2024 · 0 repositories · arXiv:2410.01127
-
Decoding Hate: Exploring Language Models' Reactions to Hate Speech 1 Oct 2024 · 0 repositories · arXiv:2410.00775
-
Deep Multimodal Fusion for Semantic Segmentation of Remote Sensing Earth Observation Data 1 Oct 2024 · 0 repositories · arXiv:2410.00469
-
Domain Aware Multi-Task Pretraining of 3D Swin Transformer for T1-weighted Brain MRI 1 Oct 2024 · 1 repository · arXiv:2410.00410
-
End-to-End Speech Recognition with Pre-trained Masked Language Model 1 Oct 2024 · 1 repository · arXiv:2410.00528
-
Exploring the Learning Capabilities of Language Models using LEVERWORLDS 1 Oct 2024 · 0 repositories · arXiv:2410.00519
-
Pediatric Wrist Fracture Detection Using Feature Context Excitation Modules in X-ray Images 1 Oct 2024 · 1 repository · arXiv:2410.01031
-
GLMHA A Guided Low-rank Multi-Head Self-Attention for Efficient Image Restoration and Spectral Reconstruction 1 Oct 2024 · 0 repositories · arXiv:2410.00380
-
GSPR: Multimodal Place Recognition Using 3D Gaussian Splatting for Autonomous Driving 1 Oct 2024 · 1 repository · arXiv:2410.00299
-
Insight: A Multi-Modal Diagnostic Pipeline using LLMs for Ocular Surface Disease Diagnosis 1 Oct 2024 · 0 repositories · arXiv:2410.00292
-
Language Enhanced Model for Eye (LEME): An Open-Source Ophthalmology-Specific Large Language Model 1 Oct 2024 · 0 repositories · arXiv:2410.03740
-
Learning Adaptive Hydrodynamic Models Using Neural ODEs in Complex Conditions 1 Oct 2024 · 0 repositories · arXiv:2410.00490
-
MAP: Unleashing Hybrid Mamba-Transformer Vision Backbone's Potential with Masked Autoregressive Pretraining 1 Oct 2024 · 0 repositories · arXiv:2410.00871
-
Multi-Scale Temporal Transformer For Speech Emotion Recognition 1 Oct 2024 · 0 repositories · arXiv:2410.00390
-
nGPT: Normalized Transformer with Representation Learning on the Hypersphere 1 Oct 2024 · 0 repositories · arXiv:2410.01131
-
Optimizing and Evaluating Enterprise Retrieval-Augmented Generation (RAG): A Content Design Perspective 1 Oct 2024 · 1 repository · arXiv:2410.12812
-
PclGPT: A Large Language Model for Patronizing and Condescending Language Detection 1 Oct 2024 · 1 repository · arXiv:2410.00361
-
Quantifying reliance on external information over parametric knowledge during Retrieval Augmented Generation (RAG) using mechanistic analysis 1 Oct 2024 · 0 repositories · arXiv:2410.00857
-
RATIONALYST: Pre-training Process-Supervision for Improving Reasoning 1 Oct 2024 · 1 repository · arXiv:2410.01044
-
Replacing Paths with Connection-Biased Attention for Knowledge Graph Completion 1 Oct 2024 · 1 repository · arXiv:2410.00876
-
Revisiting the Role of Texture in 3D Person Re-identification 1 Oct 2024 · 0 repositories · arXiv:2410.00348
-
Robust Traffic Forecasting against Spatial Shift over Years 1 Oct 2024 · 1 repository · arXiv:2410.00373Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Scene Graph Disentanglement and Composition for Generalizable Complex Image Generation 1 Oct 2024 · 0 repositories · arXiv:2410.00447
-
SCINet: Spatial and Contrast Interactive Super-Resolution Assisted Infrared UAV Target Detection 1 Oct 2024 · 2 repositories