Methods › General › Output Functions › Softmax › Papers, page 93
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 93 of 375: papers 9,201 to 9,300 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
QIANets: Quantum-Integrated Adaptive Networks for Reduced Latency and Improved Inference Times in CNN Models 14 Oct 2024 · 1 repository · arXiv:2410.10318Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Queryable Prototype Multiple Instance Learning with Vision-Language Models for Incremental Whole Slide Image Classification 14 Oct 2024 · 1 repository · arXiv:2410.10573
-
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models 14 Oct 2024 · 1 repository · arXiv:2410.10542
-
Reverse Refinement Network for Narrow Rural Road Detection in High-Resolution Satellite Imagery 14 Oct 2024 · 0 repositories · arXiv:2410.10389
-
Revisiting and Benchmarking Graph Autoencoders: A Contrastive Learning Perspective 14 Oct 2024 · 1 repository · arXiv:2410.10241
-
ROA-BEV: 2D Region-Oriented Attention for BEV-based 3D Object 14 Oct 2024 · 0 repositories · arXiv:2410.10298
-
RoCoFT: Efficient Finetuning of Large Language Models with Row-Column Updates 14 Oct 2024 · 1 repository · arXiv:2410.10075
-
Saliency Guided Optimization of Diffusion Latents 14 Oct 2024 · 0 repositories · arXiv:2410.10257
-
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers 14 Oct 2024 · 2 repositories · arXiv:2410.10629
-
SLaNC: Static LayerNorm Calibration 14 Oct 2024 · 0 repositories · arXiv:2410.10553
-
STACKFEED: Structured Textual Actor-Critic Knowledge Base Editing with FeedBack 14 Oct 2024 · 0 repositories · arXiv:2410.10584
-
The Ingredients for Robotic Diffusion Transformers 14 Oct 2024 · 0 repositories · arXiv:2410.10088
-
Towards Better Multi-head Attention via Channel-wise Sample Permutation 14 Oct 2024 · 1 repository · arXiv:2410.10914Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Transforming Game Play: A Comparative Study of DCQN and DTQN Architectures in Reinforcement Learning 14 Oct 2024 · 0 repositories · arXiv:2410.10660
-
Transparent Networks for Multivariate Time Series 14 Oct 2024 · 1 repository · arXiv:2410.10535
-
V2M: Visual 2-Dimensional Mamba for Image Representation Learning 14 Oct 2024 · 1 repository · arXiv:2410.10382
-
VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents 14 Oct 2024 · 1 repository · arXiv:2410.10594Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Watching the Watchers: Exposing Gender Disparities in Machine Translation Quality Estimation 14 Oct 2024 · 1 repository · arXiv:2410.10995
-
What Does It Mean to Be a Transformer? Insights from a Theoretical Hessian Analysis 14 Oct 2024 · 1 repository · arXiv:2410.10986Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
When Attention Sink Emerges in Language Models: An Empirical View 14 Oct 2024 · 1 repository · arXiv:2410.10781Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Will LLMs Replace the Encoder-Only Models in Temporal Relation Classification? 14 Oct 2024 · 1 repository · arXiv:2410.10476Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
3DS: Decomposed Difficulty Data Selection's Case Study on LLM Medical Domain Adaptation 13 Oct 2024 · 0 repositories · arXiv:2410.10901
-
A Comparative Study of PDF Parsing Tools Across Diverse Document Categories 13 Oct 2024 · 0 repositories · arXiv:2410.09871
-
BiDoRA: Bi-level Optimization-Based Weight-Decomposed Low-Rank Adaptation 13 Oct 2024 · 0 repositories · arXiv:2410.09758
-
Can In-context Learning Really Generalize to Out-of-distribution Tasks? 13 Oct 2024 · 0 repositories · arXiv:2410.09695
-
Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code 13 Oct 2024 · 0 repositories · arXiv:2410.09997
-
DAG-aware Transformer for Causal Effect Estimation 13 Oct 2024 · 1 repository · arXiv:2410.10044
-
Data Adaptive Few-shot Multi Label Segmentation with Foundation Model 13 Oct 2024 · 0 repositories · arXiv:2410.09759
-
Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces 13 Oct 2024 · 0 repositories · arXiv:2410.09918
-
EasyJudge: an Easy-to-use Tool for Comprehensive Response Evaluation of LLMs 13 Oct 2024 · 1 repository · arXiv:2410.09775
-
EBDM: Exemplar-guided Image Translation with Brownian-bridge Diffusion Models 13 Oct 2024 · 0 repositories · arXiv:2410.09802
-
EchoPrime: A Multi-Video View-Informed Vision-Language Model for Comprehensive Echocardiography Interpretation 13 Oct 2024 · 0 repositories · arXiv:2410.09704
-
Empowering Dysarthric Speech: Leveraging Advanced LLMs for Accurate Speech Correction and Multimodal Emotion Analysis 13 Oct 2024 · 0 repositories · arXiv:2410.12867
-
Evaluating Gender Bias of LLMs in Making Morality Judgements 13 Oct 2024 · 0 repositories · arXiv:2410.09992
-
HARDMath: A Benchmark Dataset for Challenging Problems in Applied Mathematics 13 Oct 2024 · 1 repository · arXiv:2410.09988
-
HASN: Hybrid Attention Separable Network for Efficient Image Super-resolution 13 Oct 2024 · 1 repository · arXiv:2410.09844
-
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG 13 Oct 2024 · 0 repositories · arXiv:2410.09699
-
InterMask: 3D Human Interaction Generation via Collaborative Masked Modelling 13 Oct 2024 · 1 repository · arXiv:2410.10010
-
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs 13 Oct 2024 · 0 repositories · arXiv:2410.12864
-
Joint Mixing Data Augmentation for Skeleton-based Action Recognition 13 Oct 2024 · 1 repository
-
Learning Pattern-Specific Experts for Time Series Forecasting Under Patch-level Distribution Shift 13 Oct 2024 · 1 repository · arXiv:2410.09836
-
Learning to Rank for Multiple Retrieval-Augmented Models through Iterative Utility Maximization 13 Oct 2024 · 0 repositories · arXiv:2410.09942
-
LibEER: A Comprehensive Benchmark and Algorithm Library for EEG-based Emotion Recognition 13 Oct 2024 · 2 repositories · arXiv:2410.09767
-
M2M-Gen: A Multimodal Framework for Automated Background Music Generation in Japanese Manga Using Large Language Models 13 Oct 2024 · 0 repositories · arXiv:2410.09928
-
Meta-Reinforcement Learning with Universal Policy Adaptation: Provable Near-Optimality under All-task Optimum Comparator 13 Oct 2024 · 0 repositories · arXiv:2410.09728
-
Online Multi-modal Root Cause Analysis 13 Oct 2024 · 0 repositories · arXiv:2410.10021
-
Single Ground Truth Is Not Enough: Add Linguistic Variability to Aspect-based Sentiment Analysis Evaluation 13 Oct 2024 · 0 repositories · arXiv:2410.09807
-
STA-Unet: Rethink the semantic redundant for Medical Imaging Segmentation 13 Oct 2024 · 1 repository · arXiv:2410.11578
-
TextMaster: Universal Controllable Text Edit 13 Oct 2024 · 0 repositories · arXiv:2410.09879
-
Automatic Speech Recognition with BERT and CTC Transformers: A Review 12 Oct 2024 · 0 repositories · arXiv:2410.09456
-
Beyond Exact Match: Semantically Reassessing Event Extraction by Large Language Models 12 Oct 2024 · 0 repositories · arXiv:2410.09418
-
Bi-temporal Gaussian Feature Dependency Guided Change Detection in Remote Sensing Images 12 Oct 2024 · 0 repositories · arXiv:2410.09539
-
Bridging Text and Image for Artist Style Transfer via Contrastive Learning 12 Oct 2024 · 0 repositories · arXiv:2410.09566
-
CollabEdit: Towards Non-destructive Collaborative Knowledge Editing 12 Oct 2024 · 1 repository · arXiv:2410.09508Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
Diabetic retinopathy image classification method based on GreenBen data augmentation 12 Oct 2024 · 0 repositories · arXiv:2410.09444
-
EG-SpikeFormer: Eye-Gaze Guided Transformer on Spiking Neural Networks for Medical Image Analysis 12 Oct 2024 · 0 repositories · arXiv:2410.09674
-
EmbodiedCity: A Benchmark Platform for Embodied Agent in Real-world City Environment 12 Oct 2024 · 0 repositories · arXiv:2410.09604
-
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling 12 Oct 2024 · 0 repositories · arXiv:2410.09524
-
Extended Japanese Commonsense Morality Dataset with Masked Token and Label Enhancement 12 Oct 2024 · 0 repositories · arXiv:2410.09564
-
Fine-grained Attention I/O Complexity: Comprehensive Analysis for Backward Passes 12 Oct 2024 · 0 repositories · arXiv:2410.09397
-
GPTON: Generative Pre-trained Transformers enhanced with Ontology Narration for accurate annotation of biological data 12 Oct 2024 · 0 repositories · arXiv:2410.10899
-
Hey AI Can You Grade My Essay?: Automatic Essay Grading 12 Oct 2024 · 0 repositories · arXiv:2410.09319
-
Improving 3D Finger Traits Recognition via Generalizable Neural Rendering 12 Oct 2024 · 0 repositories · arXiv:2410.09582
-
\llinstruct: An Instruction-tuned model for English Language Proficiency Assessments 12 Oct 2024 · 0 repositories · arXiv:2410.09314
-
Looped ReLU MLPs May Be All You Need as Practical Programmable Computers 12 Oct 2024 · 0 repositories · arXiv:2410.09375
-
Power-Softmax: Towards Secure LLM Inference over Encrypted Data 12 Oct 2024 · 0 repositories · arXiv:2410.09457
-
ReLU's Revival: On the Entropic Overload in Normalization-Free Large Language Models 12 Oct 2024 · 1 repository · arXiv:2410.09637Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Scaled and Inter-token Relation Enhanced Transformer for Sample-restricted Residential NILM 12 Oct 2024 · 0 repositories · arXiv:2410.12861
-
SimBrainNet: Evaluating Brain Network Similarity for Attention Disorders 12 Oct 2024 · 0 repositories · arXiv:2410.09422
-
Token Pruning using a Lightweight Background Aware Vision Transformer 12 Oct 2024 · 0 repositories · arXiv:2410.09324
-
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation 12 Oct 2024 · 1 repository · arXiv:2410.09584Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Training Dynamics of Transformers to Recognize Word Co-occurrence via Gradient Flow Analysis 12 Oct 2024 · 0 repositories · arXiv:2410.09605
-
Unraveling Movie Genres through Cross-Attention Fusion of Bi-Modal Synergy of Poster 12 Oct 2024 · 0 repositories · arXiv:2410.19764
-
Accelerated Distributed Stochastic Non-Convex Optimization over Time-Varying Directed Networks 11 Oct 2024 · 0 repositories · arXiv:2410.08508
-
A Methodology for Evaluating RAG Systems: A Case Study On Configuration Dependency Validation 11 Oct 2024 · 1 repository · arXiv:2410.08801
-
A Social Context-aware Graph-based Multimodal Attentive Learning Framework for Disaster Content Classification during Emergencies 11 Oct 2024 · 0 repositories · arXiv:2410.08814
-
Rethinking Gradient-Based Methods: Multi-Property Materials Design Beyond Differentiable Targets 11 Oct 2024 · 1 repository · arXiv:2410.08562
-
AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation 11 Oct 2024 · 1 repository · arXiv:2410.09040
-
Convolutional Neural Network Design and Evaluation for Real-Time Multivariate Time Series Fault Detection in Spacecraft Attitude Sensors 11 Oct 2024 · 0 repositories · arXiv:2410.09126
-
CoTCoNet: An Optimized Coupled Transformer-Convolutional Network with an Adaptive Graph Reconstruction for Leukemia Detection 11 Oct 2024 · 0 repositories · arXiv:2410.08797
-
Cross-Modal Bidirectional Interaction Model for Referring Remote Sensing Image Segmentation 11 Oct 2024 · 1 repository · arXiv:2410.08613
-
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation 11 Oct 2024 · 1 repository · arXiv:2410.08470
-
DeBiFormer: Vision Transformer with Deformable Agent Bi-level Routing Attention 11 Oct 2024 · 1 repository · arXiv:2410.08582
-
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08731
-
Efficiently Scanning and Resampling Spatio-Temporal Tasks with Irregular Observations 11 Oct 2024 · 0 repositories · arXiv:2410.08681
-
Encoding Agent Trajectories as Representations with Sequence Transformers 11 Oct 2024 · 0 repositories · arXiv:2410.09204
-
Enhancing Long Context Performance in LLMs Through Inner Loop Query Mechanism 11 Oct 2024 · 0 repositories · arXiv:2410.12859
-
Extra Global Attention Designation Using Keyword Detection in Sparse Transformer Architectures 11 Oct 2024 · 0 repositories · arXiv:2410.08971
-
Fine-Tuning In-House Large Language Models to Infer Differential Diagnosis from Radiology Reports 11 Oct 2024 · 0 repositories · arXiv:2410.09234
-
HorGait: A Hybrid Model for Accurate Gait Recognition in LiDAR Point Cloud Planar Projections 11 Oct 2024 · 0 repositories · arXiv:2410.08454
-
Humanity in AI: Detecting the Personality of Large Language Models 11 Oct 2024 · 0 repositories · arXiv:2410.08545
-
Hypothesis-only Biases in Large Language Model-Elicited Natural Language Inference 11 Oct 2024 · 0 repositories · arXiv:2410.08996
-
JAILJUDGE: A Comprehensive Jailbreak Judge Benchmark with Multi-Agent Enhanced Explanation Evaluation Framework 11 Oct 2024 · 1 repository · arXiv:2410.12855Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
L3Cube-MahaSum: A Comprehensive Dataset and BART Models for Abstractive Text Summarization in Marathi 11 Oct 2024 · 1 repository · arXiv:2410.09184
-
Large Language Models for Medical OSCE Assessment: A Novel Approach to Transcript Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.12858
-
Learning General Representation of 12-Lead Electrocardiogram with a Joint-Embedding Predictive Architecture 11 Oct 2024 · 1 repository · arXiv:2410.08559Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Learning Interaction-aware 3D Gaussian Splatting for One-shot Hand Avatars 11 Oct 2024 · 1 repository · arXiv:2410.08840Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Long Range Named Entity Recognition for Marathi Documents 11 Oct 2024 · 0 repositories · arXiv:2410.09192
-
Low-complexity Attention-based Unsupervised Anomalous Sound Detection exploiting Separable Convolutions and Angular Loss 11 Oct 2024 · 1 repository · arXiv:2410.08919
-
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory 11 Oct 2024 · 0 repositories · arXiv:2410.08942