Methods › General › Output Functions › Softmax › Papers, page 88
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 88 of 375: papers 8,701 to 8,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Emotion Recognition with Facial Attention and Objective Activation Functions 23 Oct 2024 · 0 repositories · arXiv:2410.17740
-
Escaping the Forest: Sparse Interpretable Neural Networks for Tabular Data 23 Oct 2024 · 1 repository · arXiv:2410.17758
-
Federated Transformer: Multi-Party Vertical Federated Learning on Practical Fuzzily Linked Data 23 Oct 2024 · 1 repository · arXiv:2410.17986Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
FIPER: Generalizable Factorized Features for Robust Low-Level Vision Models 23 Oct 2024 · 0 repositories · arXiv:2410.18083
-
From PDFs to Structured Data: Utilizing LLM Analysis in Sports Database Management 23 Oct 2024 · 0 repositories · arXiv:2410.17619
-
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction 23 Oct 2024 · 0 repositories · arXiv:2410.18160
-
Gazelle: An Instruction Dataset for Arabic Writing Assistance 23 Oct 2024 · 0 repositories · arXiv:2410.18163
-
GenDP: 3D Semantic Fields for Category-Level Generalizable Diffusion Policy 23 Oct 2024 · 0 repositories · arXiv:2410.17488
-
HCDN: A Change Detection Network for Construction Housekeeping Using Feature Fusion and Large Vision Models 23 Oct 2024 · 1 repository · arXiv:2410.17513
-
Holistic structure of neural pathways underlies brain perceptual rivalry: Physical mechanism of auditory stream segregation 23 Oct 2024 · 0 repositories · arXiv:2410.17620
-
How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization? 23 Oct 2024 · 1 repository · arXiv:2410.17594
-
Leveraging the Domain Adaptation of Retrieval Augmented Generation Models for Question Answering and Reducing Hallucination 23 Oct 2024 · 0 repositories · arXiv:2410.17783
-
Lightweight Neural App Control 23 Oct 2024 · 0 repositories · arXiv:2410.17883
-
Small Singular Values Matter: A Random Matrix Analysis of Transformer Models 23 Oct 2024 · 0 repositories · arXiv:2410.17770
-
LongRAG: A Dual-Perspective Retrieval-Augmented Generation Paradigm for Long-Context Question Answering 23 Oct 2024 · 1 repository · arXiv:2410.18050Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
MCUBERT: Memory-Efficient BERT Inference on Commodity Microcontrollers 23 Oct 2024 · 0 repositories · arXiv:2410.17957
-
MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models 23 Oct 2024 · 1 repository · arXiv:2410.17637Syntology official (archive's flag): 8 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
MiLoRA: Efficient Mixture of Low-Rank Adaptation for Large Language Models Fine-tuning 23 Oct 2024 · 0 repositories · arXiv:2410.18035
-
MojoBench: Language Modeling and Benchmarks for Mojo 23 Oct 2024 · 0 repositories · arXiv:2410.17736
-
Multi-scale feature reconstruction network for industrial anomaly detection 23 Oct 2024 · 1 repository
-
NexusIndex: Integrating Advanced Vector Indexing and Multi-Model Embeddings for Robust Fake News Detection 23 Oct 2024 · 0 repositories · arXiv:2410.18294
-
OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation 23 Oct 2024 · 1 repository · arXiv:2410.17799
-
PETAH: Parameter Efficient Task Adaptation for Hybrid Transformers in a resource-limited Context 23 Oct 2024 · 0 repositories · arXiv:2410.17661
-
PGDiffSeg: Prior-Guided Denoising Diffusion Model with Parameter-Shared Attention for Breast Cancer Segmentation 23 Oct 2024 · 0 repositories · arXiv:2410.17812
-
POD-Attention: Unlocking Full Prefill-Decode Overlap for Faster LLM Inference 23 Oct 2024 · 1 repository · arXiv:2410.18038
-
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains 23 Oct 2024 · 0 repositories · arXiv:2410.17952
-
Spiking Graph Neural Network on Riemannian Manifolds 23 Oct 2024 · 1 repository · arXiv:2410.17941Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Scaling Stick-Breaking Attention: An Efficient Implementation and In-depth Study 23 Oct 2024 · 2 repositories · arXiv:2410.17980Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Surgical Scene Segmentation by Transformer With Asymmetric Feature Enhancement 23 Oct 2024 · 1 repository · arXiv:2410.17642
-
TabDPT: Scaling Tabular Foundation Models 23 Oct 2024 · 1 repository · arXiv:2410.18164Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
TAGE: Trustworthy Attribute Group Editing for Stable Few-shot Image Generation 23 Oct 2024 · 0 repositories · arXiv:2410.17855
-
TP-Eval: Tap Multimodal LLMs' Potential in Evaluation by Customizing Prompts 23 Oct 2024 · 0 repositories · arXiv:2410.18071
-
TranSPORTmer: A Holistic Approach to Trajectory Understanding in Multi-Agent Sports 23 Oct 2024 · 0 repositories · arXiv:2410.17785
-
Value Residual Learning For Alleviating Attention Concentration In Transformers 23 Oct 2024 · 1 repository · arXiv:2410.17897Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Which Client is Reliable?: A Reliable and Personalized Prompt-based Federated Learning for Medical Image Question Answering 23 Oct 2024 · 0 repositories · arXiv:2410.17484
-
A Bayesian Perspective on the Maximum Score Problem 22 Oct 2024 · 0 repositories · arXiv:2410.17153
-
A Statistical Analysis of LLMs' Self-Evaluation Using Proverbs 22 Oct 2024 · 0 repositories · arXiv:2410.16640
-
An Eye for an AI: Evaluating GPT-4o's Visual Perception Skills and Geometric Reasoning Skills Using Computer Graphics Questions 22 Oct 2024 · 0 repositories · arXiv:2410.16991
-
Assessment of Transformer-Based Encoder-Decoder Model for Human-Like Summarization 22 Oct 2024 · 0 repositories · arXiv:2410.16842
-
Audio-to-Score Conversion Model Based on Whisper methodology 22 Oct 2024 · 0 repositories · arXiv:2410.17209
-
Automated Spinal MRI Labelling from Reports Using a Large Language Model 22 Oct 2024 · 1 repository · arXiv:2410.17235
-
Captions Speak Louder than Images (CASLIE): Generalizing Foundation Models for E-commerce from High-quality Multimodal Instruction Data 22 Oct 2024 · 0 repositories · arXiv:2410.17337
-
Dhoroni: Exploring Bengali Climate Change and Environmental Views with a Multi-Perspective News Dataset and Natural Language Processing 22 Oct 2024 · 1 repository · arXiv:2410.17225
-
DI-MaskDINO: A Joint Object Detection and Instance Segmentation Model 22 Oct 2024 · 1 repository · arXiv:2410.16707
-
Distill-SynthKG: Distilling Knowledge Graph Synthesis Workflow for Improved Coverage and Efficiency 22 Oct 2024 · 0 repositories · arXiv:2410.16597
-
DNAHLM -- DNA sequence and Human Language mixed large language Model 22 Oct 2024 · 1 repository · arXiv:2410.16917
-
Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference Under Ambiguities 22 Oct 2024 · 1 repository · arXiv:2410.17385Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Dynamic graph neural networks for enhanced volatility prediction in financial markets 22 Oct 2024 · 0 repositories · arXiv:2410.16858
-
Efficient Feature Extraction Using Light-Weight CNN Attention-Based Deep Learning Architectures for Ultrasound Fetal Plane Classification 22 Oct 2024 · 0 repositories · arXiv:2410.17396
-
End-to-End Optimization and Learning of Fair Court Schedules 22 Oct 2024 · 0 repositories · arXiv:2410.17415
-
EnvBridge: Bridging Diverse Environments with Cross-Environment Knowledge Transfer for Embodied AI 22 Oct 2024 · 0 repositories · arXiv:2410.16919
-
Exploring Possibilities of AI-Powered Legal Assistance in Bangladesh through Large Language Modeling 22 Oct 2024 · 1 repository · arXiv:2410.17210
-
FairLoRA: Unpacking Bias Mitigation in Vision Models with Fairness-Driven Low-Rank Adaptation 22 Oct 2024 · 0 repositories · arXiv:2410.17358
-
FastAttention: Extend FlashAttention2 to NPUs and Low-resource GPUs 22 Oct 2024 · 0 repositories · arXiv:2410.16663
-
From Attention to Activation: Unravelling the Enigmas of Large Language Models 22 Oct 2024 · 0 repositories · arXiv:2410.17174
-
Graph Transformers Dream of Electric Flow 22 Oct 2024 · 0 repositories · arXiv:2410.16699
-
High-Order Associative Learning Based on Memristive Circuits for Efficient Learning 22 Oct 2024 · 1 repository · arXiv:2410.16734
-
In Context Learning and Reasoning for Symbolic Regression with Large Language Models 22 Oct 2024 · 1 repository · arXiv:2410.17448
-
Interchangeable Token Embeddings for Extendable Vocabulary and Alpha-Equivalence 22 Oct 2024 · 0 repositories · arXiv:2410.17161
-
LiNo: Advancing Recursive Residual Decomposition of Linear and Nonlinear Patterns for Robust Time Series Forecasting 22 Oct 2024 · 1 repository · arXiv:2410.17159
-
Methods of improving LLM training stability 22 Oct 2024 · 0 repositories · arXiv:2410.16682
-
Optimal Design for Reward Modeling in RLHF 22 Oct 2024 · 0 repositories · arXiv:2410.17055
-
Order Matters: Exploring Order Sensitivity in Multimodal Large Language Models 22 Oct 2024 · 0 repositories · arXiv:2410.16983
-
PLDR-LLM: Large Language Model from Power Law Decoder Representations 22 Oct 2024 · 2 repositories · arXiv:2410.16703
-
Representation Shattering in Transformers: A Synthetic Study with Knowledge Editing 22 Oct 2024 · 0 repositories · arXiv:2410.17194Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Scattered Forest Search: Smarter Code Space Exploration with LLMs 22 Oct 2024 · 0 repositories · arXiv:2411.05010
-
SmartRAG: Jointly Learn RAG-Related Tasks From the Environment Feedback 22 Oct 2024 · 0 repositories · arXiv:2410.18141
-
SpikMamba: When SNN meets Mamba in Event-based Human Action Recognition 22 Oct 2024 · 1 repository · arXiv:2410.16746
-
Survival of the Fittest: Evolutionary Adaptation of Policies for Environmental Shifts 22 Oct 2024 · 0 repositories · arXiv:2410.19852
-
Tracing the Development of the Virtual Particle Concept Using Semantic Change Detection 22 Oct 2024 · 1 repository · arXiv:2410.16855
-
A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM 21 Oct 2024 · 0 repositories · arXiv:2410.15549
-
A Fusion-Driven Approach of Attention-Based CNN-BiLSTM for Protein Family Classification -- ProFamNet 21 Oct 2024 · 0 repositories · arXiv:2410.17293
-
A Troublemaker with Contagious Jailbreak Makes Chaos in Honest Towns 21 Oct 2024 · 0 repositories · arXiv:2410.16155
-
AlignVSR: Audio-Visual Cross-Modal Alignment for Visual Speech Recognition 21 Oct 2024 · 1 repository · arXiv:2410.16438
-
All You Need is an Improving Column: Enhancing Column Generation for Parallel Machine Scheduling via Transformers 21 Oct 2024 · 0 repositories · arXiv:2410.15601
-
An Efficient System for Automatic Map Storytelling -- A Case Study on Historical Maps 21 Oct 2024 · 1 repository · arXiv:2410.15780
-
An Explainable Contrastive-based Dilated Convolutional Network with Transformer for Pediatric Pneumonia Detection 21 Oct 2024 · 0 repositories · arXiv:2410.16143
-
Arithmetic Transformers Can Length-Generalize in Both Operand Length and Count 21 Oct 2024 · 1 repository · arXiv:2410.15787
-
Beyond 2:4: exploring V:N:M sparsity for efficient transformer inference on GPUs 21 Oct 2024 · 0 repositories · arXiv:2410.16135
-
Boosting Jailbreak Transferability for Large Language Models 21 Oct 2024 · 1 repository · arXiv:2410.15645
-
Building A Coding Assistant via the Retrieval-Augmented Language Model 21 Oct 2024 · 1 repository · arXiv:2410.16229
-
Calibration of ordinal regression networks 21 Oct 2024 · 0 repositories · arXiv:2410.15658
-
CamI2V: Camera-Controlled Image-to-Video Diffusion Model 21 Oct 2024 · 1 repository · arXiv:2410.15957Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
CartesianMoE: Boosting Knowledge Sharing among Experts via Cartesian Product Routing in Mixture-of-Experts 21 Oct 2024 · 1 repository · arXiv:2410.16077
-
CausalGraph2LLM: Evaluating LLMs for Causal Queries 21 Oct 2024 · 1 repository · arXiv:2410.15939
-
CompassJudger-1: All-in-one Judge Model Helps Model Evaluation and Evolution 21 Oct 2024 · 1 repository · arXiv:2410.16256
-
Deep Graph Attention Networks 21 Oct 2024 · 1 repository · arXiv:2410.15640
-
Deep Learning and Data Augmentation for Detecting Self-Admitted Technical Debt 21 Oct 2024 · 1 repository · arXiv:2410.15804
-
Developing Retrieval Augmented Generation (RAG) based LLM Systems from PDFs: An Experience Report 21 Oct 2024 · 1 repository · arXiv:2410.15944
-
Diffusion Transformer Policy 21 Oct 2024 · 1 repository · arXiv:2410.15959
-
Disambiguating Monocular Reconstruction of 3D Clothed Human with Spatial-Temporal Transformer 21 Oct 2024 · 0 repositories · arXiv:2410.16337
-
Catastrophic Failure of LLM Unlearning via Quantization 21 Oct 2024 · 1 repository · arXiv:2410.16454Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Domain-Adaptive Pre-training of Self-Supervised Foundation Models for Medical Image Classification in Gastrointestinal Endoscopy 21 Oct 2024 · 1 repository · arXiv:2410.21302
-
Enabling Energy-Efficient Deployment of Large Language Models on Memristor Crossbar: A Synergy of Large and Small 21 Oct 2024 · 0 repositories · arXiv:2410.15977
-
Enhancing SNN-based Spatio-Temporal Learning: A Benchmark Dataset and Cross-Modality Attention Model 21 Oct 2024 · 1 repository · arXiv:2410.15689
-
ExDBN: Exact learning of Dynamic Bayesian Networks 21 Oct 2024 · 0 repositories · arXiv:2410.16100
-
Exploring Pretraining via Active Forgetting for Improving Cross Lingual Transfer for Decoder Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16168
-
Focus Where It Matters: Graph Selective State Focused Attention Networks 21 Oct 2024 · 0 repositories · arXiv:2410.15849
-
FusionLungNet: Multi-scale Fusion Convolution with Refinement Network for Lung CT Image Segmentation 21 Oct 2024 · 2 repositories · arXiv:2410.15812
-
Generalized Probabilistic Attention Mechanism in Transformers 21 Oct 2024 · 0 repositories · arXiv:2410.15578