Methods › General › Output Functions › Softmax › Papers, page 167
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 167 of 375: papers 16,601 to 16,700 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
NeRCC: Nested-Regression Coded Computing for Resilient Distributed Prediction Serving Systems 6 Feb 2024 · 0 repositories · arXiv:2402.04377
-
The Hedgehog & the Porcupine: Expressive Linear Attentions with Softmax Mimicry 6 Feb 2024 · 1 repository · arXiv:2402.04347
-
The Use of a Large Language Model for Cyberbullying Detection 6 Feb 2024 · 0 repositories · arXiv:2402.04088
-
Training Language Models to Generate Text with Citations via Fine-grained Rewards 6 Feb 2024 · 1 repository · arXiv:2402.04315Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
U-shaped Vision Mamba for Single Image Dehazing 6 Feb 2024 · 1 repository · arXiv:2402.04139
-
A Survey on Transformer Compression 5 Feb 2024 · 0 repositories · arXiv:2402.05964
-
Accurate and Well-Calibrated ICD Code Assignment Through Attention Over Diverse Label Embeddings 5 Feb 2024 · 1 repository · arXiv:2402.03172Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Arabic Synonym BERT-based Adversarial Examples for Text Classification 5 Feb 2024 · 1 repository · arXiv:2402.03477
-
C-RAG: Certified Generation Risks for Retrieval-Augmented Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03181Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Reconstruct Your Previous Conversations! Comprehensively Investigating Privacy Leakage Risks in Conversations with GPT Models 5 Feb 2024 · 1 repository · arXiv:2402.02987
-
Cross-Domain Few-Shot Object Detection via Enhanced Open-Set Object Detector 5 Feb 2024 · 2 repositories · arXiv:2402.03094Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models 5 Feb 2024 · 5 repositories · arXiv:2402.03300Syntology official (archive's flag): 8 ran · 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 10 unverified (of 24 harvested samples) · 3 pointer-only (licence)
-
DiffsFormer: A Diffusion Transformer on Stock Factor Augmentation 5 Feb 2024 · 0 repositories · arXiv:2402.06656
-
Enhancing textual textbook question answering with large language models and retrieval augmented generation 5 Feb 2024 · 1 repository · arXiv:2402.05128Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Financial Report Chunking for Effective Retrieval Augmented Generation 5 Feb 2024 · 1 repository · arXiv:2402.05131
-
Focal Modulation Networks for Interpretable Sound Classification 5 Feb 2024 · 0 repositories · arXiv:2402.02754
-
Graph-enhanced Large Language Models in Asynchronous Plan Reasoning 5 Feb 2024 · 1 repository · arXiv:2402.02805Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
HAMLET: Graph Transformer Neural Operator for Partial Differential Equations 5 Feb 2024 · 0 repositories · arXiv:2402.03541
-
Harnessing PubMed User Query Logs for Post Hoc Explanations of Recommended Similar Articles 5 Feb 2024 · 0 repositories · arXiv:2402.03484
-
Illuminate: A novel approach for depression detection with explainable analysis and proactive therapy using prompt engineering 5 Feb 2024 · 0 repositories · arXiv:2402.05127
-
Is Mamba Capable of In-Context Learning? 5 Feb 2024 · 1 repository · arXiv:2402.03170Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Just Cluster It: An Approach for Exploration in High-Dimensions using Clustering and Pre-Trained Representations 5 Feb 2024 · 1 repository · arXiv:2402.03138Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
LB-KBQA: Large-language-model and BERT based Knowledge-Based Question and Answering System 5 Feb 2024 · 0 repositories · arXiv:2402.05130
-
LLM Agents in Interaction: Measuring Personality Consistency and Linguistic Alignment in Interacting Populations of Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.02896
-
MobilityGPT: Enhanced Human Mobility Modeling with a GPT model 5 Feb 2024 · 0 repositories · arXiv:2402.03264
-
Multi-Lingual Malaysian Embedding: Leveraging Large Language Models for Semantic Representations 5 Feb 2024 · 0 repositories · arXiv:2402.03053
-
On Least Square Estimation in Softmax Gating Mixture of Experts 5 Feb 2024 · 0 repositories · arXiv:2402.02952
-
SWAG: Storytelling With Action Guidance 5 Feb 2024 · 1 repository · arXiv:2402.03483
-
Time-, Memory- and Parameter-Efficient Visual Adaptation 5 Feb 2024 · 0 repositories · arXiv:2402.02887Syntology 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Toward Human-AI Alignment in Large-Scale Multi-Player Games 5 Feb 2024 · 0 repositories · arXiv:2402.03575
-
Towards Understanding the Word Sensitivity of Attention Layers: A Study via Random Features 5 Feb 2024 · 1 repository · arXiv:2402.02969Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
UniMem: Towards a Unified View of Long-Context Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03009
-
A flexible Bayesian g-formula for causal survival analyses with time-dependent confounding 4 Feb 2024 · 1 repository · arXiv:2402.02306
-
A Graph is Worth K Words: Euclideanizing Graph using Pure Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02464Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Aligner: Efficient Alignment by Learning to Correct 4 Feb 2024 · 0 repositories · arXiv:2402.02416
-
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models 4 Feb 2024 · 1 repository · arXiv:2402.02370
-
Breaking MLPerf Training: A Case Study on Optimizing BERT 4 Feb 2024 · 0 repositories · arXiv:2402.02447
-
Evaluating Large Language Models in Analysing Classroom Dialogue 4 Feb 2024 · 0 repositories · arXiv:2402.02380
-
GeReA: Question-Aware Prompt Captions for Knowledge-based Visual Question Answering 4 Feb 2024 · 1 repository · arXiv:2402.02503Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation 4 Feb 2024 · 0 repositories · arXiv:2402.14594
-
INViT: A Generalizable Routing Problem Solver with Invariant Nested View Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02317Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Key-Graph Transformer for Image Restoration 4 Feb 2024 · 0 repositories · arXiv:2402.02634
-
Minusformer: Improving Time Series Forecasting by Progressively Learning Residuals 4 Feb 2024 · 1 repository · arXiv:2402.02332
-
Pathformer: Multi-scale Transformers with Adaptive Pathways for Time Series Forecasting 4 Feb 2024 · 1 repository · arXiv:2402.05956Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
PROSAC: Provably Safe Certification for Machine Learning Models under Adversarial Attacks 4 Feb 2024 · 0 repositories · arXiv:2402.02629
-
SPCTNet: A Series-Parallel CNN and Transformer Network for 3D Medical Image Segmentation 4 Feb 2024 · 0 repositories
-
Spin: An Efficient Secure Computation Framework with GPU Acceleration 4 Feb 2024 · 0 repositories · arXiv:2402.02320
-
Timer: Generative Pre-trained Transformers Are Large Time Series Models 4 Feb 2024 · 1 repository · arXiv:2402.02368Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Unified Training of Universal Time Series Forecasting Transformers 4 Feb 2024 · 1 repository · arXiv:2402.02592
-
3D Lymphoma Segmentation on PET/CT Images via Multi-Scale Information Fusion with Cross-Attention 4 Feb 2024 · 0 repositories · arXiv:2402.02349
-
BetterV: Controlled Verilog Generation with Discriminative Guidance 3 Feb 2024 · 0 repositories · arXiv:2402.03375
-
CoFiNet: Unveiling Camouflaged Objects with Multi-Scale Finesse 3 Feb 2024 · 0 repositories · arXiv:2402.02217
-
Data Quality Matters: Suicide Intention Detection on Social Media Posts Using RoBERTa-CNN 3 Feb 2024 · 0 repositories · arXiv:2402.02262
-
DE³-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks 3 Feb 2024 · 0 repositories · arXiv:2402.05948
-
DiffVein: A Unified Diffusion Network for Finger Vein Segmentation and Authentication 3 Feb 2024 · 0 repositories · arXiv:2402.02060
-
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test 3 Feb 2024 · 0 repositories · arXiv:2402.02135
-
EffiBench: Benchmarking the Efficiency of Automatically Generated Code 3 Feb 2024 · 1 repository · arXiv:2402.02037Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
How well do LLMs cite relevant medical references? An evaluation framework and analyses 3 Feb 2024 · 1 repository · arXiv:2402.02008
-
IMUSE: IMU-based Facial Expression Capture 3 Feb 2024 · 0 repositories · arXiv:2402.03944
-
ParZC: Parametric Zero-Cost Proxies for Efficient NAS 3 Feb 2024 · 0 repositories · arXiv:2402.02105
-
ScribFormer: Transformer Makes CNN Work Better for Scribble-based Medical Image Segmentation 3 Feb 2024 · 1 repository · arXiv:2402.02029
-
TCI-Former: Thermal Conduction-Inspired Transformer for Infrared Small Target Detection 3 Feb 2024 · 0 repositories · arXiv:2402.02046
-
Topology-Informed Graph Transformer 3 Feb 2024 · 2 repositories · arXiv:2402.02005Syntology official (archive's flag): 8 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 2 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Hierarchical Structure Enhances the Convergence and Generalizability of Linear Molecular Representation 3 Feb 2024 · 1 repository · arXiv:2402.02164
-
A Data-Driven Analysis of Robust Automatic Piano Transcription 2 Feb 2024 · 0 repositories · arXiv:2402.01424
-
A Probabilistic Model Behind Self-Supervised Learning 2 Feb 2024 · 1 repository · arXiv:2402.01399Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
AGILE: Approach-based Grasp Inference Learned from Element Decomposition 2 Feb 2024 · 0 repositories · arXiv:2402.01303
-
ALERT-Transformer: Bridging Asynchronous and Synchronous Machine Learning for Real-Time Event-based Spatio-Temporal Data 2 Feb 2024 · 0 repositories · arXiv:2402.01393
-
An introduction to graphical tensor notation for mechanistic interpretability 2 Feb 2024 · 0 repositories · arXiv:2402.01790
-
BAT: Learning to Reason about Spatial Sounds with Large Language Models 2 Feb 2024 · 0 repositories · arXiv:2402.01591
-
Clarifying the Path to User Satisfaction: An Investigation into Clarification Usefulness 2 Feb 2024 · 1 repository · arXiv:2402.01934
-
COMET: Generating Commit Messages using Delta Graph Context Representation 2 Feb 2024 · 0 repositories · arXiv:2402.01841
-
ConRF: Zero-shot Stylization of 3D Scenes with Conditioned Radiation Fields 2 Feb 2024 · 1 repository · arXiv:2402.01950
-
Cross-view Masked Diffusion Transformers for Person Image Synthesis 2 Feb 2024 · 1 repository · arXiv:2402.01516Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Can LLMs perform structured graph reasoning? 2 Feb 2024 · 1 repository · arXiv:2402.01805
-
Faster Inference of Integer SWIN Transformer by Removing the GELU Activation 2 Feb 2024 · 0 repositories · arXiv:2402.01169
-
How Can Generative AI Enhance the Well-being of Blind? 2 Feb 2024 · 0 repositories · arXiv:2402.07919
-
Improving Sequential Recommendations with LLMs 2 Feb 2024 · 1 repository · arXiv:2402.01339Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Integrating Large Language Models in Causal Discovery: A Statistical Causal Approach 2 Feb 2024 · 2 repositories · arXiv:2402.01454
-
LLM-Detector: Improving AI-Generated Chinese Text Detection with Open-Source LLM Instruction Tuning 2 Feb 2024 · 1 repository · arXiv:2402.01158
-
LoTR: Low Tensor Rank Weight Adaptation 2 Feb 2024 · 0 repositories · arXiv:2402.01376
-
Predicting ATP binding sites in protein sequences using Deep Learning and Natural Language Processing 2 Feb 2024 · 0 repositories · arXiv:2402.01829
-
Retrieval Augmented End-to-End Spoken Dialog Models 2 Feb 2024 · 0 repositories · arXiv:2402.01828
-
Todyformer: Towards Holistic Dynamic Graph Transformers with Structure-Aware Tokenization 2 Feb 2024 · 0 repositories · arXiv:2402.05944
-
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks 2 Feb 2024 · 0 repositories · arXiv:2402.01176
-
Transformers Learn Nonlinear Features In Context: Nonconvex Mean-field Dynamics on the Attention Landscape 2 Feb 2024 · 0 repositories · arXiv:2402.01258
-
TravelPlanner: A Benchmark for Real-World Planning with Language Agents 2 Feb 2024 · 2 repositories · arXiv:2402.01622Syntology official (archive's flag): 13 ran · 17 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
What Will My Model Forget? Forecasting Forgotten Examples in Language Model Refinement 2 Feb 2024 · 0 repositories · arXiv:2402.01865
-
Theoretical Understanding of In-Context Learning in Shallow Transformers with Unstructured Data 1 Feb 2024 · 0 repositories · arXiv:2402.00743
-
Dendritic Learning-incorporated Vision Transformer for Image Recognition 1 Feb 2024 · 1 repository
-
FairEHR-CLP: Towards Fairness-Aware Clinical Predictions with Contrastive Learning in Multimodal Electronic Health Records 1 Feb 2024 · 0 repositories · arXiv:2402.00955
-
FuseFormer: A Transformer for Visual and Thermal Image Fusion 1 Feb 2024 · 0 repositories · arXiv:2402.00971
-
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model 1 Feb 2024 · 0 repositories · arXiv:2402.01051
-
Hierarchical Multi-Label Classification of Online Vaccine Concerns 1 Feb 2024 · 0 repositories · arXiv:2402.01783
-
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA 1 Feb 2024 · 0 repositories · arXiv:2402.01767
-
Hybrid Quantum Vision Transformers for Event Classification in High Energy Physics 1 Feb 2024 · 0 repositories · arXiv:2402.00776
-
Improving Semantic Control in Discrete Latent Spaces with Transformer Quantized Variational Autoencoders 1 Feb 2024 · 1 repository · arXiv:2402.00723
-
Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing 1 Feb 2024 · 0 repositories · arXiv:2402.00658
-
LVC-LGMC: Joint Local and Global Motion Compensation for Learned Video Compression 1 Feb 2024 · 0 repositories · arXiv:2402.00680
-
Merging Multi-Task Models via Weight-Ensembling Mixture of Experts 1 Feb 2024 · 1 repository · arXiv:2402.00433Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)