Methods › General › Output Functions › Softmax › Papers, page 159
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 159 of 375: papers 15,801 to 15,900 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Amharic LLaMA and LLaVA: Multimodal LLMs for Low Resource Languages 11 Mar 2024 · 1 repository · arXiv:2403.06354
-
Car Damage Detection and Patch-to-Patch Self-supervised Image Alignment 11 Mar 2024 · 1 repository · arXiv:2403.06674Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Development of a Reliable and Accessible Caregiving Language Model (CaLM) 11 Mar 2024 · 0 repositories · arXiv:2403.06857
-
DNGaussian: Optimizing Sparse-View 3D Gaussian Radiance Fields with Global-Local Depth Normalization 11 Mar 2024 · 1 repository · arXiv:2403.06912Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
GRITv2: Efficient and Light-weight Social Relation Recognition 11 Mar 2024 · 0 repositories · arXiv:2403.06895
-
Guiding Clinical Reasoning with Large Language Models via Knowledge Seeds 11 Mar 2024 · 0 repositories · arXiv:2403.06609
-
HDRTransDC: High Dynamic Range Image Reconstruction with Transformer Deformation Convolution 11 Mar 2024 · 0 repositories · arXiv:2403.06831
-
Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena 11 Mar 2024 · 0 repositories · arXiv:2403.06965
-
In-context Exploration-Exploitation for Reinforcement Learning 11 Mar 2024 · 0 repositories · arXiv:2403.06826
-
Multi-Scale Implicit Transformer with Re-parameterize for Arbitrary-Scale Super-Resolution 11 Mar 2024 · 0 repositories · arXiv:2403.06536
-
Multilingual Turn-taking Prediction Using Voice Activity Projection 11 Mar 2024 · 0 repositories · arXiv:2403.06487
-
Narrating Causal Graphs with Large Language Models 11 Mar 2024 · 0 repositories · arXiv:2403.07118
-
One size doesn't fit all: Predicting the Number of Examples for In-Context Learning 11 Mar 2024 · 0 repositories · arXiv:2403.06402
-
SMART: Automatically Scaling Down Language Models with Accuracy Guarantees for Reduced Processing Fees 11 Mar 2024 · 1 repository · arXiv:2403.13835
-
The pitfalls of next-token prediction 11 Mar 2024 · 1 repository · arXiv:2403.06963Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Unraveling the Mystery of Scaling Laws: Part I 11 Mar 2024 · 0 repositories · arXiv:2403.06563
-
Attacking Transformers with Feature Diversity Adversarial Perturbation 10 Mar 2024 · 0 repositories · arXiv:2403.07942
-
Attention is all you need for boosting graph convolutional neural network 10 Mar 2024 · 0 repositories · arXiv:2403.15419
-
Finding Visual Saliency in Continuous Spike Stream 10 Mar 2024 · 1 repository · arXiv:2403.06233Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
FrameQuant: Flexible Low-Bit Quantization for Transformers 10 Mar 2024 · 1 repository · arXiv:2403.06082Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Multisize Dataset Condensation 10 Mar 2024 · 1 repository · arXiv:2403.06075Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 19 harvested samples) · 15 pointer-only (licence)
-
Target-constrained Bidirectional Planning for Generation of Target-oriented Proactive Dialogue 10 Mar 2024 · 1 repository · arXiv:2403.06063
-
Towards In-Vehicle Multi-Task Facial Attribute Recognition: Investigating Synthetic Data and Vision Foundation Models 10 Mar 2024 · 0 repositories · arXiv:2403.06088
-
An Audio-textual Diffusion Model For Converting Speech Signals Into Ultrasound Tongue Imaging Data 9 Mar 2024 · 0 repositories · arXiv:2403.05820
-
AutoEval Done Right: Using Synthetic Data for Model Evaluation 9 Mar 2024 · 1 repository · arXiv:2403.07008Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ClinicalMamba: A Generative Clinical Language Model on Longitudinal Clinical Notes 9 Mar 2024 · 1 repository · arXiv:2403.05795Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Enhancing Multi-Hop Knowledge Graph Reasoning through Reward Shaping Techniques 9 Mar 2024 · 0 repositories · arXiv:2403.05801
-
General surgery vision transformer: A video pre-trained foundation model for general surgery 9 Mar 2024 · 1 repository · arXiv:2403.05949
-
TokenMark: A Modality-Agnostic Watermark for Pre-trained Transformers 9 Mar 2024 · 0 repositories · arXiv:2403.05842
-
Long-term Frame-Event Visual Tracking: Benchmark Dataset and Baseline 9 Mar 2024 · 4 repositories · arXiv:2403.05839
-
Segmentation Guided Sparse Transformer for Under-Display Camera Image Restoration 9 Mar 2024 · 0 repositories · arXiv:2403.05906
-
A Dataset and Benchmark for Hospital Course Summarization with Adapted Large Language Models 8 Mar 2024 · 1 repository · arXiv:2403.05720
-
A Novel Nuanced Conversation Evaluation Framework for Large Language Models in Mental Health 8 Mar 2024 · 0 repositories · arXiv:2403.09705
-
ActFormer: Scalable Collaborative Perception via Active Queries 8 Mar 2024 · 0 repositories · arXiv:2403.04968
-
An In-depth Evaluation of GPT-4 in Sentence Simplification with Error-based Human Assessment 8 Mar 2024 · 0 repositories · arXiv:2403.04963
-
Are Large Language Models Aligned with People's Social Intuitions for Human-Robot Interactions? 8 Mar 2024 · 1 repository · arXiv:2403.05701
-
Bias-Augmented Consistency Training Reduces Biased Reasoning in Chain-of-Thought 8 Mar 2024 · 1 repository · arXiv:2403.05518Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Binaural Speech Enhancement Using Deep Complex Convolutional Transformer Networks 8 Mar 2024 · 1 repository · arXiv:2403.05393
-
Can't Remember Details in Long Documents? You Need Some R&R 8 Mar 2024 · 1 repository · arXiv:2403.05004Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
CommitBench: A Benchmark for Commit Message Generation 8 Mar 2024 · 1 repository · arXiv:2403.05188
-
Considering Nonstationary within Multivariate Time Series with Variational Hierarchical Transformer for Forecasting 8 Mar 2024 · 1 repository · arXiv:2403.05406Syntology official (archive's flag): 12 ran · 12 ran (of which 10 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs 8 Mar 2024 · 0 repositories · arXiv:2403.05434
-
How Well Do Multi-modal LLMs Interpret CT Scans? An Auto-Evaluation Framework for Analyses 8 Mar 2024 · 0 repositories · arXiv:2403.05680
-
Denoising Autoregressive Representation Learning 8 Mar 2024 · 0 repositories · arXiv:2403.05196
-
DualBEV: Unifying Dual View Transformation with Probabilistic Correspondences 8 Mar 2024 · 1 repository · arXiv:2403.05402Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Enhancing Automatic Modulation Recognition for IoT Applications Using Transformers 8 Mar 2024 · 0 repositories · arXiv:2403.15417
-
ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models 8 Mar 2024 · 1 repository · arXiv:2403.05266Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context 8 Mar 2024 · 1 repository · arXiv:2403.05530
-
Hybridized Convolutional Neural Networks and Long Short-Term Memory for Improved Alzheimer's Disease Diagnosis from MRI Scans 8 Mar 2024 · 0 repositories · arXiv:2403.05353
-
Inverse Design of Photonic Crystal Surface Emitting Lasers is a Sequence Modeling Problem 8 Mar 2024 · 0 repositories · arXiv:2403.05149
-
JointMotion: Joint Self-Supervision for Joint Motion Prediction 8 Mar 2024 · 1 repository · arXiv:2403.05489
-
LightM-UNet: Mamba Assists in Lightweight UNet for Medical Image Segmentation 8 Mar 2024 · 1 repository · arXiv:2403.05246Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
LLM4Decompile: Decompiling Binary Code with Large Language Models 8 Mar 2024 · 1 repository · arXiv:2403.05286Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
MamMIL: Multiple Instance Learning for Whole Slide Images with State Space Models 8 Mar 2024 · 1 repository · arXiv:2403.05160
-
Med3DInsight: Enhancing 3D Medical Image Understanding with 2D Multi-Modal Large Language Models 8 Mar 2024 · 0 repositories · arXiv:2403.05141
-
PipeRAG: Fast Retrieval-Augmented Generation via Algorithm-System Co-design 8 Mar 2024 · 0 repositories · arXiv:2403.05676
-
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation 8 Mar 2024 · 1 repository · arXiv:2403.05313
-
Rule-driven News Captioning 8 Mar 2024 · 0 repositories · arXiv:2403.05101
-
Self-Supervised Multiple Instance Learning for Acute Myeloid Leukemia Classification 8 Mar 2024 · 0 repositories · arXiv:2403.05379
-
SIRST-5K: Exploring Massive Negatives Synthesis with Self-supervised Learning for Robust Infrared Small Target Detection 8 Mar 2024 · 1 repository · arXiv:2403.05416
-
Spatial-aware Transformer-GRU Framework for Enhanced Glaucoma Diagnosis from 3D OCT Imaging 8 Mar 2024 · 1 repository · arXiv:2403.05702
-
Text-to-Audio Generation Synchronized with Videos 8 Mar 2024 · 0 repositories · arXiv:2403.07938
-
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers 8 Mar 2024 · 0 repositories · arXiv:2403.05365
-
Will GPT-4 Run DOOM? 8 Mar 2024 · 0 repositories · arXiv:2403.05468
-
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers 7 Mar 2024 · 0 repositories · arXiv:2403.04200
-
Aligning GPTRec with Beyond-Accuracy Goals with Reinforcement Learning 7 Mar 2024 · 1 repository · arXiv:2403.04875
-
AO-DETR: Anti-Overlapping DETR for X-Ray Prohibited Items Detection 7 Mar 2024 · 1 repository · arXiv:2403.04309
-
Attempt Towards Stress Transfer in Speech-to-Speech Machine Translation 7 Mar 2024 · 0 repositories · arXiv:2403.04178
-
AUFormer: Vision Transformers are Parameter-Efficient Facial Action Unit Detectors 7 Mar 2024 · 1 repository · arXiv:2403.04697
-
Automating the Information Extraction from Semi-Structured Interview Transcripts 7 Mar 2024 · 1 repository · arXiv:2403.04819
-
Delving into the Trajectory Long-tail Distribution for Muti-object Tracking 7 Mar 2024 · 1 repository · arXiv:2403.04700Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Disentangled Diffusion-Based 3D Human Pose Estimation with Hierarchical Spatial and Temporal Denoiser 7 Mar 2024 · 1 repository · arXiv:2403.04444
-
Federated Recommendation via Hybrid Retrieval Augmented Generation 7 Mar 2024 · 1 repository · arXiv:2403.04256
-
Feedback-Generation for Programming Exercises With GPT-4 7 Mar 2024 · 0 repositories · arXiv:2403.04449
-
HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild 7 Mar 2024 · 1 repository · arXiv:2403.04307
-
LLMs in the Imaginarium: Tool Learning through Simulated Trial and Error 7 Mar 2024 · 1 repository · arXiv:2403.04746Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
PixArt-Σ: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation 7 Mar 2024 · 2 repositories · arXiv:2403.04692Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
RATSF: Empowering Customer Service Volume Management through Retrieval-Augmented Time-Series Forecasting 7 Mar 2024 · 0 repositories · arXiv:2403.04180
-
Self-Evaluation of Large Language Model based on Glass-box Features 7 Mar 2024 · 1 repository · arXiv:2403.04222
-
Speech Emotion Recognition Via CNN-Transformer and Multidimensional Attention Mechanism 7 Mar 2024 · 1 repository · arXiv:2403.04743
-
Telecom Language Models: Must They Be Large? 7 Mar 2024 · 0 repositories · arXiv:2403.04666
-
Yi: Open Foundation Models by 01.AI 7 Mar 2024 · 1 repository · arXiv:2403.04652Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples)
-
Assessing the Aesthetic Evaluation Capabilities of GPT-4 with Vision: Insights from Group and Individual Assessments 6 Mar 2024 · 0 repositories · arXiv:2403.03594
-
Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem 6 Mar 2024 · 1 repository · arXiv:2403.03558
-
Can Large Language Models do Analytical Reasoning? 6 Mar 2024 · 0 repositories · arXiv:2403.04031
-
Design of an Open-Source Architecture for Neural Machine Translation 6 Mar 2024 · 0 repositories · arXiv:2403.03582
-
Designing Informative Metrics for Few-Shot Example Selection 6 Mar 2024 · 0 repositories · arXiv:2403.03861
-
Enhancing ASD detection accuracy: a combined approach of machine learning and deep learning models with natural language processing 6 Mar 2024 · 1 repository · arXiv:2403.03581
-
Enhancing Price Prediction in Cryptocurrency Using Transformer Neural Network and Technical Indicators 6 Mar 2024 · 0 repositories · arXiv:2403.03606
-
FaaF: Facts as a Function for the evaluation of generated text 6 Mar 2024 · 1 repository · arXiv:2403.03888
-
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection 6 Mar 2024 · 3 repositories · arXiv:2403.03507
-
General2Specialized LLMs Translation for E-commerce 6 Mar 2024 · 0 repositories · arXiv:2403.03689
-
Guiding Enumerative Program Synthesis with Large Language Models 6 Mar 2024 · 0 repositories · arXiv:2403.03997
-
Inverse-Free Fast Natural Gradient Descent Method for Deep Learning 6 Mar 2024 · 0 repositories · arXiv:2403.03473
-
Investigation of the Impact of Synthetic Training Data in the Industrial Application of Terminal Strip Object Detection 6 Mar 2024 · 0 repositories · arXiv:2403.04809
-
Japanese-English Sentence Translation Exercises Dataset for Automatic Grading 6 Mar 2024 · 0 repositories · arXiv:2403.03396
-
Joint multi-task learning improves weakly-supervised biomarker prediction in computational pathology 6 Mar 2024 · 1 repository · arXiv:2403.03891
-
LDSF: Lightweight Dual-Stream Framework for SAR Target Recognition by Coupling Local Electromagnetic Scattering Features and Global Visual Features 6 Mar 2024 · 0 repositories · arXiv:2403.03527
-
Multi-modal Deep Learning 6 Mar 2024 · 0 repositories · arXiv:2403.03385
-
PPTC-R benchmark: Towards Evaluating the Robustness of Large Language Models for PowerPoint Task Completion 6 Mar 2024 · 1 repository · arXiv:2403.03788