Methods › General › Output Functions › Softmax › Papers, page 229
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 229 of 375: papers 22,801 to 22,900 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Commonsense Reasoning for Conversational AI: A Survey of the State of the Art 15 Feb 2023 · 0 repositories · arXiv:2302.07926
-
Confidence Score Based Speaker Adaptation of Conformer Speech Recognition Systems 15 Feb 2023 · 1 repository · arXiv:2302.07521
-
Deep Convolutional Neural Network for Plume Rise Measurements in Industrial Environments 15 Feb 2023 · 0 repositories · arXiv:2302.07416
-
Learning Performance-Improving Code Edits 15 Feb 2023 · 2 repositories · arXiv:2302.07867Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Pose-Oriented Transformer with Uncertainty-Guided Refinement for 2D-to-3D Human Pose Estimation 15 Feb 2023 · 0 repositories · arXiv:2302.07408
-
Self-Supervised Learning for Modeling Gamma-ray Variability in Blazars 15 Feb 2023 · 0 repositories · arXiv:2302.07700
-
Speculative Decoding with Big Little Decoder 15 Feb 2023 · 1 repository · arXiv:2302.07863Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
TFormer: A Transmission-Friendly ViT Model for IoT Devices 15 Feb 2023 · 0 repositories · arXiv:2302.07734
-
TiZero: Mastering Multi-Agent Football with Curriculum Learning and Self-Play 15 Feb 2023 · 1 repository · arXiv:2302.07515
-
Towards Optimal Compression: Joint Pruning and Quantization 15 Feb 2023 · 0 repositories · arXiv:2302.07612
-
Tree-Based Representation and Generation of Natural and Mathematical Language 15 Feb 2023 · 1 repository · arXiv:2302.07974
-
Uncertainty-Estimation with Normalized Logits for Out-of-Distribution Detection 15 Feb 2023 · 0 repositories · arXiv:2302.07608
-
A Modern Look at the Relationship between Sharpness and Generalization 14 Feb 2023 · 1 repository · arXiv:2302.07011
-
A Psycholinguistic Analysis of BERT's Representations of Compounds 14 Feb 2023 · 1 repository · arXiv:2302.07232
-
Deep Learning-Based Modeling of 5G Core Control Plane for 5G Network Digital Twin 14 Feb 2023 · 0 repositories · arXiv:2302.06980
-
DiffFashion: Reference-based Fashion Design with Structure-aware Transfer by Diffusion Models 14 Feb 2023 · 1 repository · arXiv:2302.06826
-
Energy Transformer 14 Feb 2023 · 4 repositories · arXiv:2302.07253Syntology official (archive's flag): 11 ran · 13 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Exploring Category Structure with Contextual Language Models and Lexical Semantic Networks 14 Feb 2023 · 0 repositories · arXiv:2302.06942
-
Few-shot learning approaches for classifying low resource domain specific software requirements 14 Feb 2023 · 0 repositories · arXiv:2302.06951
-
PolyFormer: Referring Image Segmentation as Sequential Polygon Generation 14 Feb 2023 · 1 repository · arXiv:2302.07387Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Reveal the Unknown: Out-of-Knowledge-Base Mention Discovery with Entity Linking 14 Feb 2023 · 3 repositories · arXiv:2302.07189
-
ScatterShot: Interactive In-context Example Curation for Text Transformation 14 Feb 2023 · 1 repository · arXiv:2302.07346
-
Team DETR: Guide Queries as a Professional Team in Detection Transformers 14 Feb 2023 · 1 repository · arXiv:2302.07116
-
A Comprehensive Study of Modern Architectures and Regularization Approaches on CheXpert5000 13 Feb 2023 · 0 repositories · arXiv:2302.06684
-
A Study on ReLU and Softmax in Transformer 13 Feb 2023 · 0 repositories · arXiv:2302.06461
-
A Unified View of Long-Sequence Models towards Modeling Million-Scale Dependencies 13 Feb 2023 · 0 repositories · arXiv:2302.06218
-
An Application of Deep Learning for Sweet Cherry Phenotyping using YOLO Object Detection 13 Feb 2023 · 0 repositories · arXiv:2302.06698
-
Anticipating Next Active Objects for Egocentric Videos 13 Feb 2023 · 0 repositories · arXiv:2302.06358
-
Diminished Diversity-of-Thought in a Standard Large Language Model 13 Feb 2023 · 0 repositories · arXiv:2302.07267
-
Can GPT-3 Perform Statutory Reasoning? 13 Feb 2023 · 1 repository · arXiv:2302.06100
-
CholecTriplet2022: Show me a tool and tell me the triplet -- an endoscopic vision challenge for surgical action triplet detection 13 Feb 2023 · 2 repositories · arXiv:2302.06294
-
VITR: Augmenting Vision Transformers with Relation-Focused Learning for Cross-Modal Information Retrieval 13 Feb 2023 · 0 repositories · arXiv:2302.06350
-
Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution Shifts 13 Feb 2023 · 1 repository · arXiv:2302.06495Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Detection and Segmentation of Pancreas using Morphological Snakes and Deep Convolutional Neural Networks 13 Feb 2023 · 0 repositories · arXiv:2302.06356
-
Encoding Sentence Position in Context-Aware Neural Machine Translation with Concatenation 13 Feb 2023 · 1 repository · arXiv:2302.06459
-
Inferring Player Location in Sports Matches: Multi-Agent Spatial Imputation from Limited Observations 13 Feb 2023 · 0 repositories · arXiv:2302.06569
-
Learning-Based Defect Recognitions for Autonomous UAV Inspections 13 Feb 2023 · 1 repository · arXiv:2302.06093
-
Learning to Scale Temperature in Masked Self-Attention for Image Inpainting 13 Feb 2023 · 0 repositories · arXiv:2302.06130
-
Linguistic ambiguity analysis in ChatGPT 13 Feb 2023 · 0 repositories · arXiv:2302.06426
-
One Transformer for All Time Series: Representing and Training with Time-Dependent Heterogeneous Tabular Data 13 Feb 2023 · 1 repository · arXiv:2302.06375Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Order Matters: Agent-by-agent Policy Optimization 13 Feb 2023 · 1 repository · arXiv:2302.06205Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Simple Hardware-Efficient Long Convolutions for Sequence Modeling 13 Feb 2023 · 1 repository · arXiv:2302.06646
-
STREET: A Multi-Task Structured Reasoning and Explanation Benchmark 13 Feb 2023 · 0 repositories · arXiv:2302.06729
-
Towards Local Visual Modeling for Image Captioning 13 Feb 2023 · 1 repository · arXiv:2302.06098
-
Using SHAP Values and Machine Learning to Understand Trends in the Transient Stability Limit 13 Feb 2023 · 0 repositories · arXiv:2302.06274
-
Academic Writing with GPT-3.5: Reflections on Practices, Efficacy and Transparency 12 Feb 2023 · 0 repositories · arXiv:2304.11079
-
Denoising and Prompt-Tuning for Multi-Behavior Recommendation 12 Feb 2023 · 1 repository · arXiv:2302.05862
-
Generalized Few-Shot Continual Learning with Contrastive Mixture of Adapters 12 Feb 2023 · 1 repository · arXiv:2302.05936
-
Self-supervised pseudo-colorizing of masked cells 12 Feb 2023 · 2 repositories · arXiv:2302.05968
-
Semantic Importance-Aware Communications Using Pre-trained Language Models 12 Feb 2023 · 0 repositories · arXiv:2302.07142
-
Transformer models: an introduction and catalog 12 Feb 2023 · 0 repositories · arXiv:2302.07730
-
A Brief Report on LawGPT 1.0: A Virtual Legal Assistant Based on GPT-3 11 Feb 2023 · 0 repositories · arXiv:2302.05729
-
Differentiable Outlier Detection Enable Robust Deep Multimodal Analysis 11 Feb 2023 · 1 repository · arXiv:2302.05608
-
DocILE Benchmark for Document Information Localization and Extraction 11 Feb 2023 · 1 repository · arXiv:2302.05658
-
Dual Relation Knowledge Distillation for Object Detection 11 Feb 2023 · 1 repository · arXiv:2302.05637
-
Informing clinical assessment by contextualizing post-hoc explanations of risk prediction models in type-2 diabetes 11 Feb 2023 · 0 repositories · arXiv:2302.05752
-
Rethinking Vision Transformer and Masked Autoencoder in Multimodal Face Anti-Spoofing 11 Feb 2023 · 0 repositories · arXiv:2302.05744
-
Alloprof: a new French question-answer education dataset and its use in an information retrieval case study 10 Feb 2023 · 1 repository · arXiv:2302.07738
-
BEST: BERT Pre-Training for Sign Language Recognition with Coupling Tokenization 10 Feb 2023 · 0 repositories · arXiv:2302.05075
-
Combat AI With AI: Counteract Machine-Generated Fake Restaurant Reviews on Social Media 10 Feb 2023 · 1 repository · arXiv:2302.07731
-
Dual Memory Units with Uncertainty Regulation for Weakly Supervised Video Anomaly Detection 10 Feb 2023 · 1 repository · arXiv:2302.05160Syntology official (archive's flag): 7 ran · 7 ran (of which 6 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Effective Document Image Enhancement Using tokens-to-token Transformer Network 10 Feb 2023 · 1 repository
-
FairPy: A Toolkit for Evaluation of Prediction Biases and their Mitigation in Large Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05508
-
GCNet: Probing Self-Similarity Learning for Generalized Counting Network 10 Feb 2023 · 0 repositories · arXiv:2302.05132
-
GTR-CTRL: Instrument and Genre Conditioning for Guitar-Focused Music Generation with Transformers 10 Feb 2023 · 0 repositories · arXiv:2302.05393
-
PATCorrect: Non-autoregressive Phoneme-augmented Transformer for ASR Error Correction 10 Feb 2023 · 0 repositories · arXiv:2302.05040
-
Predicting Out-of-Distribution Error with Confidence Optimal Transport 10 Feb 2023 · 0 repositories · arXiv:2302.05018
-
The Wisdom of Hindsight Makes Language Models Better Instruction Followers 10 Feb 2023 · 1 repository · arXiv:2302.05206
-
Translating Natural Language to Planning Goals with Large-Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05128
-
Better by you, better than me, chatgpt3 as writing assistance in students essays 9 Feb 2023 · 0 repositories · arXiv:2302.04536
-
Binarized Neural Machine Translation 9 Feb 2023 · 1 repository · arXiv:2302.04907Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Efficient Attention via Control Variates 9 Feb 2023 · 1 repository · arXiv:2302.04542
-
Flexible, Model-Agnostic Method for Materials Data Extraction from Text Using General Purpose Language Models 9 Feb 2023 · 0 repositories · arXiv:2302.04914
-
Generating a Structured Summary of Numerous Academic Papers: Dataset and Method 9 Feb 2023 · 1 repository · arXiv:2302.04580
-
Help the Blind See: Assistance for the Visually Impaired through Augmented Acoustic Simulation 9 Feb 2023 · 1 repository · arXiv:2303.13536
-
3D Human Pose and Shape Estimation via HybrIK-Transformer 9 Feb 2023 · 1 repository · arXiv:2302.04774
-
Optimized Hybrid Focal Margin Loss for Crack Segmentation 9 Feb 2023 · 0 repositories · arXiv:2302.04395
-
Reversible Vision Transformers 9 Feb 2023 · 4 repositories · arXiv:2302.04869Syntology official (archive's flag): 22 ran · 24 ran (of which 5 constructed an object rather than computing a result; 23 with no instrument failure: 0 honoured, 0 violated, 23 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 29 harvested samples) · 26 pointer-only (licence)
-
Adapting Pre-trained Vision Transformers from 2D to 3D through Weight Inflation Improves Medical Image Segmentation 8 Feb 2023 · 1 repository · arXiv:2302.04303
-
An Empirical Study of Uniform-Architecture Knowledge Distillation in Document Ranking 8 Feb 2023 · 0 repositories · arXiv:2302.04112
-
Attending to Graph Transformers 8 Feb 2023 · 1 repository · arXiv:2302.04181Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 4 pointer-only (licence)
-
EvoText: Enhancing Natural Language Generation Models via Self-Escalation Learning for Up-to-Date Knowledge and Improved Performance 8 Feb 2023 · 0 repositories · arXiv:2302.03896
-
Channelformer: Attention based Neural Solution for Wireless Channel Estimation and Effective Online Training 8 Feb 2023 · 1 repository · arXiv:2302.04368
-
CRL+: A Novel Semi-Supervised Deep Active Contrastive Representation Learning-Based Text Classification Model for Insurance Data 8 Feb 2023 · 0 repositories · arXiv:2302.04343
-
Dual-interest Factorization-heads Attention for Sequential Recommendation 8 Feb 2023 · 1 repository · arXiv:2302.03965
-
Efficient Joint Learning for Clinical Named Entity Recognition and Relation Extraction Using Fourier Networks: A Use Case in Adverse Drug Events 8 Feb 2023 · 1 repository · arXiv:2302.04185
-
Prompting for Multimodal Hateful Meme Classification 8 Feb 2023 · 0 repositories · arXiv:2302.04156
-
Short-Term Memory Convolutions 8 Feb 2023 · 0 repositories · arXiv:2302.04331
-
SwinCross: Cross-modal Swin Transformer for Head-and-Neck Tumor Segmentation in PET/CT Images 8 Feb 2023 · 0 repositories · arXiv:2302.03861
-
A Deep Learning-based in silico Framework for Optimization on Retinal Prosthetic Stimulation 7 Feb 2023 · 0 repositories · arXiv:2302.03570
-
Impact of velocity and impact angle on football shot accuracy during fundamental trainings 7 Feb 2023 · 0 repositories · arXiv:2302.03426
-
LUT-NN: Empower Efficient Neural Network Inference with Centroid Learning and Table Lookup 7 Feb 2023 · 0 repositories · arXiv:2302.03213
-
OSRT: Omnidirectional Image Super-Resolution with Distortion-aware Transformer 7 Feb 2023 · 1 repository · arXiv:2302.03453Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Reliable Natural Language Understanding with Large Language Models and Answer Set Programming 7 Feb 2023 · 0 repositories · arXiv:2302.03780
-
VertXNet: An Ensemble Method for Vertebrae Segmentation and Identification of Spinal X-Ray 7 Feb 2023 · 0 repositories · arXiv:2302.03476
-
What do Language Models know about word senses? Zero-Shot WSD with Language Models and Domain Inventories 7 Feb 2023 · 0 repositories · arXiv:2302.03353
-
What Matters In The Structured Pruning of Generative Language Models? 7 Feb 2023 · 1 repository · arXiv:2302.03773Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples)
-
AIM: Adapting Image Models for Efficient Video Action Recognition 6 Feb 2023 · 1 repository · arXiv:2302.03024
-
Context-Gloss Augmentation for Improving Arabic Target Sense Verification 6 Feb 2023 · 0 repositories · arXiv:2302.03126
-
Controllable Lexical Simplification for English 6 Feb 2023 · 1 repository · arXiv:2302.02900