Methods › General › Output Functions › Softmax › Papers, page 268
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 268 of 375: papers 26,701 to 26,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
NeW CRFs: Neural Window Fully-connected CRFs for Monocular Depth Estimation 3 Mar 2022 · 1 repository · arXiv:2203.01502Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
On Learning Contrastive Representations for Learning with Noisy Labels 3 Mar 2022 · 1 repository · arXiv:2203.01785
-
Polarity Sampling: Quality and Diversity Control of Pre-Trained Generative Networks via Singular Values 3 Mar 2022 · 1 repository · arXiv:2203.01993
-
Selective Residual M-Net for Real Image Denoising 3 Mar 2022 · 1 repository · arXiv:2203.01645
-
BoMD: Bag of Multi-label Descriptors for Noisy Chest X-ray Classification 3 Mar 2022 · 2 repositories · arXiv:2203.01937
-
Measuring Self-Supervised Representation Quality for Downstream Classification using Discriminative Features 3 Mar 2022 · 0 repositories · arXiv:2203.01881
-
ViTransPAD: Video Transformer using convolution and self-attention for Face Presentation Attack Detection 3 Mar 2022 · 0 repositories · arXiv:2203.01562
-
3DCTN: 3D Convolution-Transformer Network for Point Cloud Classification 2 Mar 2022 · 1 repository · arXiv:2203.00828
-
ADVISE: ADaptive Feature Relevance and VISual Explanations for Convolutional Neural Networks 2 Mar 2022 · 1 repository · arXiv:2203.01289
-
Aggregated Pyramid Vision Transformer: Split-transform-merge Strategy for Image Recognition without Convolutions 2 Mar 2022 · 0 repositories · arXiv:2203.00960
-
Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic Segmentation 2 Mar 2022 · 1 repository · arXiv:2203.01452
-
Class Re-Activation Maps for Weakly-Supervised Semantic Segmentation 2 Mar 2022 · 1 repository · arXiv:2203.00962Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Contextual Attention Network: Transformer Meets U-Net 2 Mar 2022 · 3 repositories · arXiv:2203.01932
-
D^2ETR: Decoder-Only DETR with Computationally Efficient Cross-Scale Attention 2 Mar 2022 · 0 repositories · arXiv:2203.00860
-
Discontinuous Constituency and BERT: A Case Study of Dutch 2 Mar 2022 · 1 repository · arXiv:2203.01063
-
DN-DETR: Accelerate DETR Training by Introducing Query DeNoising 2 Mar 2022 · 17 repositories · arXiv:2203.01305Syntology official (archive's flag): 10 ran · 17 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 11 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 9 pointer-only (licence)
-
FastFold: Reducing AlphaFold Training Time from 11 Days to 67 Hours 2 Mar 2022 · 1 repository · arXiv:2203.00854
-
Adaptive Discriminative Regularization for Visual Classification 2 Mar 2022 · 0 repositories · arXiv:2203.00833
-
Instance-aware multi-object self-supervision for monocular depth prediction 2 Mar 2022 · 0 repositories · arXiv:2203.00809
-
MTet: Multi-domain Translation for English-Vietnamese 2 Mar 2022 · 1 repository
-
Parameter-Efficient Mixture-of-Experts Architecture for Pre-trained Language Models 2 Mar 2022 · 2 repositories · arXiv:2203.01104Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Protecting Celebrities from DeepFake with Identity Consistency Transformer 2 Mar 2022 · 1 repository · arXiv:2203.01318Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Temporal Context Matters: Enhancing Single Image Prediction with Disease Progression Representations 2 Mar 2022 · 0 repositories · arXiv:2203.01933
-
X-Trans2Cap: Cross-Modal Knowledge Transfer using Transformer for 3D Dense Captioning 2 Mar 2022 · 1 repository · arXiv:2203.00843
-
BERT-LID: Leveraging BERT to Improve Spoken Language Identification 1 Mar 2022 · 1 repository · arXiv:2203.00328
-
Colon Nuclei Instance Segmentation using a Probabilistic Two-Stage Detector 1 Mar 2022 · 0 repositories · arXiv:2203.01321
-
DeepNet: Scaling Transformers to 1,000 Layers 1 Mar 2022 · 6 repositories · arXiv:2203.00555Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 3 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
E-LANG: Energy-Based Joint Inferencing of Super and Swift Language Models 1 Mar 2022 · 0 repositories · arXiv:2203.00748
-
Exploring and Adapting Chinese GPT to Pinyin Input Method 1 Mar 2022 · 1 repository · arXiv:2203.00249
-
Fast-R2D2: A Pretrained Recursive Neural Network based on Pruned CKY for Grammar Induction and Text Representation 1 Mar 2022 · 2 repositories · arXiv:2203.00281Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
HyperPrompt: Prompt-based Task-Conditioning of Transformers 1 Mar 2022 · 0 repositories · arXiv:2203.00759
-
"Is Whole Word Masking Always Better for Chinese BERT?": Probing on Chinese Grammatical Error Correction 1 Mar 2022 · 0 repositories · arXiv:2203.00286
-
Robots Autonomously Detecting People: A Multimodal Deep Contrastive Learning Method Robust to Intraclass Variations 1 Mar 2022 · 0 repositories · arXiv:2203.00187
-
Temporal Perceiver: A General Architecture for Arbitrary Boundary Detection 1 Mar 2022 · 0 repositories · arXiv:2203.00307
-
Transformer Grammars: Augmenting Transformer Language Models with Syntactic Inductive Biases at Scale 1 Mar 2022 · 0 repositories · arXiv:2203.00633
-
Wearable Sensor-Based Human Activity Recognition with Transformer Model 1 Mar 2022 · 1 repository
-
A Data-scalable Transformer for Medical Image Segmentation: Architecture, Model Efficiency, and Benchmark 28 Feb 2022 · 2 repositories · arXiv:2203.00131Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
CTformer: Convolution-free Token2Token Dilated Vision Transformer for Low-dose CT Denoising 28 Feb 2022 · 2 repositories · arXiv:2202.13517
-
Deep, Deep Learning with BART 28 Feb 2022 · 1 repository · arXiv:2202.14005
-
DropIT: Dropping Intermediate Tensors for Memory-Efficient DNN Training 28 Feb 2022 · 1 repository · arXiv:2202.13808
-
Filter-enhanced MLP is All You Need for Sequential Recommendation 28 Feb 2022 · 2 repositories · arXiv:2202.13556
-
LiLT: A Simple yet Effective Language-Independent Layout Transformer for Structured Document Understanding 28 Feb 2022 · 5 repositories · arXiv:2202.13669Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
LISA: Learning Interpretable Skill Abstractions from Language 28 Feb 2022 · 1 repository · arXiv:2203.00054
-
Spatio-temporal Vision Transformer for Super-resolution Microscopy 28 Feb 2022 · 1 repository · arXiv:2203.00030
-
SUNet: Swin Transformer UNet for Image Denoising 28 Feb 2022 · 2 repositories · arXiv:2202.14009
-
The impact of lexical and grammatical processing on generating code from natural language 28 Feb 2022 · 2 repositories · arXiv:2202.13972
-
Using Multi-scale SwinTransformer-HTC with Data augmentation in CoNIC Challenge 28 Feb 2022 · 0 repositories · arXiv:2202.13588
-
A Computer Vision-assisted Approach to Automated Real-Time Road Infrastructure Management 27 Feb 2022 · 0 repositories · arXiv:2202.13285
-
DXM-TransFuse U-net: Dual Cross-Modal Transformer Fusion U-net for Automated Nerve Identification 27 Feb 2022 · 0 repositories · arXiv:2202.13304
-
Enhancing Legal Argument Mining with Domain Pre-training and Neural Networks 27 Feb 2022 · 1 repository · arXiv:2202.13457
-
Spatial-Temporal Attention Fusion Network for short-term passenger flow prediction on holidays in urban rail transit systems 27 Feb 2022 · 0 repositories · arXiv:2203.00007
-
Short-term passenger flow prediction for multi-traffic modes: A Transformer and residual network based multi-task learning method 27 Feb 2022 · 0 repositories · arXiv:2203.00422
-
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation 27 Feb 2022 · 2 repositories · arXiv:2202.13393
-
A Systematic Evaluation of Large Language Models of Code 26 Feb 2022 · 3 repositories · arXiv:2202.13169Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
An End-to-End Transformer Model for Crowd Localization 26 Feb 2022 · 1 repository · arXiv:2202.13065
-
Analysis of Visual Reasoning on One-Stage Object Detection 26 Feb 2022 · 0 repositories · arXiv:2202.13115
-
Bi-directional Joint Neural Networks for Intent Classification and Slot Filling 26 Feb 2022 · 0 repositories · arXiv:2202.13079
-
Real-World Blind Super-Resolution via Feature Matching with Implicit High-Resolution Priors 26 Feb 2022 · 2 repositories · arXiv:2202.13142
-
Multi-image Super-resolution via Quality Map Associated Attention Network 26 Feb 2022 · 0 repositories · arXiv:2202.13124
-
Multi-Level Contrastive Learning for Cross-Lingual Alignment 26 Feb 2022 · 0 repositories · arXiv:2202.13083
-
A Data-Driven Column Generation Algorithm For Bin Packing Problem in Manufacturing Industry 25 Feb 2022 · 0 repositories · arXiv:2202.12466
-
A Hardware-Aware System for Accelerating Deep Neural Network Optimization 25 Feb 2022 · 0 repositories · arXiv:2202.12954
-
APEACH: Attacking Pejorative Expressions with Analysis on Crowd-Generated Hate Speech Evaluation Datasets 25 Feb 2022 · 1 repository · arXiv:2202.12459
-
Extracting Effective Subnetworks with Gumbel-Softmax 25 Feb 2022 · 1 repository · arXiv:2202.12986
-
Rethinking the Role of Demonstrations: What Makes In-Context Learning Work? 25 Feb 2022 · 2 repositories · arXiv:2202.12837
-
Self-Supervised and Interpretable Anomaly Detection using Network Transformers 25 Feb 2022 · 0 repositories · arXiv:2202.12997
-
BERTVision -- A Parameter-Efficient Approach for Question Answering 24 Feb 2022 · 1 repository · arXiv:2202.12210Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Finding Inverse Document Frequency Information in BERT 24 Feb 2022 · 0 repositories · arXiv:2202.12191
-
Finite-Sum Coupled Compositional Stochastic Optimization: Theory and Applications 24 Feb 2022 · 0 repositories · arXiv:2202.12396
-
From Natural Language to Simulations: Applying GPT-3 Codex to Automate Simulation Modeling of Logistics Systems 24 Feb 2022 · 1 repository · arXiv:2202.12107
-
Instantaneous Physiological Estimation using Video Transformers 24 Feb 2022 · 1 repository · arXiv:2202.12368
-
openFEAT: Improving Speaker Identification by Open-set Few-shot Embedding Adaptation with Transformer 24 Feb 2022 · 0 repositories · arXiv:2202.12349
-
Pretraining without Wordpieces: Learning Over a Vocabulary of Millions of Words 24 Feb 2022 · 0 repositories · arXiv:2202.12142
-
Probing BERT's priors with serial reproduction chains 24 Feb 2022 · 1 repository · arXiv:2202.12226Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Retriever: Learning Content-Style Representation as a Token-Level Bipartite Graph 24 Feb 2022 · 2 repositories · arXiv:2202.12307Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Sky Computing: Accelerating Geo-distributed Computing in Federated Learning 24 Feb 2022 · 1 repository · arXiv:2202.11836
-
Transformers in Medical Image Analysis: A Review 24 Feb 2022 · 0 repositories · arXiv:2202.12165
-
TrimBERT: Tailoring BERT for Trade-offs 24 Feb 2022 · 0 repositories · arXiv:2202.12411
-
Using calibrator to improve robustness in Machine Reading Comprehension 24 Feb 2022 · 0 repositories · arXiv:2202.11865
-
A Bayesian Deep Learning Approach to Near-Term Climate Prediction 23 Feb 2022 · 0 repositories · arXiv:2202.11244
-
A Differential Attention Fusion Model Based on Transformer for Time Series Forecasting 23 Feb 2022 · 0 repositories · arXiv:2202.11402
-
Consistent Dropout for Policy Gradient Reinforcement Learning 23 Feb 2022 · 0 repositories · arXiv:2202.11818
-
FastRPB: a Scalable Relative Positional Encoding for Long Sequence Tasks 23 Feb 2022 · 1 repository · arXiv:2202.11364
-
Integration of neural network and fuzzy logic decision making compared with bilayered neural network in the simulation of daily dew point temperature 23 Feb 2022 · 0 repositories · arXiv:2202.12256
-
ISDA: Position-Aware Instance Segmentation with Deformable Attention 23 Feb 2022 · 1 repository · arXiv:2202.12251
-
Paying U-Attention to Textures: Multi-Stage Hourglass Vision Transformer for Universal Texture Synthesis 23 Feb 2022 · 0 repositories · arXiv:2202.11703
-
Refining the state-of-the-art in Machine Translation, optimizing NMT for the JA <-> EN language pair by leveraging personal domain expertise 23 Feb 2022 · 0 repositories · arXiv:2202.11669
-
Skeleton Sequence and RGB Frame Based Multi-Modality Feature Fusion Network for Action Recognition 23 Feb 2022 · 0 repositories · arXiv:2202.11374
-
Think Global, Act Local: Dual-scale Graph Transformer for Vision-and-Language Navigation 23 Feb 2022 · 1 repository · arXiv:2202.11742Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
When do GANs replicate? On the choice of dataset size 23 Feb 2022 · 1 repository · arXiv:2202.11765
-
A New Generation of Perspective API: Efficient Multilingual Character-level Transformers 22 Feb 2022 · 0 repositories · arXiv:2202.11176
-
Exploiting long-term temporal dynamics for video captioning 22 Feb 2022 · 0 repositories · arXiv:2202.10828
-
GroupViT: Semantic Segmentation Emerges from Text Supervision 22 Feb 2022 · 6 repositories · arXiv:2202.11094
-
Improving CTC-based speech recognition via knowledge transferring from pre-trained language models 22 Feb 2022 · 1 repository · arXiv:2203.03582
-
JAMES: Normalizing Job Titles with Multi-Aspect Graph Embeddings and Reasoning 22 Feb 2022 · 0 repositories · arXiv:2202.10739
-
Learning Cluster Patterns for Abstractive Summarization 22 Feb 2022 · 0 repositories · arXiv:2202.10967
-
One-shot Scene Graph Generation 22 Feb 2022 · 1 repository · arXiv:2202.10824
-
Social Computational Design Method for Generating Product Shapes with GAN and Transformer Models 22 Feb 2022 · 0 repositories · arXiv:2202.10774
-
Socialformer: Social Network Inspired Long Document Modeling for Document Ranking 22 Feb 2022 · 1 repository · arXiv:2202.10870
-
Tracking perovskite crystallization via deep learning-based feature detection on 2D X-ray scattering data 22 Feb 2022 · 1 repository · arXiv:2202.10983