Methods › General › Output Functions › Softmax › Papers, page 269
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 269 of 375: papers 26,801 to 26,900 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Embarrassingly Simple Performance Prediction for Abductive Natural Language Inference 21 Feb 2022 · 1 repository · arXiv:2202.10408
-
Formal Analysis of the Sampling Behaviour of Stochastic Event-Triggered Control 21 Feb 2022 · 0 repositories · arXiv:2202.10178
-
Items from Psychometric Tests as Training Data for Personality Profiling Models of Twitter Users 21 Feb 2022 · 0 repositories · arXiv:2202.10415
-
Rethinking the Zigzag Flattening for Image Reading 21 Feb 2022 · 0 repositories · arXiv:2202.10240
-
S3T: Self-Supervised Pre-training with Swin Transformer for Music Classification 21 Feb 2022 · 1 repository · arXiv:2202.10139
-
ViTAEv2: Vision Transformer Advanced by Exploring Inductive Bias for Image Recognition and Beyond 21 Feb 2022 · 8 repositories · arXiv:2202.10108Syntology community repositories only · 22 ran (of which 8 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 25 harvested samples) · 4 pointer-only (licence)
-
ARM3D: Attention-based relation module for indoor 3D object detection 20 Feb 2022 · 1 repository · arXiv:2202.09715
-
Contextual Semantic Embeddings for Ontology Subsumption Prediction 20 Feb 2022 · 2 repositories · arXiv:2202.09791
-
Generalized Bayesian Additive Regression Trees Models: Beyond Conditional Conjugacy 20 Feb 2022 · 0 repositories · arXiv:2202.09924
-
Do Transformers know symbolic rules, and would we know if they did? 19 Feb 2022 · 0 repositories · arXiv:2203.00162
-
GCNET: graph-based prediction of stock price movement using graph convolutional network 19 Feb 2022 · 1 repository · arXiv:2203.11091
-
TransDreamer: Reinforcement Learning with Transformer World Models 19 Feb 2022 · 0 repositories · arXiv:2202.09481Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Survey of Vision-Language Pre-Trained Models 18 Feb 2022 · 0 repositories · arXiv:2202.10936
-
Evaluating the Construct Validity of Text Embeddings with Application to Survey Questions 18 Feb 2022 · 1 repository · arXiv:2202.09166
-
Mixture-of-Experts with Expert Choice Routing 18 Feb 2022 · 0 repositories · arXiv:2202.09368
-
Task Specific Attention is one more thing you need for object detection 18 Feb 2022 · 1 repository · arXiv:2202.09048
-
Unleashing the Power of Transformer for Graphs 18 Feb 2022 · 0 repositories · arXiv:2202.10581
-
Multi-Scale Hybrid Vision Transformer for Learning Gastric Histology: AI-Based Decision Support System for Gastric Cancer Treatment 17 Feb 2022 · 0 repositories · arXiv:2202.08510
-
cosFormer: Rethinking Softmax in Attention 17 Feb 2022 · 3 repositories · arXiv:2202.08791
-
ST-MoE: Designing Stable and Transferable Sparse Expert Models 17 Feb 2022 · 3 repositories · arXiv:2202.08906Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
General Cyclical Training of Neural Networks 17 Feb 2022 · 1 repository · arXiv:2202.08835
-
Graph Masked Autoencoders with Transformers 17 Feb 2022 · 1 repository · arXiv:2202.08391
-
Improving English to Sinhala Neural Machine Translation using Part-of-Speech Tag 17 Feb 2022 · 0 repositories · arXiv:2202.08882
-
Mirror-Yolo: A Novel Attention Focus, Instance Segmentation and Mirror Detection Model 17 Feb 2022 · 0 repositories · arXiv:2202.08498
-
Revisiting Over-smoothing in BERT from the Perspective of Graph 17 Feb 2022 · 0 repositories · arXiv:2202.08625
-
SGPT: GPT Sentence Embeddings for Semantic Search 17 Feb 2022 · 1 repository · arXiv:2202.08904Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Transformer for Graphs: An Overview from Architecture Perspective 17 Feb 2022 · 1 repository · arXiv:2202.08455Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TraSeTR: Track-to-Segment Transformer with Contrastive Query for Instance-level Instrument Segmentation in Robotic Surgery 17 Feb 2022 · 0 repositories · arXiv:2202.08453
-
When BERT Meets Quantum Temporal Convolution Learning for Text Classification in Heterogeneous Computing 17 Feb 2022 · 0 repositories · arXiv:2203.03550
-
A Survey of Pretraining on Graphs: Taxonomy, Methods, and Applications 16 Feb 2022 · 3 repositories · arXiv:2202.07893
-
ActionFormer: Localizing Moments of Actions with Transformers 16 Feb 2022 · 1 repository · arXiv:2202.07925
-
Cyclical Focal Loss 16 Feb 2022 · 1 repository · arXiv:2202.08978
-
EdgeFormer: A Parameter-Efficient Transformer for On-Device Seq2seq Generation 16 Feb 2022 · 1 repository · arXiv:2202.07959Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Meta Knowledge Distillation 16 Feb 2022 · 0 repositories · arXiv:2202.07940
-
No One Left Behind: Inclusive Federated Learning over Heterogeneous Devices 16 Feb 2022 · 0 repositories · arXiv:2202.08036
-
Probing Pretrained Models of Source Code 16 Feb 2022 · 1 repository · arXiv:2202.08975
-
Reducing Overconfidence Predictions for Autonomous Driving Perception 16 Feb 2022 · 0 repositories · arXiv:2202.07825
-
The learning phases in NN: From Fitting the Majority to Fitting a Few 16 Feb 2022 · 0 repositories · arXiv:2202.08299
-
The NLP Task Effectiveness of Long-Range Transformers 16 Feb 2022 · 0 repositories · arXiv:2202.07856
-
A Survey on Dynamic Neural Networks for Natural Language Processing 15 Feb 2022 · 0 repositories · arXiv:2202.07101
-
A Survey on Model Compression and Acceleration for Pretrained Language Models 15 Feb 2022 · 0 repositories · arXiv:2202.07105
-
BLUE at Memotion 2.0 2022: You have my Image, my Text and my Transformer 15 Feb 2022 · 0 repositories · arXiv:2202.07543
-
Defending against Reconstruction Attacks with Rényi Differential Privacy 15 Feb 2022 · 0 repositories · arXiv:2202.07623
-
DualConv: Dual Convolutional Kernels for Lightweight Deep Neural Networks 15 Feb 2022 · 1 repository · arXiv:2202.07481
-
HiMA: A Fast and Scalable History-based Memory Access Engine for Differentiable Neural Computer 15 Feb 2022 · 0 repositories · arXiv:2202.07275
-
Improving the repeatability of deep learning models with Monte Carlo dropout 15 Feb 2022 · 1 repository · arXiv:2202.07562
-
One Configuration to Rule Them All? Towards Hyperparameter Transfer in Topic Models using Multi-Objective Bayesian Optimization 15 Feb 2022 · 1 repository · arXiv:2202.07631
-
Personalized Prompt Learning for Explainable Recommendation 15 Feb 2022 · 1 repository · arXiv:2202.07371
-
Predicting on the Edge: Identifying Where a Larger Model Does Better 15 Feb 2022 · 0 repositories · arXiv:2202.07652
-
SODAR: Segmenting Objects by DynamicallyAggregating Neighboring Mask Representations 15 Feb 2022 · 1 repository · arXiv:2202.07402
-
Taking a Step Back with KCal: Multi-Class Kernel-Based Calibration for Deep Neural Networks 15 Feb 2022 · 0 repositories · arXiv:2202.07679
-
Tomayto, Tomahto. Beyond Token-level Answer Equivalence for Question Answering Evaluation 15 Feb 2022 · 1 repository · arXiv:2202.07654
-
Toxic Comments Hunter : Score Severity of Toxic Comments 15 Feb 2022 · 0 repositories · arXiv:2203.03548
-
Transformers in Time Series: A Survey 15 Feb 2022 · 11 repositories · arXiv:2202.07125
-
Unreasonable Effectiveness of Last Hidden Layer Activations for Adversarial Robustness 15 Feb 2022 · 0 repositories · arXiv:2202.07342
-
ViNTER: Image Narrative Generation with Emotion-Arc-Aware Transformer 15 Feb 2022 · 0 repositories · arXiv:2202.07305
-
XAI for Transformers: Better Explanations through Conservative Propagation 15 Feb 2022 · 1 repository · arXiv:2202.07304Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
CodeFill: Multi-token Code Completion by Jointly Learning from Structure and Naming Sequences 14 Feb 2022 · 1 repository · arXiv:2202.06689
-
Faster hyperspectral image classification based on selective kernel mechanism using deep convolutional networks 14 Feb 2022 · 1 repository · arXiv:2202.06458
-
Geometric Transformer for Fast and Robust Point Cloud Registration 14 Feb 2022 · 2 repositories · arXiv:2202.06688Syntology community repositories only · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 9 harvested samples) · 9 pointer-only (licence)
-
Handcrafted Histological Transformer (H2T): Unsupervised Representation of Whole Slide Images 14 Feb 2022 · 1 repository · arXiv:2202.07001Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
I-Tuning: Tuning Frozen Language Models with Image for Lightweight Image Captioning 14 Feb 2022 · 0 repositories · arXiv:2202.06574
-
Mixing and Shifting: Exploiting Global and Local Dependencies in Vision MLPs 14 Feb 2022 · 2 repositories · arXiv:2202.06510Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 6 pointer-only (licence)
-
Punctuation restoration in Swedish through fine-tuned KB-BERT 14 Feb 2022 · 0 repositories · arXiv:2202.06769
-
QA4QG: Using Question Answering to Constrain Multi-Hop Question Generation 14 Feb 2022 · 1 repository · arXiv:2202.06538
-
Research on Dual Channel News Headline Classification Based on ERNIE Pre-training Model 14 Feb 2022 · 0 repositories · arXiv:2202.06600
-
Sequence-to-Sequence Resources for Catalan 14 Feb 2022 · 1 repository · arXiv:2202.06871
-
Source Code Summarization with Structural Relative Position Guided Transformer 14 Feb 2022 · 1 repository · arXiv:2202.06521
-
Transformer Memory as a Differentiable Search Index 14 Feb 2022 · 1 repository · arXiv:2202.06991
-
UserBERT: Modeling Long- and Short-Term User Preferences via Self-Supervision 14 Feb 2022 · 0 repositories · arXiv:2202.07605
-
What Do They Capture? -- A Structural Analysis of Pre-Trained Language Models for Source Code 14 Feb 2022 · 1 repository · arXiv:2202.06840
-
Assessment of contextualised representations in detecting outcome phrases in clinical trials 13 Feb 2022 · 0 repositories · arXiv:2203.03547
-
BViT: Broad Attention based Vision Transformer 13 Feb 2022 · 1 repository · arXiv:2202.06268
-
ET-BERT: A Contextualized Datagram Representation with Pre-training Transformers for Encrypted Traffic Classification 13 Feb 2022 · 1 repository · arXiv:2202.06335
-
LighTN: Light-weight Transformer Network for Performance-overhead Tradeoff in Point Cloud Downsampling 13 Feb 2022 · 0 repositories · arXiv:2202.06263
-
LMN at SemEval-2022 Task 11: A Transformer-based System for English Named Entity Recognition 13 Feb 2022 · 0 repositories · arXiv:2203.03546
-
A multi-task semi-supervised framework for Text2Graph & Graph2Text 12 Feb 2022 · 1 repository · arXiv:2202.06041
-
Automatic Issue Classifier: A Transfer Learning Framework for Classifying Issue Reports 12 Feb 2022 · 1 repository · arXiv:2202.06149
-
Benchmark Assessment for DeepSpeed Optimization Library 12 Feb 2022 · 0 repositories · arXiv:2202.12831
-
Maximizing Communication Efficiency for Large-scale Training via 0/1 Adam 12 Feb 2022 · 1 repository · arXiv:2202.06009
-
Multi-direction and Multi-scale Pyramid in Transformer for Video-based Pedestrian Retrieval 12 Feb 2022 · 1 repository · arXiv:2202.06014
-
Multi-direction and Multi-scale Pyramid in Transformer for Video-based Pedestrian Retrieval 12 Feb 2022 · 1 repository
-
Constrained Optimization with Dynamic Bound-scaling for Effective NLPBackdoor Defense 11 Feb 2022 · 1 repository · arXiv:2202.05749
-
HaT5: Hate Language Identification using Text-to-Text Transfer Transformer 11 Feb 2022 · 0 repositories · arXiv:2202.05690
-
Including Facial Expressions in Contextual Embeddings for Sign Language Generation 11 Feb 2022 · 0 repositories · arXiv:2202.05383
-
Vehicle and License Plate Recognition with Novel Dataset for Toll Collection 11 Feb 2022 · 3 repositories · arXiv:2202.05631
-
White-Box Attacks on Hate-speech BERT Classifiers in German with Explicit and Implicit Character Level Defense 11 Feb 2022 · 1 repository · arXiv:2202.05778
-
A Multi-task Learning Framework for Product Ranking with BERT 10 Feb 2022 · 0 repositories · arXiv:2202.05317
-
AA-TransUNet: Attention Augmented TransUNet For Nowcasting Tasks 10 Feb 2022 · 1 repository · arXiv:2202.04996
-
Quantune: Post-training Quantization of Convolutional Neural Networks using Extreme Gradient Boosting for Fast Deployment 10 Feb 2022 · 1 repository · arXiv:2202.05048
-
Slovene SuperGLUE Benchmark: Translation and Evaluation 10 Feb 2022 · 0 repositories · arXiv:2202.04994
-
A Local Geometric Interpretation of Feature Extraction in Deep Feedforward Neural Networks 9 Feb 2022 · 0 repositories · arXiv:2202.04632
-
Can Open Domain Question Answering Systems Answer Visual Knowledge Questions? 9 Feb 2022 · 0 repositories · arXiv:2202.04306
-
Deep Feature Rotation for Multimodal Image Style Transfer 9 Feb 2022 · 1 repository · arXiv:2202.04426
-
Social Media as an Instant Source of Feedback on Water Quality 9 Feb 2022 · 0 repositories · arXiv:2202.04462
-
pNLP-Mixer: an Efficient all-MLP Architecture for Language 9 Feb 2022 · 1 repository · arXiv:2202.04350
-
The Volcspeech system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge 9 Feb 2022 · 0 repositories · arXiv:2202.04261
-
CALM: Contrastive Aligned Audio-Language Multirate and Multimodal Representations 8 Feb 2022 · 0 repositories · arXiv:2202.03587
-
Do Language Models Learn Position-Role Mappings? 8 Feb 2022 · 0 repositories · arXiv:2202.03611
-
Efficacy of Transformer Networks for Classification of Raw EEG Data 8 Feb 2022 · 0 repositories · arXiv:2202.05170