Methods › General › Output Functions › Softmax › Papers, page 314
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 314 of 375: papers 31,301 to 31,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fully Non-autoregressive Neural Machine Translation: Tricks of the Trade 31 Dec 2020 · 1 repository · arXiv:2012.15833
-
KART: Parameterization of Privacy Leakage Scenarios from Pre-trained Language Models 31 Dec 2020 · 1 repository · arXiv:2101.00036
-
Fast WordPiece Tokenization 31 Dec 2020 · 1 repository · arXiv:2012.15524
-
Making Pre-trained Language Models Better Few-shot Learners 31 Dec 2020 · 9 repositories · arXiv:2012.15723Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 9 harvested samples) · 7 pointer-only (licence)
-
MiniLMv2: Multi-Head Self-Attention Relation Distillation for Compressing Pretrained Transformers 31 Dec 2020 · 2 repositories · arXiv:2012.15828
-
Rethinking Semantic Segmentation from a Sequence-to-Sequence Perspective with Transformers 31 Dec 2020 · 5 repositories · arXiv:2012.15840
-
Revisiting Robust Neural Machine Translation: A Transformer Case Study 31 Dec 2020 · 0 repositories · arXiv:2012.15710
-
Studying Strategically: Learning to Mask for Closed-book QA 31 Dec 2020 · 0 repositories · arXiv:2012.15856
-
The Pile: An 800GB Dataset of Diverse Text for Language Modeling 31 Dec 2020 · 22 repositories · arXiv:2101.00027Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
TransTrack: Multiple Object Tracking with Transformer 31 Dec 2020 · 2 repositories · arXiv:2012.15460
-
Unified Mandarin TTS Front-end Based on Distilled BERT Model 31 Dec 2020 · 1 repository · arXiv:2012.15404
-
UNKs Everywhere: Adapting Multilingual Language Models to New Scripts 31 Dec 2020 · 2 repositories · arXiv:2012.15562
-
Verb Knowledge Injection for Multilingual Event Processing 31 Dec 2020 · 0 repositories · arXiv:2012.15421
-
XLM-T: Scaling up Multilingual Machine Translation with Pretrained Cross-lingual Transformer Encoders 31 Dec 2020 · 0 repositories · arXiv:2012.15547
-
ECONET: Effective Continual Pretraining of Language Models for Event Temporal Reasoning 30 Dec 2020 · 2 repositories · arXiv:2012.15283
-
Deriving Contextualised Semantic Features from BERT (and Other Transformer Model) Embeddings 30 Dec 2020 · 0 repositories · arXiv:2012.15353
-
Improving BERT with Syntax-aware Local Attention 30 Dec 2020 · 1 repository · arXiv:2012.15150
-
Optimizing Deeper Transformers on Small Datasets 30 Dec 2020 · 1 repository · arXiv:2012.15355
-
Out of Order: How Important Is The Sequential Order of Words in a Sentence in Natural Language Understanding Tasks? 30 Dec 2020 · 0 repositories · arXiv:2012.15180
-
SemGloVe: Semantic Co-occurrences for GloVe from BERT 30 Dec 2020 · 3 repositories · arXiv:2012.15197
-
Towards Unsupervised Deep Image Enhancement with Generative Adversarial Network 30 Dec 2020 · 1 repository · arXiv:2012.15020
-
Transformer for Image Quality Assessment 30 Dec 2020 · 0 repositories · arXiv:2101.01097
-
UnNatural Language Inference 30 Dec 2020 · 1 repository · arXiv:2101.00010
-
A Hierarchical Transformer with Speaker Modeling for Emotion Recognition in Conversation 29 Dec 2020 · 1 repository · arXiv:2012.14781
-
Kaleidoscope: An Efficient, Learnable Representation For All Structured Linear Maps 29 Dec 2020 · 2 repositories · arXiv:2012.14966Syntology official (archive's flag): 16 ran · 16 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 7 where Syntology's instrument failed) · 7 unverified (of 23 harvested samples)
-
LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding 29 Dec 2020 · 9 repositories · arXiv:2012.14740
-
Multiple Structural Priors Guided Self Attention Network for Language Understanding 29 Dec 2020 · 0 repositories · arXiv:2012.14642
-
Object sorting using faster R-CNN 29 Dec 2020 · 0 repositories · arXiv:2012.14840
-
Presenting a Dataset for Collaborator Recommending Systems in Academic Social Network: a Case Study on ReseachGate 29 Dec 2020 · 0 repositories · arXiv:2101.01141
-
Robust Dialogue Utterance Rewriting as Sequence Tagging 29 Dec 2020 · 1 repository · arXiv:2012.14535
-
Code Summarization with Structure-induced Transformer 29 Dec 2020 · 1 repository · arXiv:2012.14710
-
Understanding and Improving Encoder Layer Fusion in Sequence-to-Sequence Learning 29 Dec 2020 · 1 repository · arXiv:2012.14768
-
A Paragraph-level Multi-task Learning Model for Scientific Fact-Verification 28 Dec 2020 · 1 repository · arXiv:2012.14500
-
BURT: BERT-inspired Universal Representation from Learning Meaningful Segment 28 Dec 2020 · 0 repositories · arXiv:2012.14320
-
Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models 28 Dec 2020 · 2 repositories · arXiv:2012.14252
-
Red Dragon AI at TextGraphs 2020 Shared Task: LIT : LSTM-Interleaved Transformer for Multi-Hop Explanation Ranking 28 Dec 2020 · 1 repository · arXiv:2012.14164
-
Syntax-Enhanced Pre-trained Model 28 Dec 2020 · 1 repository · arXiv:2012.14116
-
TransPose: Keypoint Localization via Transformer 28 Dec 2020 · 1 repository · arXiv:2012.14214
-
A multi-task learning network using shared BERT models for aspect-based sentiment analysis 27 Dec 2020 · 0 repositories
-
ALP-KD: Attention-Based Layer Projection for Knowledge Distillation 27 Dec 2020 · 0 repositories · arXiv:2012.14022
-
An Embarrassingly Simple Model for Dialogue Relation Extraction 27 Dec 2020 · 1 repository · arXiv:2012.13873
-
Ellipse Regression with Predicted Uncertainties for Accurate Multi-View 3D Object Estimation 27 Dec 2020 · 0 repositories · arXiv:2101.05212
-
Inserting Information Bottlenecks for Attribution in Transformers 27 Dec 2020 · 1 repository · arXiv:2012.13838
-
Learning Light-Weight Translation Models from Deep Transformer 27 Dec 2020 · 1 repository · arXiv:2012.13866Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
MeDAL: Medical Abbreviation Disambiguation Dataset for Natural Language Understanding Pretraining 27 Dec 2020 · 1 repository · arXiv:2012.13978
-
Portfolio Optimization with 2D Relative-Attentional Gated Transformer 27 Dec 2020 · 0 repositories · arXiv:2101.03138
-
SG-Net: Syntax Guided Transformer for Language Representation 27 Dec 2020 · 0 repositories · arXiv:2012.13915
-
Sparse Adversarial Attack to Object Detection 26 Dec 2020 · 1 repository · arXiv:2012.13692
-
Inception Convolution with Efficient Dilation Search 25 Dec 2020 · 1 repository · arXiv:2012.13587
-
Detecting Hateful Memes Using a Multimodal Deep Ensemble 24 Dec 2020 · 1 repository · arXiv:2012.13235
-
Global Context Networks 24 Dec 2020 · 3 repositories · arXiv:2012.13375
-
I like fish, especially dolphins: Addressing Contradictions in Dialogue Modeling 24 Dec 2020 · 0 repositories · arXiv:2012.13391
-
Seed Phenotyping on Neural Networks using Domain Randomization and Transfer Learning 24 Dec 2020 · 0 repositories · arXiv:2012.13259
-
On the Granularity of Explanations in Model Agnostic NLP Interpretability 24 Dec 2020 · 1 repository · arXiv:2012.13189
-
To what extent do human explanations of model behavior align with actual model behavior? 24 Dec 2020 · 0 repositories · arXiv:2012.13354
-
A Survey on Visual Transformer 23 Dec 2020 · 0 repositories · arXiv:2012.12556
-
Bridging Textual and Tabular Data for Cross-Domain Text-to-SQL Semantic Parsing 23 Dec 2020 · 2 repositories · arXiv:2012.12627
-
Detecting Hate Speech in Memes Using Multimodal Deep Learning Approaches: Prize-winning solution to Hateful Memes Challenge 23 Dec 2020 · 1 repository · arXiv:2012.12975
-
Future-Guided Incremental Transformer for Simultaneous Translation 23 Dec 2020 · 0 repositories · arXiv:2012.12465
-
Self-Supervised Hyperboloid Representations from Logical Queries over Knowledge Graphs 23 Dec 2020 · 1 repository · arXiv:2012.13023
-
SWA Object Detection 23 Dec 2020 · 2 repositories · arXiv:2012.12645
-
Training data-efficient image transformers & distillation through attention 23 Dec 2020 · 40 repositories · arXiv:2012.12877Syntology community repositories only · 12 ran (of which 1 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 19 harvested samples) · 3 pointer-only (licence)
-
Vehicle Re-identification Based on Dual Distance Center Loss 23 Dec 2020 · 0 repositories · arXiv:2012.12519
-
Applying Wav2vec2.0 to Speech Recognition in Various Low-resource Languages 22 Dec 2020 · 0 repositories · arXiv:2012.12121
-
Domain Adaptation of NMT models for English-Hindi Machine Translation Task at AdapMT ICON 2020 22 Dec 2020 · 0 repositories · arXiv:2012.12112
-
FcaNet: Frequency Channel Attention Networks 22 Dec 2020 · 7 repositories · arXiv:2012.11879
-
Improved Biomedical Word Embeddings in the Transformer Era 22 Dec 2020 · 1 repository · arXiv:2012.11808
-
Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning 22 Dec 2020 · 2 repositories · arXiv:2012.13255Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Learning to Retrieve Entity-Aware Knowledge and Generate Responses with Copy Mechanism for Task-Oriented Dialogue Systems 22 Dec 2020 · 1 repository · arXiv:2012.11937
-
Molecular CT: Unifying Geometry and Representation Learning for Molecules at Different Scales 22 Dec 2020 · 0 repositories · arXiv:2012.11816
-
Multi-Head Self-Attention with Role-Guided Masks 22 Dec 2020 · 1 repository · arXiv:2012.12366
-
Recognizing Emotion Cause in Conversations 22 Dec 2020 · 1 repository · arXiv:2012.11820
-
Uncertainty and Surprisal Jointly Deliver the Punchline: Exploiting Incongruity-Based Features for Humor Recognition 22 Dec 2020 · 0 repositories · arXiv:2012.12007
-
3D Object Detection with Pointformer 21 Dec 2020 · 1 repository · arXiv:2012.11409
-
A Graph Reasoning Network for Multi-turn Response Selection via Customized Pre-training 21 Dec 2020 · 0 repositories · arXiv:2012.11099
-
Cross-domain Retrieval in the Legal and Patent Domains: a Reproducibility Study 21 Dec 2020 · 1 repository · arXiv:2012.11405
-
Domain specific BERT representation for Named Entity Recognition of lab protocol 21 Dec 2020 · 1 repository · arXiv:2012.11145
-
Encoding Syntactic Knowledge in Transformer Encoder for Intent Detection and Slot Filling 21 Dec 2020 · 0 repositories · arXiv:2012.11689
-
RealFormer: Transformer Likes Residual Attention 21 Dec 2020 · 5 repositories · arXiv:2012.11747
-
Infrared image pedestrian target detection based on Yolov3 and migration learning 21 Dec 2020 · 0 repositories · arXiv:2012.11185
-
Leveraging ParsBERT and Pretrained mT5 for Persian Abstractive Text Summarization 21 Dec 2020 · 1 repository · arXiv:2012.11204
-
Sub-Linear Memory: How to Make Performers SLiM 21 Dec 2020 · 2 repositories · arXiv:2012.11346
-
Towards Incorporating Entity-specific Knowledge Graph Information in Predicting Drug-Drug Interactions 21 Dec 2020 · 0 repositories · arXiv:2012.11142
-
A hybrid deep-learning approach for complex biochemical named entity recognition 20 Dec 2020 · 0 repositories · arXiv:2012.10824
-
Adaptive Bi-directional Attention: Exploring Multi-Granularity Representations for Machine Reading Comprehension 20 Dec 2020 · 0 repositories · arXiv:2012.10877
-
AdnFM: An Attentive DenseNet based Factorization Machine for CTR Prediction 20 Dec 2020 · 0 repositories · arXiv:2012.10820
-
Color Channel Perturbation Attacks for Fooling Convolutional Neural Networks and A Defense Against Such Attacks 20 Dec 2020 · 1 repository · arXiv:2012.14456
-
Computer Vision based Accident Detection for Autonomous Vehicles 20 Dec 2020 · 0 repositories · arXiv:2012.10870
-
Computer Vision based Animal Collision Avoidance Framework for Autonomous Vehicles 20 Dec 2020 · 0 repositories · arXiv:2012.10878
-
Multi-Head Linear Attention Generative Adversarial Network for Thin Cloud Removal 20 Dec 2020 · 0 repositories · arXiv:2012.10898
-
Breaking Writer's Block: Low-cost Fine-tuning of Natural Language Generation Models 19 Dec 2020 · 0 repositories · arXiv:2101.03216
-
Enabling Retrain-free Deep Neural Network Pruning using Surrogate Lagrangian Relaxation 18 Dec 2020 · 0 repositories · arXiv:2012.10079
-
An Empirical Study of Using Pre-trained BERT Models for Vietnamese Relation Extraction Task at VLSP 2020 18 Dec 2020 · 1 repository · arXiv:2012.10275
-
HateXplain: A Benchmark Dataset for Explainable Hate Speech Detection 18 Dec 2020 · 6 repositories · arXiv:2012.10289
-
NeurST: Neural Speech Translation Toolkit 18 Dec 2020 · 1 repository · arXiv:2012.10018
-
On Modality Bias in the TVQA Dataset 18 Dec 2020 · 1 repository · arXiv:2012.10210
-
Regularized Attentive Capsule Network for Overlapped Relation Extraction 18 Dec 2020 · 0 repositories · arXiv:2012.10187
-
SCNet: Training Inference Sample Consistency for Instance Segmentation 18 Dec 2020 · 2 repositories · arXiv:2012.10150
-
Toward Streaming ASR with Non-Autoregressive Insertion-based Model 18 Dec 2020 · 0 repositories · arXiv:2012.10128
-
A Generalization of Transformer Networks to Graphs 17 Dec 2020 · 3 repositories · arXiv:2012.09699Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)