Methods › General › Output Functions › Softmax › Papers, page 348
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 348 of 375: papers 34,701 to 34,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
English to Hindi Multi-modal Neural Machine Translation and Hindi Image Captioning 1 Nov 2019 · 0 repositories
-
Enhanced Transformer Model for Data-to-Text Generation 1 Nov 2019 · 0 repositories
-
Enhancing BERT for Lexical Normalization 1 Nov 2019 · 0 repositories
-
Evaluating BERT for natural language inference: A case study on the CommitmentBank 1 Nov 2019 · 0 repositories
-
Event Causality Recognition Exploiting Multiple Annotators' Judgments and Background Knowledge 1 Nov 2019 · 0 repositories
-
Experimenting with Power Divergences for Language Modeling 1 Nov 2019 · 0 repositories
-
Extract and Aggregate: A Novel Domain-Independent Approach to Factual Data Verification 1 Nov 2019 · 0 repositories
-
Extractive NarrativeQA with Heuristic Pre-Training 1 Nov 2019 · 0 repositories
-
FASPell: A Fast, Adaptable, Simple, Powerful Chinese Spell Checker Based On DAE-Decoder Paradigm 1 Nov 2019 · 1 repository
-
Fine-Grained Propaganda Detection with Fine-Tuned BERT 1 Nov 2019 · 0 repositories
-
Fine-tune BERT with Sparse Self-Attention Mechanism 1 Nov 2019 · 0 repositories
-
From Monolingual to Multilingual FAQ Assistant using Multilingual Co-training 1 Nov 2019 · 0 repositories
-
Fully Unsupervised Crosslingual Semantic Textual Similarity Metric Based on BERT for Identifying Parallel Data 1 Nov 2019 · 0 repositories
-
GEM: Generative Enhanced Model for adversarial attacks 1 Nov 2019 · 0 repositories
-
Generalizing Question Answering System with Pre-trained Language Model Fine-tuning 1 Nov 2019 · 0 repositories
-
HIT-SCIR at MRP 2019: A Unified Pipeline for Meaning Representation Parsing via Efficient Training and Effective Encoding 1 Nov 2019 · 0 repositories
-
How well do NLI models capture verb veridicality? 1 Nov 2019 · 0 repositories
-
Idiap NMT System for WAT 2019 Multimodal Translation Task 1 Nov 2019 · 0 repositories
-
IIT-KGP at COIN 2019: Using pre-trained Language Models for modeling Machine Comprehension 1 Nov 2019 · 0 repositories
-
Improved Differentiable Architecture Search for Language Modeling and Named Entity Recognition 1 Nov 2019 · 1 repository
-
Improving Answer Selection and Answer Triggering using Hard Negatives 1 Nov 2019 · 0 repositories
-
Improving Generalization of Transformer for Speech Recognition with Parallel Schedule Sampling and Relative Positional Embedding 1 Nov 2019 · 0 repositories · arXiv:1911.00203
-
Improving Natural Language Understanding by Reverse Mapping Bytepair Encoding 1 Nov 2019 · 0 repositories
-
Improving Pre-Trained Multilingual Model with Vocabulary Expansion 1 Nov 2019 · 0 repositories
-
Inspecting Unification of Encoding and Matching with Transformer: A Case Study of Machine Reading Comprehension 1 Nov 2019 · 0 repositories
-
JBNU at MRP 2019: Multi-level Biaffine Attention for Semantic Dependency Parsing 1 Nov 2019 · 0 repositories
-
Jeff Da at COIN - Shared Task: BIG MOOD: Relating Transformers to Explicit Commonsense Knowledge 1 Nov 2019 · 0 repositories
-
JUSTDeep at NLP4IF 2019 Task 1: Propaganda Detection using Ensemble Deep Learning Models 1 Nov 2019 · 0 repositories
-
Kernelized Bayesian Softmax for Text Generation 1 Nov 2019 · 1 repository · arXiv:1911.00274
-
Learning to Generate Word- and Phrase-Embeddings for Efficient Phrase-Based Neural Machine Translation 1 Nov 2019 · 0 repositories
-
LexicalAT: Lexical-Based Adversarial Reinforcement Training for Robust Sentiment Classification 1 Nov 2019 · 0 repositories
-
Long Warm-up and Self-Training: Training Strategies of NICT-2 NMT System at WAT-2019 1 Nov 2019 · 0 repositories
-
LTRC-MT Simple & Effective Hindi-English Neural Machine Translation Systems at WAT 2019 1 Nov 2019 · 0 repositories
-
Mixed Multi-Head Self-Attention for Neural Machine Translation 1 Nov 2019 · 0 repositories
-
MrMep: Joint Extraction of Multiple Relations and Multiple Entity Pairs Based on Triplet Attention 1 Nov 2019 · 0 repositories
-
MulCode: A Multiplicative Multi-way Model for Compressing Neural Language Model 1 Nov 2019 · 0 repositories
-
Multi-View Domain Adapted Sentence Embeddings for Low-Resource Unsupervised Duplicate Question Detection 1 Nov 2019 · 0 repositories
-
Named Entity Recognition - Is There a Glass Ceiling? 1 Nov 2019 · 0 repositories
-
Natural Language Generation for Effective Knowledge Distillation 1 Nov 2019 · 1 repository
-
No, you're not alone: A better way to find people with similar experiences on Reddit 1 Nov 2019 · 0 repositories
-
NSIT@NLP4IF-2019: Propaganda Detection from News Articles using Transfer Learning 1 Nov 2019 · 0 repositories
-
On Sentence Representations for Propaganda Detection: From Handcrafted Features to Word Embeddings 1 Nov 2019 · 0 repositories
-
On the Relation between Position Information and Sentence Length in Neural Machine Translation 1 Nov 2019 · 0 repositories
-
Our Neural Machine Translation Systems for WAT 2019 1 Nov 2019 · 0 repositories
-
Pingan Smart Health and SJTU at COIN - Shared Task: utilizing Pre-trained Language Models and Common-sense Knowledge in Machine Reading Tasks 1 Nov 2019 · 0 repositories
-
Policy Preference Detection in Parliamentary Debate Motions 1 Nov 2019 · 0 repositories
-
Pre-Training BERT on Domain Resources for Short Answer Grading 1 Nov 2019 · 0 repositories
-
Question Answering Using Hierarchical Attention on Top of BERT Features 1 Nov 2019 · 0 repositories
-
Recurrent Positional Embedding for Neural Machine Translation 1 Nov 2019 · 0 repositories
-
Recycling a Pre-trained BERT Encoder for Neural Machine Translation 1 Nov 2019 · 0 repositories
-
Reevaluating Argument Component Extraction in Low Resource Settings 1 Nov 2019 · 0 repositories
-
Relation Extraction among Multiple Entities Using a Dual Pointer Network with a Multi-Head Attention Mechanism 1 Nov 2019 · 0 repositories
-
Relation Module for Non-Answerable Predictions on Reading Comprehension 1 Nov 2019 · 0 repositories
-
Sarah's Participation in WAT 2019 1 Nov 2019 · 0 repositories
-
Selecting, Planning, and Rewriting: A Modular Approach for Data-to-Document Generation and Translation 1 Nov 2019 · 0 repositories
-
Self-Adaptive Scaling for Learnable Residual Structure 1 Nov 2019 · 0 repositories
-
Semi-Supervised Semantic Role Labeling with Cross-View Training 1 Nov 2019 · 0 repositories
-
Sentence-Level Propaganda Detection in News Articles with Transfer Learning and BERT-BiLSTM-Capsule Model 1 Nov 2019 · 0 repositories
-
SUDA-Alibaba at MRP 2019: Graph-Based Models with BERT 1 Nov 2019 · 0 repositories
-
SUM-QE: a BERT-based Summary Quality Estimation Model 1 Nov 2019 · 0 repositories
-
Supervised neural machine translation based on data augmentation and improved training & inference process 1 Nov 2019 · 0 repositories
-
SYSTRAN @ WAT 2019: Russian-Japanese News Commentary task 1 Nov 2019 · 0 repositories
-
SYSTRAN @ WNGT 2019: DGT Task 1 Nov 2019 · 0 repositories
-
Team DOMLIN: Exploiting Evidence Enhancement for the FEVER Shared Task 1 Nov 2019 · 0 repositories
-
The Concordia NLG Surface Realizer at SRST 2019 1 Nov 2019 · 0 repositories
-
The Feasibility of Embedding Based Automatic Evaluation for Single Document Summarization 1 Nov 2019 · 0 repositories
-
Transfer Learning in Biomedical Named Entity Recognition: An Evaluation of BERT in the PharmaCoNER task 1 Nov 2019 · 0 repositories
-
Transformer and seq2seq model for Paraphrase Generation 1 Nov 2019 · 0 repositories
-
Transformer-based Model for Single Documents Neural Summarization 1 Nov 2019 · 0 repositories
-
Transformer Dissection: An Unified Understanding for Transformer's Attention via the Lens of Kernel 1 Nov 2019 · 0 repositories
-
``Transforming'' Delete, Retrieve, Generate Approach for Controlled Text Style Transfer 1 Nov 2019 · 0 repositories
-
TUPA at MRP 2019: A Multi-Task Baseline System 1 Nov 2019 · 0 repositories
-
Unsupervised Labeled Parsing with Deep Inside-Outside Recursive Autoencoders 1 Nov 2019 · 0 repositories
-
Visual Detection with Context for Document Layout Analysis 1 Nov 2019 · 0 repositories
-
What Does This Word Mean? Explaining Contextualized Embeddings with Natural Language Definition 1 Nov 2019 · 0 repositories
-
When Choosing Plausible Alternatives, Clever Hans can be Clever 1 Nov 2019 · 0 repositories · arXiv:1911.00225
-
Attention Is All You Need for Chinese Word Segmentation 31 Oct 2019 · 1 repository · arXiv:1910.14537
-
Device-Circuit-Architecture Co-Exploration for Computing-in-Memory Neural Accelerators 31 Oct 2019 · 0 repositories · arXiv:1911.00139
-
DiaNet: BERT and Hierarchical Attention Multi-Task Learning of Fine-Grained Dialect 31 Oct 2019 · 0 repositories · arXiv:1910.14243
-
Do Multi-hop Readers Dream of Reasoning Chains? 31 Oct 2019 · 1 repository · arXiv:1910.14520Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Document-level Neural Machine Translation with Associated Memory Network 31 Oct 2019 · 0 repositories · arXiv:1910.14528
-
Human-centric Metric for Accelerating Pathology Reports Annotation 31 Oct 2019 · 0 repositories · arXiv:1911.01226
-
Image-Conditioned Graph Generation for Road Network Extraction 31 Oct 2019 · 3 repositories · arXiv:1910.14388
-
LIMIT-BERT : Linguistic Informed Multi-Task BERT 31 Oct 2019 · 0 repositories · arXiv:1910.14296
-
Multi-Stage Document Ranking with BERT 31 Oct 2019 · 3 repositories · arXiv:1910.14424Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
NAT: Neural Architecture Transformer for Accurate and Compact Architectures 31 Oct 2019 · 1 repository · arXiv:1910.14488
-
Neural Assistant: Joint Action Prediction, Response Generation, and Latent Knowledge Reasoning 31 Oct 2019 · 1 repository · arXiv:1910.14613
-
On Neural Architecture Search for Resource-Constrained Hardware Platforms 31 Oct 2019 · 0 repositories · arXiv:1911.00105
-
On the Interaction Between Deep Detectors and Siamese Trackers in Video Surveillance 31 Oct 2019 · 0 repositories · arXiv:1910.14552
-
Parameter Sharing Decoder Pair for Auto Composing 31 Oct 2019 · 0 repositories · arXiv:1910.14270
-
Positional Attention-based Frame Identification with BERT: A Deep Learning Approach to Target Disambiguation and Semantic Frame Selection 31 Oct 2019 · 0 repositories · arXiv:1910.14549
-
Masked Language Model Scoring 31 Oct 2019 · 6 repositories · arXiv:1910.14659Syntology official: no sample here; runs from other or unrecorded repositories · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Transfer Learning from Transformers to Fake News Challenge Stance Detection (FNC-1) Task 31 Oct 2019 · 0 repositories · arXiv:1910.14353
-
Very high resolution Airborne PolSAR Image Classification using Convolutional Neural Networks 31 Oct 2019 · 0 repositories · arXiv:1910.14578
-
Visual Appearance Based Person Retrieval in Unconstrained Environment Videos 31 Oct 2019 · 0 repositories · arXiv:1910.14565
-
An Augmented Transformer Architecture for Natural Language Generation Tasks 30 Oct 2019 · 0 repositories · arXiv:1910.13634
-
Discourse-Aware Neural Extractive Text Summarization 30 Oct 2019 · 1 repository · arXiv:1910.14142Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Generalization in multitask deep neural classifiers: a statistical physics approach 30 Oct 2019 · 0 repositories · arXiv:1910.13593
-
Lightweight and Efficient End-to-End Speech Recognition Using Low-Rank Transformer 30 Oct 2019 · 0 repositories · arXiv:1910.13923
-
Lsh-sampling Breaks the Computation Chicken-and-egg Loop in Adaptive Stochastic Gradient Estimation 30 Oct 2019 · 0 repositories · arXiv:1910.14162