Methods › General › Output Functions › Softmax › Papers, page 304
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 304 of 375: papers 30,301 to 30,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Variational Relational Point Completion Network 20 Apr 2021 · 1 repository · arXiv:2104.10154
-
WASSA@IITK at WASSA 2021: Multi-task Learning and Transformer Finetuning for Emotion Classification and Empathy Prediction 20 Apr 2021 · 0 repositories · arXiv:2104.09827
-
A novel time-frequency Transformer based on self-attention mechanism and its application in fault diagnosis of rolling bearings 19 Apr 2021 · 0 repositories · arXiv:2104.09079
-
Advanced Long-context End-to-end Speech Recognition Using Context-expanded Transformers 19 Apr 2021 · 0 repositories · arXiv:2104.09426
-
BigGreen at SemEval-2021 Task 1: Lexical Complexity Prediction with Assembly Models 19 Apr 2021 · 1 repository · arXiv:2104.09040
-
BM-NAS: Bilevel Multimodal Neural Architecture Search 19 Apr 2021 · 1 repository · arXiv:2104.09379
-
Code Structure Guided Transformer for Source Code Summarization 19 Apr 2021 · 0 repositories · arXiv:2104.09340
-
ELECTRAMed: a new pre-trained language representation model for biomedical NLP 19 Apr 2021 · 2 repositories · arXiv:2104.09585
-
Extracting Temporal Event Relation with Syntax-guided Graph Transformer 19 Apr 2021 · 1 repository · arXiv:2104.09570
-
Improving Transformer-Kernel Ranking Model Using Conformer and Query Term Independence 19 Apr 2021 · 0 repositories · arXiv:2104.09393
-
Memory Efficient 3D U-Net with Reversible Mobile Inverted Bottlenecks for Brain Tumor Segmentation 19 Apr 2021 · 0 repositories · arXiv:2104.09648
-
Modeling "Newsworthiness" for Lead-Generation Across Corpora 19 Apr 2021 · 0 repositories · arXiv:2104.09653
-
Multi-Modal Fusion Transformer for End-to-End Autonomous Driving 19 Apr 2021 · 2 repositories · arXiv:2104.09224Syntology official (archive's flag): 3 ran · 9 ran (of which 7 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
Neural Language Models with Distant Supervision to Identify Major Depressive Disorder from Clinical Notes 19 Apr 2021 · 0 repositories · arXiv:2104.09644
-
OCTIS: Comparing and Optimizing Topic models is Simple! 19 Apr 2021 · 1 repository
-
Operationalizing a National Digital Library: The Case for a Norwegian Transformer Model 19 Apr 2021 · 2 repositories · arXiv:2104.09617
-
Probing for Bridging Inference in Transformer Language Models 19 Apr 2021 · 1 repository · arXiv:2104.09400
-
Sentiment Classification in Swahili Language Using Multilingual BERT 19 Apr 2021 · 0 repositories · arXiv:2104.09006
-
TeamUNCC@LT-EDI-EACL2021: Hope Speech Detection using Transfer Learning with Transformers 19 Apr 2021 · 1 repository
-
TetraPackNet: Four-Corner-Based Object Detection in Logistics Use-Cases 19 Apr 2021 · 0 repositories · arXiv:2104.09123
-
TransCrowd: weakly-supervised crowd counting with transformers 19 Apr 2021 · 1 repository · arXiv:2104.09116
-
A Token-level Reference-free Hallucination Detection Benchmark for Free-form Text Generation 18 Apr 2021 · 2 repositories · arXiv:2104.08704Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Attention-based Clinical Note Summarization 18 Apr 2021 · 1 repository · arXiv:2104.08942
-
CEAR: Cross-Entity Aware Reranker for Knowledge Base Completion 18 Apr 2021 · 0 repositories · arXiv:2104.08741
-
A Simple and Effective Positional Encoding for Transformers 18 Apr 2021 · 0 repositories · arXiv:2104.08698
-
Distributed NLI: Learning to Predict Human Opinion Distributions for Language Reasoning 18 Apr 2021 · 1 repository · arXiv:2104.08676
-
Dual-View Distilled BERT for Sentence Embedding 18 Apr 2021 · 0 repositories · arXiv:2104.08675
-
Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity 18 Apr 2021 · 2 repositories · arXiv:2104.08786Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples)
-
FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks 18 Apr 2021 · 1 repository · arXiv:2104.08815
-
GPT3Mix: Leveraging Large-scale Language Models for Text Augmentation 18 Apr 2021 · 1 repository · arXiv:2104.08826Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Knowledge Neurons in Pretrained Transformers 18 Apr 2021 · 3 repositories · arXiv:2104.08696Syntology official (archive's flag): 1 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Language in a (Search) Box: Grounding Language Learning in Real-World Human-Machine Interaction 18 Apr 2021 · 0 repositories · arXiv:2104.08874
-
MT6: Multilingual Pretrained Text-to-Text Transformer with Translation Pairs 18 Apr 2021 · 2 repositories · arXiv:2104.08692
-
Cross-Task Generalization via Natural Language Crowdsourcing Instructions 18 Apr 2021 · 3 repositories · arXiv:2104.08773Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Cross-Attention is All You Need: Adapting Pretrained Transformers for Machine Translation 18 Apr 2021 · 1 repository · arXiv:2104.08771Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Rethinking Network Pruning -- under the Pre-train and Fine-tune Paradigm 18 Apr 2021 · 1 repository · arXiv:2104.08682Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Revealing Persona Biases in Dialogue Systems 18 Apr 2021 · 1 repository · arXiv:2104.08728
-
SimCSE: Simple Contrastive Learning of Sentence Embeddings 18 Apr 2021 · 23 repositories · arXiv:2104.08821Syntology community repositories only · 17 ran (of which 9 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 5 where Syntology's instrument failed) · 13 unverified (of 30 harvested samples) · 19 pointer-only (licence)
-
The Power of Scale for Parameter-Efficient Prompt Tuning 18 Apr 2021 · 12 repositories · arXiv:2104.08691Syntology community repositories only · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
When Does Pretraining Help? Assessing Self-Supervised Learning for Law and the CaseHOLD Dataset 18 Apr 2021 · 2 repositories · arXiv:2104.08671Syntology official (archive's flag): 1 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
Zero-shot Cross-lingual Transfer of Neural Machine Translation with Multilingual Pretrained Encoders 18 Apr 2021 · 1 repository · arXiv:2104.08757Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A multilabel approach to morphosyntactic probing 17 Apr 2021 · 0 repositories · arXiv:2104.08464
-
ASBERT: Siamese and Triplet network embedding for open question answering 17 Apr 2021 · 0 repositories · arXiv:2104.08558
-
Co-BERT: A Context-Aware BERT Retrieval Model Incorporating Local and Query-specific Context 17 Apr 2021 · 0 repositories · arXiv:2104.08523
-
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP 17 Apr 2021 · 2 repositories · arXiv:2104.08620
-
Frequency-based Distortions in Contextualized Word Embeddings 17 Apr 2021 · 0 repositories · arXiv:2104.08465
-
Three-level Hierarchical Transformer Networks for Long-sequence and Multiple Clinical Documents Classification 17 Apr 2021 · 1 repository · arXiv:2104.08444
-
Higher Order Recurrent Space-Time Transformer for Video Action Prediction 17 Apr 2021 · 1 repository · arXiv:2104.08665
-
Identifying the Limits of Cross-Domain Knowledge Transfer for Pretrained Models 17 Apr 2021 · 1 repository · arXiv:2104.08410
-
Improving Zero-Shot Cross-Lingual Transfer Learning via Robust Training 17 Apr 2021 · 1 repository · arXiv:2104.08645
-
Multi-source Neural Topic Modeling in Multi-view Embedding Spaces 17 Apr 2021 · 1 repository · arXiv:2104.08551
-
RefineMask: Towards High-Quality Instance Segmentation with Fine-Grained Features 17 Apr 2021 · 1 repository · arXiv:2104.08569
-
The Topic Confusion Task: A Novel Scenario for Authorship Attribution 17 Apr 2021 · 0 repositories · arXiv:2104.08530
-
Towards Efficient Convolutional Network Models with Filter Distribution Templates 17 Apr 2021 · 0 repositories · arXiv:2104.08446
-
UPB at SemEval-2021 Task 5: Virtual Adversarial Training for Toxic Spans Detection 17 Apr 2021 · 0 repositories · arXiv:2104.08635
-
Vision Transformer Pruning 17 Apr 2021 · 2 repositories · arXiv:2104.08500Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Zero-shot Slot Filling with DPR and RAG 17 Apr 2021 · 2 repositories · arXiv:2104.08610
-
An Adversarially-Learned Turing Test for Dialog Generation Models 16 Apr 2021 · 1 repository · arXiv:2104.08231
-
An Analysis of a BERT Deep Learning Strategy on a Technology Assisted Review Task 16 Apr 2021 · 0 repositories · arXiv:2104.08340
-
Memorisation versus Generalisation in Pre-trained Language Models 16 Apr 2021 · 1 repository · arXiv:2105.00828Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Comparison of Grammatical Error Correction Using Back-Translation Models 16 Apr 2021 · 0 repositories · arXiv:2104.07848
-
Editing Factual Knowledge in Language Models 16 Apr 2021 · 3 repositories · arXiv:2104.08164Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Fast, Effective, and Self-Supervised: Transforming Masked Language Models into Universal Lexical and Sentence Encoders 16 Apr 2021 · 1 repository · arXiv:2104.08027Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Condenser: a Pre-training Architecture for Dense Retrieval 16 Apr 2021 · 1 repository · arXiv:2104.08253Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
Membership Inference Attack Susceptibility of Clinical Language Models 16 Apr 2021 · 0 repositories · arXiv:2104.08305
-
Probing Across Time: What Does RoBERTa Know and When? 16 Apr 2021 · 1 repository · arXiv:2104.07885Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Counter-Interference Adapter for Multilingual Machine Translation 16 Apr 2021 · 1 repository · arXiv:2104.08154
-
Surface Form Competition: Why the Highest Probability Answer Isn't Always Right 16 Apr 2021 · 2 repositories · arXiv:2104.08315Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Temporal Adaptation of BERT and Performance on Downstream Document Classification: Insights from Social Media 16 Apr 2021 · 2 repositories · arXiv:2104.08116
-
Text2App: A Framework for Creating Android Apps from Text Descriptions 16 Apr 2021 · 2 repositories · arXiv:2104.08301
-
Towards Variable-Length Textual Adversarial Attacks 16 Apr 2021 · 0 repositories · arXiv:2104.08139
-
A Sample-Based Training Method for Distantly Supervised Relation Extraction with Pre-Trained Transformers 15 Apr 2021 · 0 repositories · arXiv:2104.07512
-
A Survey of Recent Abstract Summarization Techniques 15 Apr 2021 · 0 repositories · arXiv:2105.00824
-
Adaptive Sparse Transformer for Multilingual Translation 15 Apr 2021 · 0 repositories · arXiv:2104.07358
-
Are Multilingual BERT models robust? A Case Study on Adversarial Attacks for Multilingual Question Answering 15 Apr 2021 · 0 repositories · arXiv:2104.07646
-
BERT based Transformers lead the way in Extraction of Health Information from Social Media 15 Apr 2021 · 1 repository · arXiv:2104.07367
-
Cross-domain Speech Recognition with Unsupervised Character-level Distribution Matching 15 Apr 2021 · 1 repository · arXiv:2104.07491
-
Robust Optimization for Multilingual Translation with Imbalanced Data 15 Apr 2021 · 0 repositories · arXiv:2104.07639
-
Does BERT Pretrained on Clinical Notes Reveal Sensitive Data? 15 Apr 2021 · 4 repositories · arXiv:2104.07762Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Emotion Dynamics Modeling via BERT 15 Apr 2021 · 0 repositories · arXiv:2104.07252
-
Estimation of atrial fibrillation from lead-I ECGs: Comparison with cardiologists and machine learning model (CurAlive), a clinical validation study 15 Apr 2021 · 0 repositories · arXiv:2104.07427
-
ExplaGraphs: An Explanation Graph Generation Task for Structured Commonsense Reasoning 15 Apr 2021 · 1 repository · arXiv:2104.07644Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
How to Train BERT with an Academic Budget 15 Apr 2021 · 4 repositories · arXiv:2104.07705
-
Meta Faster R-CNN: Towards Accurate Few-Shot Object Detection with Attentive Feature Alignment 15 Apr 2021 · 2 repositories · arXiv:2104.07719
-
NT5?! Training T5 to Perform Numerical Reasoning 15 Apr 2021 · 1 repository · arXiv:2104.07307
-
Points as Queries: Weakly Semi-supervised Object Detection by Points 15 Apr 2021 · 1 repository · arXiv:2104.07434Syntology 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Natural Language Understanding with Privacy-Preserving BERT 15 Apr 2021 · 0 repositories · arXiv:2104.07504
-
Rethinking Text Line Recognition Models 15 Apr 2021 · 0 repositories · arXiv:2104.07787
-
Self-supervised Video Object Segmentation by Motion Grouping 15 Apr 2021 · 0 repositories · arXiv:2104.07658
-
Shoulder Implant X-Ray Manufacturer Classification: Exploring with Vision Transformer 15 Apr 2021 · 1 repository · arXiv:2104.07667
-
SINA-BERT: A pre-trained Language Model for Analysis of Medical Texts in Persian 15 Apr 2021 · 0 repositories · arXiv:2104.07613
-
Syntax-Aware Graph-to-Graph Transformer for Semantic Role Labelling 15 Apr 2021 · 0 repositories · arXiv:2104.07704
-
Text Guide: Improving the quality of long text classification by a text selection method based on feature importance 15 Apr 2021 · 1 repository · arXiv:2104.07225
-
TorontoCL at CMCL 2021 Shared Task: RoBERTa with Multi-Stage Fine-Tuning for Eye-Tracking Prediction 15 Apr 2021 · 1 repository · arXiv:2104.07244
-
Ultra-High Dimensional Sparse Representations with Binarization for Efficient Text Retrieval 15 Apr 2021 · 0 repositories · arXiv:2104.07198
-
UIT-E10dot3 at SemEval-2021 Task 5: Toxic Spans Detection with Named Entity Recognition and Question-Answering Approaches 15 Apr 2021 · 0 repositories · arXiv:2104.07376
-
Vision Transformer using Low-level Chest X-ray Feature Corpus for COVID-19 Diagnosis and Severity Quantification 15 Apr 2021 · 0 repositories · arXiv:2104.07235
-
An Interpretability Illusion for BERT 14 Apr 2021 · 0 repositories · arXiv:2104.07143
-
An Introduction of mini-AlphaStar 14 Apr 2021 · 1 repository · arXiv:2104.06890
-
Decoupled Spatial-Temporal Transformer for Video Inpainting 14 Apr 2021 · 1 repository · arXiv:2104.06637