Methods › General › Output Functions › Softmax › Papers, page 244
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 244 of 375: papers 24,301 to 24,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SUMBot: Summarizing Context in Open-Domain Dialogue Systems 12 Oct 2022 · 0 repositories · arXiv:2210.06496
-
Towards Theoretically Inspired Neural Initialization Optimization 12 Oct 2022 · 1 repository · arXiv:2210.05956
-
Uplift and Upsample: Efficient 3D Human Pose Estimation with Uplifting Transformers 12 Oct 2022 · 2 repositories · arXiv:2210.06110Syntology official (archive's flag): 7 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
ZITS++: Image Inpainting by Improving the Incremental Transformer on Structural Priors 12 Oct 2022 · 2 repositories · arXiv:2210.05950Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
A Win-win Deal: Towards Sparse and Robust Pre-trained Language Models 11 Oct 2022 · 1 repository · arXiv:2210.05211Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
An Exploration of Hierarchical Attention Transformers for Efficient Long Document Classification 11 Oct 2022 · 0 repositories · arXiv:2210.05529
-
Are Pretrained Multilingual Models Equally Fair Across Languages? 11 Oct 2022 · 1 repository · arXiv:2210.05457
-
CLIP also Understands Text: Prompting CLIP for Phrase Understanding 11 Oct 2022 · 0 repositories · arXiv:2210.05836
-
Reliable Conditioning of Behavioral Cloning for Offline Reinforcement Learning 11 Oct 2022 · 1 repository · arXiv:2210.05158
-
Enriching Biomedical Knowledge for Low-resource Language Through Large-Scale Translation 11 Oct 2022 · 1 repository · arXiv:2210.05598
-
LARF: Two-level Attention-based Random Forests with a Mixture of Contamination Models 11 Oct 2022 · 1 repository · arXiv:2210.05168
-
Memory transformers for full context and high-resolution 3D Medical Segmentation 11 Oct 2022 · 0 repositories · arXiv:2210.05313
-
Mixture of Attention Heads: Selecting Attention Heads Per Token 11 Oct 2022 · 2 repositories · arXiv:2210.05144Syntology official (archive's flag): 5 ran · 13 ran (of which 2 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Multilingual BERT has an accent: Evaluating English influences on fluency in multilingual models 11 Oct 2022 · 0 repositories · arXiv:2210.05619
-
On the Interpolation of Contextualized Term-based Ranking with BM25 for Query-by-Example Retrieval 11 Oct 2022 · 1 repository · arXiv:2210.05512
-
On the Use of Semantically-Aligned Speech Representations for Spoken Language Understanding 11 Oct 2022 · 0 repositories · arXiv:2210.05291
-
Point Transformer V2: Grouped Vector Attention and Partition-based Pooling 11 Oct 2022 · 2 repositories · arXiv:2210.05666
-
Reflection of Thought: Inversely Eliciting Numerical Reasoning in Language Models via Solving Linear Systems 11 Oct 2022 · 0 repositories · arXiv:2210.05075
-
SaiT: Sparse Vision Transformers through Adaptive Token Pruning 11 Oct 2022 · 1 repository · arXiv:2210.05832
-
Streaming Punctuation for Long-form Dictation with Transformers 11 Oct 2022 · 0 repositories · arXiv:2210.05756
-
T5 for Hate Speech, Augmented Data and Ensemble 11 Oct 2022 · 1 repository · arXiv:2210.05480
-
Understanding the Failure of Batch Normalization for Transformers in NLP 11 Oct 2022 · 1 repository · arXiv:2210.05153Syntology official (archive's flag): 9 ran · 9 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Viterbi Decoding of Directed Acyclic Transformer for Non-Autoregressive Machine Translation 11 Oct 2022 · 1 repository · arXiv:2210.05193Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Vote'n'Rank: Revision of Benchmarking with Social Choice Theory 11 Oct 2022 · 1 repository · arXiv:2210.05769Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 17 harvested samples)
-
A Memory Transformer Network for Incremental Learning 10 Oct 2022 · 0 repositories · arXiv:2210.04485
-
BoundaryFace: A mining framework with noise label self-correction for Face Recognition 10 Oct 2022 · 1 repository · arXiv:2210.04567
-
Characterization of anomalous diffusion through convolutional transformers 10 Oct 2022 · 0 repositories · arXiv:2210.04959
-
DCVQE: A Hierarchical Transformer for Video Quality Assessment 10 Oct 2022 · 0 repositories · arXiv:2210.04377
-
Deep object detection for waterbird monitoring using aerial imagery 10 Oct 2022 · 1 repository · arXiv:2210.04868
-
DEPTWEET: A Typology for Social Media Texts to Detect Depression Severities 10 Oct 2022 · 1 repository · arXiv:2210.05372
-
Empowering the Fact-checkers! Automatic Identification of Claim Spans on Twitter 10 Oct 2022 · 1 repository · arXiv:2210.04710
-
Ensemble Learning using Transformers and Convolutional Networks for Masked Face Recognition 10 Oct 2022 · 1 repository · arXiv:2210.04816
-
FS-DETR: Few-Shot DEtection TRansformer with prompting and without re-training 10 Oct 2022 · 0 repositories · arXiv:2210.04845
-
LAPFormer: A Light and Accurate Polyp Segmentation Transformer 10 Oct 2022 · 0 repositories · arXiv:2210.04393
-
LMQFormer: A Laplace-Prior-Guided Mask Query Transformer for Lightweight Snow Removal 10 Oct 2022 · 1 repository · arXiv:2210.04787
-
MMT: Image-guided Story Ending Generation with Multimodal Memory Transformer 10 Oct 2022 · 1 repository
-
Multi-CLS BERT: An Efficient Alternative to Traditional Ensembling 10 Oct 2022 · 1 repository · arXiv:2210.05043
-
REV: Information-Theoretic Evaluation of Free-Text Rationales 10 Oct 2022 · 1 repository · arXiv:2210.04982Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Revisiting adapters with adversarial training 10 Oct 2022 · 0 repositories · arXiv:2210.04886
-
SCAM! Transferring humans between images with Semantic Cross Attention Modulation 10 Oct 2022 · 1 repository · arXiv:2210.04883
-
The Minimum Wage as an Anchor: Effects on Determinations of Fairness by Humans and AI 10 Oct 2022 · 0 repositories · arXiv:2210.10585
-
Uncertainty Quantification with Pre-trained Language Models: A Large-Scale Empirical Analysis 10 Oct 2022 · 1 repository · arXiv:2210.04714Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Visual Prompt Tuning for Test-time Domain Adaptation 10 Oct 2022 · 0 repositories · arXiv:2210.04831
-
A Transformer-based deep neural network model for SSVEP classification 9 Oct 2022 · 2 repositories · arXiv:2210.04172
-
AMPose: Alternately Mixed Global-Local Attention Model for 3D Human Pose Estimation 9 Oct 2022 · 0 repositories · arXiv:2210.04216
-
ASDOT: Any-Shot Data-to-Text Generation with Pretrained Language Models 9 Oct 2022 · 1 repository · arXiv:2210.04325
-
CHARD: Clinical Health-Aware Reasoning Across Dimensions for Text Generation Models 9 Oct 2022 · 1 repository · arXiv:2210.04191
-
ConTra: (Con)text (Tra)nsformer for Cross-Modal Video Retrieval 9 Oct 2022 · 1 repository · arXiv:2210.04341
-
Controllable Dialogue Simulation with In-Context Learning 9 Oct 2022 · 1 repository · arXiv:2210.04185
-
Deep Span Representations for Named Entity Recognition 9 Oct 2022 · 1 repository · arXiv:2210.04182
-
ELIGN: Expectation Alignment as a Multi-Agent Intrinsic Reward 9 Oct 2022 · 1 repository · arXiv:2210.04365
-
Fine-Grained Detection of Solidarity for Women and Migrants in 155 Years of German Parliamentary Debates 9 Oct 2022 · 2 repositories · arXiv:2210.04359
-
Fine-Tuning Pre-trained Transformers into Decaying Fast Weights 9 Oct 2022 · 1 repository · arXiv:2210.04243Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Better Pre-Training by Reducing Representation Confusion 9 Oct 2022 · 0 repositories · arXiv:2210.04246
-
KSAT: Knowledge-infused Self Attention Transformer -- Integrating Multiple Domain-Specific Contexts 9 Oct 2022 · 0 repositories · arXiv:2210.04307
-
Learning Texture Transformer Network for Light Field Super-Resolution 9 Oct 2022 · 0 repositories · arXiv:2210.09293
-
Multi-Objective Personalized Product Retrieval in Taobao Search 9 Oct 2022 · 0 repositories · arXiv:2210.04170
-
Spread Love Not Hate: Undermining the Importance of Hateful Pre-training for Hate Speech Detection 9 Oct 2022 · 1 repository · arXiv:2210.04267
-
Strong Gravitational Lensing Parameter Estimation with Vision Transformer 9 Oct 2022 · 1 repository · arXiv:2210.04143Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Test-time Recalibration of Conformal Predictors Under Distribution Shift Based on Unlabeled Examples 9 Oct 2022 · 1 repository · arXiv:2210.04166
-
Transformer-based Flood Scene Segmentation for Developing Countries 9 Oct 2022 · 0 repositories · arXiv:2210.04218
-
VoLTA: Vision-Language Transformer with Weakly-Supervised Local-Feature Alignment 9 Oct 2022 · 1 repository · arXiv:2210.04135Syntology official (archive's flag): 10 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 4 pointer-only (licence)
-
AlphaTuning: Quantization-Aware Parameter-Efficient Adaptation of Large-Scale Pre-Trained Language Models 8 Oct 2022 · 0 repositories · arXiv:2210.03858
-
Detaching and Boosting: Dual Engine for Scale-Invariant Self-Supervised Monocular Depth Estimation 8 Oct 2022 · 1 repository · arXiv:2210.03952Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
FBNet: Feedback Network for Point Cloud Completion 8 Oct 2022 · 1 repository · arXiv:2210.03974
-
Hierarchical Graph Transformer with Adaptive Node Sampling 8 Oct 2022 · 1 repository · arXiv:2210.03930Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Improving Data-Efficient Fossil Segmentation via Model Editing 8 Oct 2022 · 0 repositories · arXiv:2210.03879
-
KG-MTT-BERT: Knowledge Graph Enhanced BERT for Multi-Type Medical Text Classification 8 Oct 2022 · 0 repositories · arXiv:2210.03970
-
On Task-Adaptive Pretraining for Dialogue Response Selection 8 Oct 2022 · 0 repositories · arXiv:2210.04073
-
Short Text Pre-training with Extended Token Classification for E-commerce Query Understanding 8 Oct 2022 · 0 repositories · arXiv:2210.03915
-
Towards Light Weight Object Detection System 8 Oct 2022 · 0 repositories · arXiv:2210.03861
-
A Closer Look at Hardware-Friendly Weight Quantization 7 Oct 2022 · 1 repository · arXiv:2210.03671
-
Automatic Chain of Thought Prompting in Large Language Models 7 Oct 2022 · 5 repositories · arXiv:2210.03493Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
DABERT: Dual Attention Enhanced BERT for Semantic Matching 7 Oct 2022 · 0 repositories · arXiv:2210.03454
-
Generating Quizzes to Support Training on Quality Management and Assurance in Space Science and Engineering 7 Oct 2022 · 0 repositories · arXiv:2210.03427
-
How Large Language Models are Transforming Machine-Paraphrased Plagiarism 7 Oct 2022 · 3 repositories · arXiv:2210.03568
-
Knowledge Injected Prompt Based Fine-tuning for Multi-label Few-shot ICD Coding 7 Oct 2022 · 1 repository · arXiv:2210.03304Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Measuring and Narrowing the Compositionality Gap in Language Models 7 Oct 2022 · 1 repository · arXiv:2210.03350
-
LLMEffiChecker: Understanding and Testing Efficiency Degradation of Large Language Models 7 Oct 2022 · 1 repository · arXiv:2210.03696
-
PCAE: A Framework of Plug-in Conditional Auto-Encoder for Controllable Text Generation 7 Oct 2022 · 1 repository · arXiv:2210.03496
-
PS-ARM: An End-to-End Attention-aware Relation Mixer Network for Person Search 7 Oct 2022 · 1 repository · arXiv:2210.03433
-
Time-Space Transformers for Video Panoptic Segmentation 7 Oct 2022 · 0 repositories · arXiv:2210.03546
-
UU-Tax at SemEval-2022 Task 3: Improving the generalizability of language models for taxonomy classification through data augmentation 7 Oct 2022 · 1 repository · arXiv:2210.03378
-
Binding Language Models in Symbolic Languages 6 Oct 2022 · 4 repositories · arXiv:2210.02875Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Brain structure can mediate or moderate the relationship of behavior to brain function and transcriptome. A preliminary study 6 Oct 2022 · 0 repositories · arXiv:2210.03195
-
ByteTransformer: A High-Performance Transformer Boosted for Variable-Length Inputs 6 Oct 2022 · 1 repository · arXiv:2210.03052
-
Explainable Verbal Deception Detection using Transformers 6 Oct 2022 · 0 repositories · arXiv:2210.03080
-
Focal and Global Spatial-Temporal Transformer for Skeleton-based Action Recognition 6 Oct 2022 · 0 repositories · arXiv:2210.02693
-
Forecasting Bitcoin volatility spikes from whale transactions and CryptoQuant data using Synthesizer Transformer models 6 Oct 2022 · 1 repository · arXiv:2211.08281
-
Gastrointestinal Disorder Detection with a Transformer Based Approach 6 Oct 2022 · 0 repositories · arXiv:2210.03168
-
Generalization Properties of Retrieval-based Models 6 Oct 2022 · 0 repositories · arXiv:2210.02617
-
Guess the Instruction! Flipped Learning Makes Language Models Stronger Zero-Shot Learners 6 Oct 2022 · 1 repository · arXiv:2210.02969
-
Improving the Domain Adaptation of Retrieval Augmented Generation (RAG) Models for Open Domain Question Answering 6 Oct 2022 · 1 repository · arXiv:2210.02627
-
Interpreting County Level COVID-19 Infection and Feature Sensitivity using Deep Learning Time Series Models 6 Oct 2022 · 1 repository · arXiv:2210.03258
-
Join-Chain Network: A Logical Reasoning View of the Multi-head Attention in Transformer 6 Oct 2022 · 0 repositories · arXiv:2210.02729
-
LungViT: Ensembling Cascade of Texture Sensitive Hierarchical Vision Transformers for Cross-Volume Chest CT Image-to-Image Translation 6 Oct 2022 · 0 repositories · arXiv:2210.02625
-
Mask3D: Mask Transformer for 3D Semantic Instance Segmentation 6 Oct 2022 · 1 repository · arXiv:2210.03105Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Matching Text and Audio Embeddings: Exploring Transfer-learning Strategies for Language-based Audio Retrieval 6 Oct 2022 · 0 repositories · arXiv:2210.02833
-
MechRetro is a chemical-mechanism-driven graph learning framework for interpretable retrosynthesis prediction and pathway planning 6 Oct 2022 · 1 repository · arXiv:2210.02630
-
Melody Infilling with User-Provided Structural Context 6 Oct 2022 · 1 repository · arXiv:2210.02829