Methods › General › Output Functions › Softmax › Papers, page 236
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 236 of 375: papers 23,501 to 23,600 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RCDT: Relational Remote Sensing Change Detection with Transformer 9 Dec 2022 · 1 repository · arXiv:2212.04869
-
Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints 9 Dec 2022 · 1 repository · arXiv:2212.05055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Turing Deception 9 Dec 2022 · 0 repositories · arXiv:2212.06721
-
TRBLLmaker -- Transformer Reads Between Lyrics Lines maker 9 Dec 2022 · 0 repositories · arXiv:2212.04917
-
Visual Detection of Personal Protective Equipment and Safety Gear on Industry Workers 9 Dec 2022 · 0 repositories · arXiv:2212.04794
-
DC-MBR: Distributional Cooling for Minimum Bayesian Risk Decoding 8 Dec 2022 · 0 repositories · arXiv:2212.04205
-
Explain to me like I am five -- Sentence Simplification Using Transformers 8 Dec 2022 · 1 repository · arXiv:2212.04595
-
Federated Learning for Inference at Anytime and Anywhere 8 Dec 2022 · 0 repositories · arXiv:2212.04084
-
Group Generalized Mean Pooling for Vision Transformer 8 Dec 2022 · 0 repositories · arXiv:2212.04114
-
Harnessing the Power of Multi-Task Pretraining for Ground-Truth Level Natural Language Explanations 8 Dec 2022 · 1 repository · arXiv:2212.04231Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models 8 Dec 2022 · 1 repository · arXiv:2212.04088
-
NP4G : Network Programming for Generalization 8 Dec 2022 · 1 repository · arXiv:2212.11118
-
NRTR: Neuron Reconstruction with Transformer from 3D Optical Microscopy Images 8 Dec 2022 · 0 repositories · arXiv:2212.04163
-
The Role of AI in Drug Discovery: Challenges, Opportunities, and Strategies 8 Dec 2022 · 0 repositories · arXiv:2212.08104
-
Memorization of Named Entities in Fine-tuned BERT Models 7 Dec 2022 · 1 repository · arXiv:2212.03749
-
DeepSpeed Data Efficiency: Improving Deep Learning Model Quality and Training Efficiency via Efficient Data Sampling and Routing 7 Dec 2022 · 1 repository · arXiv:2212.03597
-
Gaussian Radar Transformer for Semantic Segmentation in Noisy Radar Data 7 Dec 2022 · 0 repositories · arXiv:2212.03690
-
Hierarchical multimodal transformers for Multi-Page DocVQA 7 Dec 2022 · 1 repository · arXiv:2212.05935
-
Learning-To-Embed: Adopting Transformer based models for E-commerce Products Representation Learning 7 Dec 2022 · 0 repositories · arXiv:2212.03725
-
Multimodal Vision Transformers with Forced Attention for Behavior Analysis 7 Dec 2022 · 0 repositories · arXiv:2212.03968
-
Multiple Object Tracking Challenge Technical Report for Team MT_IoT 7 Dec 2022 · 1 repository · arXiv:2212.03586
-
SimVTP: Simple Video Text Pre-training with Masked Autoencoders 7 Dec 2022 · 0 repositories · arXiv:2212.03490
-
Slimmable Pruned Neural Networks 7 Dec 2022 · 1 repository · arXiv:2212.03415
-
TweetDrought: A Deep-Learning Drought Impacts Recognizer based on Twitter Data 7 Dec 2022 · 0 repositories · arXiv:2212.04001
-
ViTPose++: Vision Transformer for Generic Body Pose Estimation 7 Dec 2022 · 2 repositories · arXiv:2212.04246Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A K-variate Time Series Is Worth K Words: Evolution of the Vanilla Transformer Architecture for Long-term Multivariate Time Series Forecasting 6 Dec 2022 · 0 repositories · arXiv:2212.02789
-
AbHE: All Attention-based Homography Estimation 6 Dec 2022 · 0 repositories · arXiv:2212.03029
-
Adaptive Testing of Computer Vision Models 6 Dec 2022 · 1 repository · arXiv:2212.02774Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
An advanced YOLOv3 method for small object detection 6 Dec 2022 · 0 repositories · arXiv:2212.02809
-
Controlled Text Generation using T5 based Encoder-Decoder Soft Prompt Tuning and Analysis of the Utility of Generated Text in AI 6 Dec 2022 · 0 repositories · arXiv:2212.02924
-
Counterfactual reasoning: Do language models need world knowledge for causal understanding? 6 Dec 2022 · 1 repository · arXiv:2212.03278
-
CySecBERT: A Domain-Adapted Language Model for the Cybersecurity Domain 6 Dec 2022 · 0 repositories · arXiv:2212.02974
-
Document-Level Abstractive Summarization 6 Dec 2022 · 1 repository · arXiv:2212.03013
-
Vision Transformer Computation and Resilience for Dynamic Inference 6 Dec 2022 · 0 repositories · arXiv:2212.02687
-
Event-based Monocular Dense Depth Estimation with Recurrent Transformers 6 Dec 2022 · 0 repositories · arXiv:2212.02791
-
FacT: Factor-Tuning for Lightweight Adaptation on Vision Transformer 6 Dec 2022 · 1 repository · arXiv:2212.03145Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Hybrid Model using Feature Extraction and Non-linear SVM for Brain Tumor Classification 6 Dec 2022 · 0 repositories · arXiv:2212.02794
-
IncepFormer: Efficient Inception Transformer with Pyramid Pooling for Semantic Segmentation 6 Dec 2022 · 1 repository · arXiv:2212.03035
-
LUNA: Language Understanding with Number Augmentations on Transformers via Number Plugins and Pre-training 6 Dec 2022 · 1 repository · arXiv:2212.02691Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MobilePTX: Sparse Coding for Pneumothorax Detection Given Limited Training Examples 6 Dec 2022 · 0 repositories · arXiv:2212.03282
-
Modern French Poetry Generation with RoBERTa and GPT-2 6 Dec 2022 · 0 repositories · arXiv:2212.02911
-
Open World DETR: Transformer based Open World Object Detection 6 Dec 2022 · 0 repositories · arXiv:2212.02969
-
Pretrained Diffusion Models for Unified Human Motion Synthesis 6 Dec 2022 · 0 repositories · arXiv:2212.02837
-
Semantic-aware Message Broadcasting for Efficient Unsupervised Domain Adaptation 6 Dec 2022 · 1 repository · arXiv:2212.02739
-
Semantic-Conditional Diffusion Networks for Image Captioning 6 Dec 2022 · 2 repositories · arXiv:2212.03099
-
Simple Baseline for Weather Forecasting Using Spatiotemporal Context Aggregation Network 6 Dec 2022 · 1 repository · arXiv:2212.02952Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Reinforcement Learning for Molecular Dynamics Optimization: A Stochastic Pontryagin Maximum Principle Approach 6 Dec 2022 · 1 repository · arXiv:2212.03320
-
Style transfer and classification in hebrew news items 6 Dec 2022 · 0 repositories · arXiv:2212.03019
-
UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression 6 Dec 2022 · 2 repositories · arXiv:2212.02746Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples) · 4 pointer-only (licence)
-
Video Object of Interest Segmentation 6 Dec 2022 · 0 repositories · arXiv:2212.02871
-
3D-LatentMapper: View Agnostic Single-View Reconstruction of 3D Shapes 5 Dec 2022 · 0 repositories · arXiv:2212.02184
-
Audio-Driven Co-Speech Gesture Video Generation 5 Dec 2022 · 0 repositories · arXiv:2212.02350
-
Automatic Generation of Factual News Headlines in Finnish 5 Dec 2022 · 0 repositories · arXiv:2212.02170
-
Evince the artifacts of Spoof Speech by blending Vocal Tract and Voice Source Features 5 Dec 2022 · 0 repositories · arXiv:2212.02013
-
FBLNet: FeedBack Loop Network for Driver Attention Prediction 5 Dec 2022 · 0 repositories · arXiv:2212.02096
-
LMEC: Learnable Multiplicative Absolute Position Embedding Based Conformer for Speech Recognition 5 Dec 2022 · 1 repository · arXiv:2212.02099
-
Mask Matching Transformer for Few-Shot Segmentation 5 Dec 2022 · 1 repository · arXiv:2301.01208
-
Retrieval as Attention: End-to-end Learning of Retrieval and Reading within a Single Transformer 5 Dec 2022 · 1 repository · arXiv:2212.02027
-
This changes to that : Combining causal and non-causal explanations to generate disease progression in capsule endoscopy 5 Dec 2022 · 1 repository · arXiv:2212.02506
-
Unifying Vision, Text, and Layout for Universal Document Processing 5 Dec 2022 · 5 repositories · arXiv:2212.02623Syntology official: no sample here; runs from other or unrecorded repositories · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 2 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
Video Games as a Corpus: Sentiment Analysis using Fallout New Vegas Dialog 5 Dec 2022 · 0 repositories · arXiv:2212.02168
-
Joint Self-Supervised Image-Volume Representation Learning with Intra-Inter Contrastive Clustering 4 Dec 2022 · 0 repositories · arXiv:2212.01893
-
Languages You Know Influence Those You Learn: Impact of Language Characteristics on Multi-Lingual Text-to-Text Transfer 4 Dec 2022 · 0 repositories · arXiv:2212.01757
-
A Domain-specific Perceptual Metric via Contrastive Self-supervised Representation: Applications on Natural and Medical Images 3 Dec 2022 · 0 repositories · arXiv:2212.01577
-
Exploring Stochastic Autoregressive Image Modeling for Visual Representation 3 Dec 2022 · 1 repository · arXiv:2212.01610
-
Exploring the Limits of Differentially Private Deep Learning with Group-wise Clipping 3 Dec 2022 · 0 repositories · arXiv:2212.01539
-
Global memory transformer for processing long documents 3 Dec 2022 · 0 repositories · arXiv:2212.01650
-
iEnhancer-ELM: improve enhancer identification by extracting position-related multiscale contextual information based on enhancer language models 3 Dec 2022 · 1 repository · arXiv:2212.01495
-
Recognition and Prediction of Surgical Gestures and Trajectories Using Transformer Models in Robot-Assisted Surgery 3 Dec 2022 · 0 repositories · arXiv:2212.01683
-
Avoiding spurious correlations via logit correction 2 Dec 2022 · 1 repository · arXiv:2212.01433Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
ColD Fusion: Collaborative Descent for Distributed Multitask Finetuning 2 Dec 2022 · 0 repositories · arXiv:2212.01378
-
Event knowledge in large language models: the gap between the impossible and the unlikely 2 Dec 2022 · 1 repository · arXiv:2212.01488
-
FECAM: Frequency Enhanced Channel Attention Mechanism for Time Series Forecasting 2 Dec 2022 · 1 repository · arXiv:2212.01209
-
Improving Training and Inference of Face Recognition Models via Random Temperature Scaling 2 Dec 2022 · 0 repositories · arXiv:2212.01015
-
Nonparametric Masked Language Modeling 2 Dec 2022 · 1 repository · arXiv:2212.01349
-
Relation-Aware Language-Graph Transformer for Question Answering 2 Dec 2022 · 1 repository · arXiv:2212.00975
-
Deciphering RNA Secondary Structure Prediction: A Probabilistic K-Rook Matching Perspective 2 Dec 2022 · 1 repository · arXiv:2212.14041Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 1 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
Multi-scale Transformer Network with Edge-aware Pre-training for Cross-Modality MR Image Synthesis 2 Dec 2022 · 2 repositories · arXiv:2212.01108
-
SumREN: Summarizing Reported Speech about Events in News 2 Dec 2022 · 1 repository · arXiv:2212.01146
-
Tackling Low-Resourced Sign Language Translation: UPC at WMT-SLT 22 2 Dec 2022 · 1 repository · arXiv:2212.01140
-
Towards Diverse, Relevant and Coherent Open-Domain Dialogue Generation via Hybrid Latent Variables 2 Dec 2022 · 0 repositories · arXiv:2212.01145
-
a survey on GPT-3 1 Dec 2022 · 0 repositories · arXiv:2212.00857
-
Adapted Multimodal BERT with Layer-wise Fusion for Sentiment Analysis 1 Dec 2022 · 0 repositories · arXiv:2212.00678
-
CHAPTER: Exploiting Convolutional Neural Network Adapters for Self-supervised Speech Models 1 Dec 2022 · 0 repositories · arXiv:2212.01282
-
Concealed Object Detection for Passive Millimeter-Wave Security Imaging Based on Task-Aligned Detection Transformer 1 Dec 2022 · 0 repositories · arXiv:2212.00313
-
CUNI Non-Autoregressive System for the WMT 22 Efficient Translation Shared Task 1 Dec 2022 · 0 repositories · arXiv:2212.00477
-
Data-Efficient Finetuning Using Cross-Task Nearest Neighbors 1 Dec 2022 · 1 repository · arXiv:2212.00196Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Distilling Reasoning Capabilities into Smaller Language Models 1 Dec 2022 · 1 repository · arXiv:2212.00193
-
Explainable Artificial Intelligence for Improved Modeling of Processes 1 Dec 2022 · 1 repository · arXiv:2212.00695
-
Ghost-free High Dynamic Range Imaging via Hybrid CNN-Transformer and Structure Tensor 1 Dec 2022 · 1 repository · arXiv:2212.00595
-
GMM-IL: Image Classification using Incrementally Learnt, Independent Probabilistic Models for Small Sample Sizes 1 Dec 2022 · 0 repositories · arXiv:2212.00572
-
Learning Progressive Modality-shared Transformers for Effective Visible-Infrared Person Re-identification 1 Dec 2022 · 1 repository · arXiv:2212.00226
-
A Novel Framework for Decentralized Dynamic Resource Allocation Using Voronoi Tessellations 30 Nov 2022 · 0 repositories · arXiv:2212.00140
-
BudgetLongformer: Can we Cheaply Pretrain a SotA Legal Language Model From Scratch? 30 Nov 2022 · 0 repositories · arXiv:2211.17135
-
DSNet: a simple yet efficient network with dual-stream attention for lesion segmentation 30 Nov 2022 · 0 repositories · arXiv:2211.16950
-
ExtremeBERT: A Toolkit for Accelerating Pretraining of Customized BERT 30 Nov 2022 · 1 repository · arXiv:2211.17201Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
HEAT: Hardware-Efficient Automatic Tensor Decomposition for Transformer Compression 30 Nov 2022 · 0 repositories · arXiv:2211.16749
-
Location analysis of players in UEFA EURO 2020 and 2022 using generalized valuation of defense by estimating probabilities 30 Nov 2022 · 1 repository · arXiv:2212.00021
-
Optimizing Explanations by Network Canonization and Hyperparameter Search 30 Nov 2022 · 0 repositories · arXiv:2211.17174
-
Part-based Face Recognition with Vision Transformers 30 Nov 2022 · 1 repository · arXiv:2212.00057Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)