Methods › General › Output Functions › Softmax › Papers, page 240
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 240 of 375: papers 23,901 to 24,000 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Active Example Selection for In-Context Learning 8 Nov 2022 · 1 repository · arXiv:2211.04486Syntology official (archive's flag): 10 ran · 10 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples)
-
Conciseness: An Overlooked Language Task 8 Nov 2022 · 0 repositories · arXiv:2211.04126
-
Discover, Explanation, Improvement: An Automatic Slice Detection Framework for Natural Language Processing 8 Nov 2022 · 0 repositories · arXiv:2211.04476
-
Going for GOAL: A Resource for Grounded Football Commentaries 8 Nov 2022 · 1 repository · arXiv:2211.04534
-
Linear Self-Attention Approximation via Trainable Feedforward Kernel 8 Nov 2022 · 0 repositories · arXiv:2211.04076
-
Pushing the limits of self-supervised speaker verification using regularized distillation framework 8 Nov 2022 · 1 repository · arXiv:2211.04168
-
SimOn: A Simple Framework for Online Temporal Action Localization 8 Nov 2022 · 1 repository · arXiv:2211.04905
-
Splitting expands the application range of Vision Transformer -- variable Vision Transformer (vViT) 8 Nov 2022 · 0 repositories · arXiv:2211.03992
-
AD-BERT: Using Pre-trained contextualized embeddings to Predict the Progression from Mild Cognitive Impairment to Alzheimer's Disease 7 Nov 2022 · 0 repositories · arXiv:2212.06042
-
Automatic Number Plate Recognition (ANPR) with YOLOv3-CNN 7 Nov 2022 · 0 repositories · arXiv:2211.05229
-
Retrieval augmentation of large language models for lay language generation 7 Nov 2022 · 1 repository · arXiv:2211.03818Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
CoNMix for Source-free Single and Multi-target Domain Adaptation 7 Nov 2022 · 1 repository · arXiv:2211.03876Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining 7 Nov 2022 · 0 repositories · arXiv:2211.03594
-
How Much Does Attention Actually Attend? Questioning the Importance of Attention in Pretrained Transformers 7 Nov 2022 · 1 repository · arXiv:2211.03495
-
Interpreting deep learning output for out-of-distribution detection 7 Nov 2022 · 0 repositories · arXiv:2211.03637
-
Hear The Flow: Optical Flow-Based Self-Supervised Visual Sound Source Localization 6 Nov 2022 · 1 repository · arXiv:2211.03019
-
Sequential Transformer for End-to-End Person Search 6 Nov 2022 · 0 repositories · arXiv:2211.04323
-
Suffix Retrieval-Augmented Language Modeling 6 Nov 2022 · 1 repository · arXiv:2211.03053
-
Wall Street Tree Search: Risk-Aware Planning for Offline Reinforcement Learning 6 Nov 2022 · 0 repositories · arXiv:2211.04583
-
A Robust and Low Complexity Deep Learning Model for Remote Sensing Image Classification 5 Nov 2022 · 0 repositories · arXiv:2211.02820
-
Accurate and Reliable Methods for 5G UAV Jamming Identification With Calibrated Uncertainty 5 Nov 2022 · 0 repositories · arXiv:2211.02924
-
ESKNet-An enhanced adaptive selection kernel convolution for breast tumors segmentation 5 Nov 2022 · 1 repository · arXiv:2211.02915
-
Inductive Graph Transformer for Delivery Time Estimation 5 Nov 2022 · 1 repository · arXiv:2211.02863
-
Learning to Infer from Unlabeled Data: A Semi-supervised Learning Approach for Robust Natural Language Inference 5 Nov 2022 · 1 repository · arXiv:2211.02971
-
A Transformer Architecture for Online Gesture Recognition of Mathematical Expressions 4 Nov 2022 · 0 repositories · arXiv:2211.02643
-
A Weakly-Supervised Streaming Multilingual Speech Model with Truly Zero-Shot Capability 4 Nov 2022 · 0 repositories · arXiv:2211.02499
-
BERT-Deep CNN: State-of-the-Art for Sentiment Analysis of COVID-19 Tweets 4 Nov 2022 · 0 repositories · arXiv:2211.09733
-
BERT for Long Documents: A Case Study of Automated ICD Coding 4 Nov 2022 · 0 repositories · arXiv:2211.02519
-
CCATMos: Convolutional Context-aware Transformer Network for Non-intrusive Speech Quality Assessment 4 Nov 2022 · 0 repositories · arXiv:2211.02577
-
Continuous Prompt Tuning Based Textual Entailment Model for E-commerce Entity Typing 4 Nov 2022 · 1 repository · arXiv:2211.02483
-
Deep learning for structural health monitoring: An application to heritage structures 4 Nov 2022 · 0 repositories · arXiv:2211.10351
-
Fraudulent User Detection Via Behavior Information Aggregation Network (BIAN) On Large-Scale Financial Social Network 4 Nov 2022 · 0 repositories · arXiv:2211.06315
-
Generation of Chinese classical poetry based on pre-trained model 4 Nov 2022 · 0 repositories · arXiv:2211.02541
-
OSIC: A New One-Stage Image Captioner Coined 4 Nov 2022 · 0 repositories · arXiv:2211.02321
-
Patch DCT vs LeNet 4 Nov 2022 · 0 repositories · arXiv:2211.02392
-
RCDPT: Radar-Camera fusion Dense Prediction Transformer 4 Nov 2022 · 1 repository · arXiv:2211.02432
-
Real-Time Target Sound Extraction 4 Nov 2022 · 1 repository · arXiv:2211.02250Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
SPEAKER VGG CCT: Cross-corpus Speech Emotion Recognition with Speaker Embedding and Vision Transformers 4 Nov 2022 · 1 repository · arXiv:2211.02366
-
SSDA-YOLO: Semi-supervised Domain Adaptive YOLO for Cross-Domain Object Detection 4 Nov 2022 · 1 repository · arXiv:2211.02213
-
Towards Asteroid Detection in Microlensing Surveys with Deep Learning 4 Nov 2022 · 1 repository · arXiv:2211.02239
-
WaveNets: Wavelet Channel Attention Networks 4 Nov 2022 · 1 repository · arXiv:2211.02695
-
Alternative formulations for gilthead seabream diets: towards a more sustainable production 3 Nov 2022 · 0 repositories · arXiv:2211.02430
-
Channel-Aware Pretraining of Joint Encoder-Decoder Self-Supervised Model for Telephonic-Speech ASR 3 Nov 2022 · 0 repositories · arXiv:2211.01669
-
Crosslingual Generalization through Multitask Finetuning 3 Nov 2022 · 1 repository · arXiv:2211.01786Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Evaluating a Synthetic Image Dataset Generated with Stable Diffusion 3 Nov 2022 · 0 repositories · arXiv:2211.01777
-
Transformers on Multilingual Clause-Level Morphology 3 Nov 2022 · 1 repository · arXiv:2211.01736
-
FedTP: Federated Learning by Transformer Personalization 3 Nov 2022 · 1 repository · arXiv:2211.01572Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Fine-Tuning Language Models via Epistemic Neural Networks 3 Nov 2022 · 1 repository · arXiv:2211.01568Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Graph-Based Multi-Camera Soccer Player Tracker 3 Nov 2022 · 0 repositories · arXiv:2211.02125
-
Pangu-Weather: A 3D High-Resolution Model for Fast and Accurate Global Weather Forecast 3 Nov 2022 · 6 repositories · arXiv:2211.02556
-
PolyBuilding: Polygon Transformer for End-to-End Building Extraction 3 Nov 2022 · 0 repositories · arXiv:2211.01589
-
Rethinking Hierarchies in Pre-trained Plain Vision Transformer 3 Nov 2022 · 0 repositories · arXiv:2211.01785
-
SAP-DETR: Bridging the Gap Between Salient Points and Queries-Based Transformer Detector for Fast Model Convergency 3 Nov 2022 · 1 repository · arXiv:2211.02006
-
Scaling Multimodal Pre-Training via Cross-Modality Gradient Harmonization 3 Nov 2022 · 0 repositories · arXiv:2211.02077
-
Self Similarity Matrix based CNN Filter Pruning 3 Nov 2022 · 0 repositories · arXiv:2211.01814
-
Using Large Pre-Trained Language Model to Assist FDA in Premarket Medical Device 3 Nov 2022 · 0 repositories · arXiv:2212.01217
-
Attention-based Neural Cellular Automata 2 Nov 2022 · 0 repositories · arXiv:2211.01233
-
BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder 2 Nov 2022 · 0 repositories · arXiv:2211.00792
-
eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers 2 Nov 2022 · 2 repositories · arXiv:2211.01324Syntology 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
MAST: Multiscale Audio Spectrogram Transformers 2 Nov 2022 · 1 repository · arXiv:2211.01515
-
MPCFormer: fast, performant and private Transformer inference with MPC 2 Nov 2022 · 1 repository · arXiv:2211.01452
-
Multi-level Distillation of Semantic Knowledge for Pre-training Multilingual Language Model 2 Nov 2022 · 0 repositories · arXiv:2211.01200
-
Pop2Piano : Pop Audio-based Piano Cover Generation 2 Nov 2022 · 4 repositories · arXiv:2211.00895Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Processing Long Legal Documents with Pre-trained Transformers: Modding LegalBERT and Longformer 2 Nov 2022 · 0 repositories · arXiv:2211.00974
-
RegCLR: A Self-Supervised Framework for Tabular Representation Learning in the Wild 2 Nov 2022 · 0 repositories · arXiv:2211.01165
-
SIMD-size aware weight regularization for fast neural vocoding on CPU 2 Nov 2022 · 0 repositories · arXiv:2211.00898
-
Data Level Lottery Ticket Hypothesis for Vision Transformers 2 Nov 2022 · 1 repository · arXiv:2211.01484
-
Transformer-based encoder-encoder architecture for Spoken Term Detection 2 Nov 2022 · 0 repositories · arXiv:2211.01089
-
TSAA: A Two-Stage Anchor Assignment Method towards Anchor Drift in Crowded Object Detection 2 Nov 2022 · 0 repositories · arXiv:2211.00826
-
WITT: A Wireless Image Transmission Transformer for Semantic Communications 2 Nov 2022 · 2 repositories · arXiv:2211.00937
-
ClassActionPrediction: A Challenging Benchmark for Legal Judgment Prediction of Class Action Cases in the US 1 Nov 2022 · 1 repository · arXiv:2211.00582
-
FRSUM: Towards Faithful Abstractive Summarization via Enhancing Factual Robustness 1 Nov 2022 · 0 repositories · arXiv:2211.00294
-
Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small 1 Nov 2022 · 7 repositories · arXiv:2211.00593Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Investigating Content-Aware Neural Text-To-Speech MOS Prediction Using Prosodic and Linguistic Features 1 Nov 2022 · 0 repositories · arXiv:2211.00342
-
KAMEL : Knowledge Analysis with Multitoken Entities in Language Models 1 Nov 2022 · 1 repository
-
An Empirical Study on Data Leakage and Generalizability of Link Prediction Models for Issues and Commits 1 Nov 2022 · 0 repositories · arXiv:2211.00381
-
Pixel-Wise Contrastive Distillation 1 Nov 2022 · 1 repository · arXiv:2211.00218
-
Two-stage LLM Fine-tuning with Less Specialization and More Generalization 1 Nov 2022 · 0 repositories · arXiv:2211.00635
-
Reduce, Reuse, Recycle: Improving Training Efficiency with Distillation 1 Nov 2022 · 0 repositories · arXiv:2211.00683
-
T5lephone: Bridging Speech and Text Self-supervised Models for Spoken Language Understanding via Phoneme level T5 1 Nov 2022 · 1 repository · arXiv:2211.00586
-
Text-Only Training for Image Captioning using Noise-Injected CLIP 1 Nov 2022 · 4 repositories · arXiv:2211.00575Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
VID-Trans-ReID: Enhanced Video Transformers for Person Re-identification 1 Nov 2022 · 1 repository
-
ViT-DeiT: An Ensemble Model for Breast Cancer Histopathological Images Classification 1 Nov 2022 · 0 repositories · arXiv:2211.00749
-
AdaMix: Mixture-of-Adaptations for Parameter-efficient Model Tuning 31 Oct 2022 · 1 repository · arXiv:2210.17451Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Controllable Factuality in Document-Grounded Dialog Systems Using a Noisy Channel Model 31 Oct 2022 · 1 repository · arXiv:2210.17418
-
Efficient Document Retrieval by End-to-End Refining and Quantizing BERT Embedding with Contrastive Product Quantization 31 Oct 2022 · 1 repository · arXiv:2210.17170
-
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers 31 Oct 2022 · 17 repositories · arXiv:2210.17323Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Joint Audio/Text Training for Transformer Rescorer of Streaming Speech Recognition 31 Oct 2022 · 0 repositories · arXiv:2211.00174
-
Leveraging Pre-trained Models for Failure Analysis Triplets Generation 31 Oct 2022 · 0 repositories · arXiv:2210.17497
-
Multi-Camera Calibration Free BEV Representation for 3D Object Detection 31 Oct 2022 · 0 repositories · arXiv:2210.17252
-
Probabilistic Decomposition Transformer for Time Series Forecasting 31 Oct 2022 · 1 repository · arXiv:2210.17393
-
Probability-Dependent Gradient Decay in Large Margin Softmax 31 Oct 2022 · 0 repositories · arXiv:2210.17145
-
QNet: A Quantum-native Sequence Encoder Architecture 31 Oct 2022 · 0 repositories · arXiv:2210.17262
-
QuaLA-MiniLM: a Quantized Length Adaptive MiniLM 31 Oct 2022 · 2 repositories · arXiv:2210.17114
-
Revisiting Attention Weights as Explanations from an Information Theoretic Perspective 31 Oct 2022 · 0 repositories · arXiv:2211.07714
-
SDCL: Self-Distillation Contrastive Learning for Chinese Spell Checking 31 Oct 2022 · 0 repositories · arXiv:2210.17168
-
SEVGGNet-LSTM: a fused deep learning model for ECG classification 31 Oct 2022 · 0 repositories · arXiv:2210.17111
-
Spatial-Temporal Synchronous Graph Transformer network (STSGT) for COVID-19 forecasting 31 Oct 2022 · 1 repository · arXiv:2211.00082
-
SSD-LM: Semi-autoregressive Simplex-based Diffusion Language Model for Text Generation and Modular Control 31 Oct 2022 · 2 repositories · arXiv:2210.17432
-
Structured State Space Decoder for Speech Recognition and Synthesis 31 Oct 2022 · 0 repositories · arXiv:2210.17098