Methods › General › Output Functions › Softmax › Papers, page 208
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 208 of 375: papers 20,701 to 20,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fine-Tuning Language Models for Scientific Writing Support 19 Jun 2023 · 1 repository · arXiv:2306.10974
-
Generative Sequential Recommendation with GPTRec 19 Jun 2023 · 0 repositories · arXiv:2306.11114
-
High-dimensional Contextual Bandit Problem without Sparsity 19 Jun 2023 · 0 repositories · arXiv:2306.11017
-
Learn to Accumulate Evidence from All Training Samples: Theory and Practice 19 Jun 2023 · 1 repository · arXiv:2306.11113
-
Multi-task Learning for Radar Signal Characterisation 19 Jun 2023 · 1 repository · arXiv:2306.13105
-
Multitrack Music Transcription with a Time-Frequency Perceiver 19 Jun 2023 · 0 repositories · arXiv:2306.10785
-
NAR-Former V2: Rethinking Transformer for Universal Neural Network Representation Learning 19 Jun 2023 · 1 repository · arXiv:2306.10792Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
RaViTT: Random Vision Transformer Tokens 19 Jun 2023 · 0 repositories · arXiv:2306.10959
-
RedMotion: Motion Prediction via Redundancy Reduction 19 Jun 2023 · 3 repositories · arXiv:2306.10840Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SynerGPT: In-Context Learning for Personalized Drug Synergy Prediction and Drug Design 19 Jun 2023 · 0 repositories · arXiv:2307.11694
-
TeleViT: Teleconnection-driven Transformers Improve Subseasonal to Seasonal Wildfire Forecasting 19 Jun 2023 · 1 repository · arXiv:2306.10940
-
Temporal Data Meets LLM -- Explainable Financial Time Series Forecasting 19 Jun 2023 · 0 repositories · arXiv:2306.11025
-
Transformer Training Strategies for Forecasting Multiple Load Time Series 19 Jun 2023 · 1 repository · arXiv:2306.10891
-
Enhanced Masked Image Modeling for Analysis of Dental Panoramic Radiographs 18 Jun 2023 · 1 repository · arXiv:2306.10623
-
Gender Bias in Transformer Models: A comprehensive survey 18 Jun 2023 · 0 repositories · arXiv:2306.10530
-
Instant Soup: Cheap Pruning Ensembles in A Single Pass Can Draw Lottery Tickets from Large Models 18 Jun 2023 · 1 repository · arXiv:2306.10460Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Mixed-Curvature Transformers for Graph Representation Learning papersreview 18 Jun 2023 · 0 repositories
-
News Verifiers Showdown: A Comparative Performance Evaluation of ChatGPT 3.5, ChatGPT 4.0, Bing AI, and Bard in News Fact-Checking 18 Jun 2023 · 0 repositories · arXiv:2306.17176
-
Summarization from Leaderboards to Practice: Choosing A Representation Backbone and Ensuring Robustness 18 Jun 2023 · 0 repositories · arXiv:2306.10555
-
Enhancing social network hate detection using back translation and GPT-3 augmentations during training and test-time 17 Jun 2023 · 1 repository
-
Snowman: A Million-scale Chinese Commonsense Knowledge Graph Distilled from Foundation Model 17 Jun 2023 · 0 repositories · arXiv:2306.10241
-
AD-AutoGPT: An Autonomous GPT for Alzheimer's Disease Infodemiology 16 Jun 2023 · 0 repositories · arXiv:2306.10095
-
Building Blocks for a Complex-Valued Transformer Architecture 16 Jun 2023 · 0 repositories · arXiv:2306.09827
-
Is Self-Repair a Silver Bullet for Code Generation? 16 Jun 2023 · 1 repository · arXiv:2306.09896
-
End-to-End Vectorized HD-map Construction with Piecewise Bezier Curve 16 Jun 2023 · 1 repository · arXiv:2306.09700
-
Evaluating Superhuman Models with Consistency Checks 16 Jun 2023 · 2 repositories · arXiv:2306.09983
-
GPT4 is Slightly Helpful for Peer-Review Assistance: A Pilot Study 16 Jun 2023 · 2 repositories · arXiv:2307.05492
-
Investigating Masking-based Data Generation in Language Models 16 Jun 2023 · 0 repositories · arXiv:2307.00008
-
MultiWave: Multiresolution Deep Architectures through Wavelet Decomposition for Multivariate Time Series Prediction 16 Jun 2023 · 1 repository · arXiv:2306.10164
-
PAtt-Lite: Lightweight Patch and Attention MobileNet for Challenging Facial Expression Recognition 16 Jun 2023 · 1 repository · arXiv:2306.09626
-
Revealing the impact of social circumstances on the selection of cancer therapy through natural language processing of social work notes 16 Jun 2023 · 0 repositories · arXiv:2306.09877
-
Robot Learning with Sensorimotor Pre-training 16 Jun 2023 · 0 repositories · arXiv:2306.10007
-
Systematic Architectural Design of Scale Transformed Attention Condenser DNNs via Multi-Scale Class Representational Response Similarity Analysis 16 Jun 2023 · 0 repositories · arXiv:2306.10128
-
TSNet-SAC: Leveraging Transformers for Efficient Task Scheduling 16 Jun 2023 · 0 repositories · arXiv:2307.07445
-
BED: Bi-Encoder-Based Detectors for Out-of-Distribution Detection 15 Jun 2023 · 1 repository · arXiv:2306.08852
-
Block-State Transformers 15 Jun 2023 · 0 repositories · arXiv:2306.09539
-
ChessGPT: Bridging Policy Learning and Language Modeling 15 Jun 2023 · 1 repository · arXiv:2306.09200Syntology official (archive's flag): 14 ran · 14 ran (of which 8 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 20 harvested samples)
-
DEYOv2: Rank Feature with Greedy Matching for End-to-End Object Detection 15 Jun 2023 · 0 repositories · arXiv:2306.09165
-
Distillation Strategies for Discriminative Speech Recognition Rescoring 15 Jun 2023 · 0 repositories · arXiv:2306.09452
-
Efficient Token-Guided Image-Text Retrieval with Consistent Multimodal Contrastive Training 15 Jun 2023 · 1 repository · arXiv:2306.08789Syntology 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples) · 2 pointer-only (licence)
-
Enlarged Large Margin Loss for Imbalanced Classification 15 Jun 2023 · 1 repository · arXiv:2306.09132
-
Evaluating Data Attribution for Text-to-Image Models 15 Jun 2023 · 2 repositories · arXiv:2306.09345Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Explaining Legal Concepts with Augmented Large Language Models (GPT-4) 15 Jun 2023 · 0 repositories · arXiv:2306.09525
-
Explore, Establish, Exploit: Red Teaming Language Models from Scratch 15 Jun 2023 · 3 repositories · arXiv:2306.09442Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring the MIT Mathematics and EECS Curriculum Using Large Language Models 15 Jun 2023 · 0 repositories · arXiv:2306.08997
-
Fast Training of Diffusion Models with Masked Transformers 15 Jun 2023 · 1 repository · arXiv:2306.09305Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
Leveraging Human Salience to Improve Calorie Estimation 15 Jun 2023 · 1 repository · arXiv:2306.09527
-
Mapping Researcher Activity based on Publication Data by means of Transformers 15 Jun 2023 · 0 repositories · arXiv:2306.09049
-
MPSA-DenseNet: A novel deep learning model for English accent classification 15 Jun 2023 · 0 repositories · arXiv:2306.08798
-
Recurrent Action Transformer with Memory 15 Jun 2023 · 1 repository · arXiv:2306.09459Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Relational Temporal Graph Reasoning for Dual-task Dialogue Language Understanding 15 Jun 2023 · 0 repositories · arXiv:2306.09114
-
Dissecting Multimodality in VideoQA Transformer Models by Impairing Modality Fusion 15 Jun 2023 · 0 repositories · arXiv:2306.08889
-
Rosetta Neurons: Mining the Common Units in a Model Zoo 15 Jun 2023 · 0 repositories · arXiv:2306.09346
-
Seeing the Pose in the Pixels: Learning Pose-Aware Representations in Vision Transformers 15 Jun 2023 · 1 repository · arXiv:2306.09331
-
SLAMB: Accelerated Large Batch Training with Sparse Communication 15 Jun 2023 · 1 repository
-
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization 15 Jun 2023 · 0 repositories · arXiv:2306.09222
-
The pop song generator: designing an online course to teach collaborative, creative AI 15 Jun 2023 · 0 repositories · arXiv:2306.10069
-
Thrilled by Your Progress! Large Language Models (GPT-4) No Longer Struggle to Pass Assessments in Higher Education Programming Courses 15 Jun 2023 · 0 repositories · arXiv:2306.10073
-
Ensembled Prediction Intervals for Causal Outcomes Under Hidden Confounding 15 Jun 2023 · 0 repositories · arXiv:2306.09520
-
DiffAug: A Diffuse-and-Denoise Augmentation for Training Robust Classifiers 15 Jun 2023 · 0 repositories · arXiv:2306.09192
-
ViP: A Differentially Private Foundation Model for Computer Vision 15 Jun 2023 · 1 repository · arXiv:2306.08842Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A semantically enhanced dual encoder for aspect sentiment triplet extraction 14 Jun 2023 · 1 repository · arXiv:2306.08373
-
Assessing the Effectiveness of GPT-3 in Detecting False Political Statements: A Case Study on the LIAR Dataset 14 Jun 2023 · 1 repository · arXiv:2306.08190
-
Building a Corpus for Biomedical Relation Extraction of Species Mentions 14 Jun 2023 · 0 repositories · arXiv:2306.08403
-
Improving Selective Visual Question Answering by Learning from Your Peers 14 Jun 2023 · 1 repository · arXiv:2306.08751
-
Language models are not naysayers: An analysis of language models on negation benchmarks 14 Jun 2023 · 1 repository · arXiv:2306.08189
-
LoSh: Long-Short Text Joint Prediction Network for Referring Video Object Segmentation 14 Jun 2023 · 1 repository · arXiv:2306.08736
-
M^2UNet: MetaFormer Multi-scale Upsampling Network for Polyp Segmentation 14 Jun 2023 · 0 repositories · arXiv:2306.08600
-
MCR-Data2vec 2.0: Improving Self-supervised Speech Pre-training via Model-level Consistency Regularization 14 Jun 2023 · 0 repositories · arXiv:2306.08463
-
MUBen: Benchmarking the Uncertainty of Molecular Representation Models 14 Jun 2023 · 2 repositories · arXiv:2306.10060Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Multimodal Optimal Transport-based Co-Attention Transformer with Global Structure Consistency for Survival Prediction 14 Jun 2023 · 3 repositories · arXiv:2306.08330Syntology official: harvested, nothing ran · 10 ran (of which 6 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Research on Named Entity Recognition in Improved transformer with R-Drop structure 14 Jun 2023 · 0 repositories · arXiv:2306.08315
-
The Expressive Leaky Memory Neuron: an Efficient and Expressive Phenomenological Neuron Model Can Solve Long-Horizon Tasks 14 Jun 2023 · 1 repository · arXiv:2306.16922Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Towards AGI in Computer Vision: Lessons Learned from GPT and Large Language Models 14 Jun 2023 · 0 repositories · arXiv:2306.08641
-
TSMixer: Lightweight MLP-Mixer Model for Multivariate Time Series Forecasting 14 Jun 2023 · 1 repository · arXiv:2306.09364Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Unraveling the ARC Puzzle: Mimicking Human Solutions with Object-Centric Decision Transformer 14 Jun 2023 · 0 repositories · arXiv:2306.08204
-
When to Use Efficient Self Attention? Profiling Text, Speech and Image Transformer Variants 14 Jun 2023 · 1 repository · arXiv:2306.08667
-
World-to-Words: Grounded Open Vocabulary Acquisition through Fast Mapping in Vision-Language Models 14 Jun 2023 · 1 repository · arXiv:2306.08685Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Area is all you need: repeatable elements make stronger adversarial attacks 13 Jun 2023 · 0 repositories · arXiv:2306.07768
-
arXiVeri: Automatic table verification with GPT 13 Jun 2023 · 1 repository · arXiv:2306.07968
-
Can ChatGPT Enable ITS? The Case of Mixed Traffic Control via Reinforcement Learning 13 Jun 2023 · 1 repository · arXiv:2306.08094
-
Detection and classification of faults aimed at preventive maintenance of PV systems 13 Jun 2023 · 0 repositories · arXiv:2306.08004
-
Enhancing Social Network Hate Detection Using Back Translation and GPT-3 Augmentations During Training and Test-Time 13 Jun 2023 · 1 repository
-
FLamE: Few-shot Learning from Natural Language Explanations 13 Jun 2023 · 0 repositories · arXiv:2306.08042
-
GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition 13 Jun 2023 · 0 repositories · arXiv:2306.07848
-
h2oGPT: Democratizing Large Language Models 13 Jun 2023 · 2 repositories · arXiv:2306.08161
-
Human-Like Intuitive Behavior and Reasoning Biases Emerged in Language Models -- and Disappeared in GPT-4 13 Jun 2023 · 0 repositories · arXiv:2306.07622
-
Improving Zero-Shot Detection of Low Prevalence Chest Pathologies using Domain Pre-trained Language Models 13 Jun 2023 · 1 repository · arXiv:2306.08000
-
MolCAP: Molecular Chemical reActivity pretraining and prompted-finetuning enhanced molecular representation learning 13 Jun 2023 · 0 repositories · arXiv:2306.09187
-
Monolingual and Cross-Lingual Knowledge Transfer for Topic Classification 13 Jun 2023 · 0 repositories · arXiv:2306.07797
-
Reviving Shift Equivariance in Vision Transformers 13 Jun 2023 · 0 repositories · arXiv:2306.07470
-
Semi-supervised learning made simple with self-supervised clustering 13 Jun 2023 · 1 repository · arXiv:2306.07483Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Discrete Graph Auto-Encoder 13 Jun 2023 · 0 repositories · arXiv:2306.07735
-
XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models 13 Jun 2023 · 1 repository · arXiv:2306.07971
-
A Survey of Vision-Language Pre-training from the Lens of Multimodal Machine Translation 12 Jun 2023 · 0 repositories · arXiv:2306.07198
-
AerialFormer: Multi-resolution Transformer for Aerial Image Segmentation 12 Jun 2023 · 1 repository · arXiv:2306.06842
-
AI-Generated Image Detection using a Cross-Attention Enhanced Dual-Stream Network 12 Jun 2023 · 1 repository · arXiv:2306.07005
-
CD-CTFM: A Lightweight CNN-Transformer Network for Remote Sensing Cloud Detection Fusing Multiscale Features 12 Jun 2023 · 0 repositories · arXiv:2306.07186
-
Enhancing COVID-19 Diagnosis through Vision Transformer-Based Analysis of Chest X-ray Images 12 Jun 2023 · 0 repositories · arXiv:2306.06914
-
Exploring Attention Mechanisms for Multimodal Emotion Recognition in an Emergency Call Center Corpus 12 Jun 2023 · 0 repositories · arXiv:2306.07115