Methods › General › Output Functions › Softmax › Papers, page 207
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 207 of 375: papers 20,601 to 20,700 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Constraint-aware and Ranking-distilled Token Pruning for Efficient Transformer Inference 26 Jun 2023 · 1 repository · arXiv:2306.14393
-
CST-YOLO: A Novel Method for Blood Cell Detection Based on Improved YOLOv7 and CNN-Swin Transformer 26 Jun 2023 · 1 repository · arXiv:2306.14590
-
DNABERT-2: Efficient Foundation Model and Benchmark For Multi-Species Genome 26 Jun 2023 · 6 repositories · arXiv:2306.15006Syntology community repositories only · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 10 unverified (of 23 harvested samples) · 1 pointer-only (licence)
-
Exploring the Robustness of Large Language Models for Solving Programming Problems 26 Jun 2023 · 0 repositories · arXiv:2306.14583
-
FeSViBS: Federated Split Learning of Vision Transformer with Block Sampling 26 Jun 2023 · 1 repository · arXiv:2306.14638
-
Large Multimodal Models: Notes on CVPR 2023 Tutorial 26 Jun 2023 · 0 repositories · arXiv:2306.14895
-
LM4HPC: Towards Effective Language Model Application in High-Performance Computing 26 Jun 2023 · 0 repositories · arXiv:2306.14979
-
LongCoder: A Long-Range Pre-trained Language Model for Code Completion 26 Jun 2023 · 1 repository · arXiv:2306.14893Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ParameterNet: Parameters Are All You Need 26 Jun 2023 · 0 repositories · arXiv:2306.14525
-
Supervised Pretraining Can Learn In-Context Reinforcement Learning 26 Jun 2023 · 0 repositories · arXiv:2306.14892
-
ViNT: A Foundation Model for Visual Navigation 26 Jun 2023 · 1 repository · arXiv:2306.14846
-
A Web-based Mpox Skin Lesion Detection System Using State-of-the-art Deep Learning Models Considering Racial Diversity 25 Jun 2023 · 1 repository · arXiv:2306.14169
-
Adaptive Window Pruning for Efficient Local Motion Deblurring 25 Jun 2023 · 0 repositories · arXiv:2306.14268
-
Addressing Cold Start Problem for End-to-end Automatic Speech Scoring 25 Jun 2023 · 0 repositories · arXiv:2306.14310
-
G-STO: Sequential Main Shopping Intention Detection via Graph-Regularized Stochastic Transformer 25 Jun 2023 · 0 repositories · arXiv:2306.14314
-
Interactive Design by Integrating a Large Pre-Trained Language Model and Building Information Modeling 25 Jun 2023 · 0 repositories · arXiv:2306.14165
-
Let's Do a Thought Experiment: Using Counterfactuals to Improve Moral Reasoning 25 Jun 2023 · 0 repositories · arXiv:2306.14308
-
Multi-Scale Cross Contrastive Learning for Semi-Supervised Medical Image Segmentation 25 Jun 2023 · 0 repositories · arXiv:2306.14293
-
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices 25 Jun 2023 · 0 repositories · arXiv:2306.14263
-
Steganographic Capacity of Deep Learning Models 25 Jun 2023 · 0 repositories · arXiv:2306.17189
-
Switch-BERT: Learning to Model Multimodal Interactions by Switching Attention and Input 25 Jun 2023 · 0 repositories · arXiv:2306.14182
-
Action Q-Transformer: Visual Explanation in Deep Reinforcement Learning with Encoder-Decoder Model using Action Query 24 Jun 2023 · 0 repositories · arXiv:2306.13879
-
Can GPT-4 Support Analysis of Textual Data in Tasks Requiring Highly Specialized Domain Expertise? 24 Jun 2023 · 0 repositories · arXiv:2306.13906
-
Comparison of Pre-trained Language Models for Turkish Address Parsing 24 Jun 2023 · 0 repositories · arXiv:2306.13947
-
Emotion Flip Reasoning in Multiparty Conversations 24 Jun 2023 · 0 repositories · arXiv:2306.13959
-
Fusing Multimodal Signals on Hyper-complex Space for Extreme Abstractive Text Summarization (TL;DR) of Scientific Contents 24 Jun 2023 · 1 repository · arXiv:2306.13968
-
IERL: Interpretable Ensemble Representation Learning -- Combining CrowdSourced Knowledge and Distributed Semantic Representations 24 Jun 2023 · 0 repositories · arXiv:2306.13865
-
Is Pre-training Truly Better Than Meta-Learning? 24 Jun 2023 · 0 repositories · arXiv:2306.13841
-
L3Cube-MahaSent-MD: A Multi-domain Marathi Sentiment Analysis Dataset and Transformer Models 24 Jun 2023 · 1 repository · arXiv:2306.13888
-
Large Language Models as Sous Chefs: Revising Recipes with GPT-3 24 Jun 2023 · 1 repository · arXiv:2306.13986
-
Large Sequence Models for Sequential Decision-Making: A Survey 24 Jun 2023 · 0 repositories · arXiv:2306.13945
-
Math Word Problem Solving by Generating Linguistic Variants of Problem Statements 24 Jun 2023 · 1 repository · arXiv:2306.13899
-
My Boli: Code-mixed Marathi-English Corpora, Pretrained Language Models and Evaluation Benchmarks 24 Jun 2023 · 1 repository · arXiv:2306.14030
-
On the Uses of Large Language Models to Interpret Ambiguous Cyberattack Descriptions 24 Jun 2023 · 0 repositories · arXiv:2306.14062
-
Partitioning-Guided K-Means: Extreme Empty Cluster Resolution for Extreme Model Compression 24 Jun 2023 · 0 repositories · arXiv:2306.14031
-
Waypoint Transformer: Reinforcement Learning via Supervised Learning with Intermediate Targets 24 Jun 2023 · 0 repositories · arXiv:2306.14069
-
Abstractive Text Summarization for Resumes With Cutting Edge NLP Transformers and LSTM 23 Jun 2023 · 0 repositories · arXiv:2306.13315
-
Bridging the Performance Gap between DETR and R-CNN for Graphical Object Detection in Document Images 23 Jun 2023 · 0 repositories · arXiv:2306.13526
-
Cross-Language Speech Emotion Recognition Using Multimodal Dual Attention Transformers 23 Jun 2023 · 0 repositories · arXiv:2306.13804
-
Efficient Online Processing with Deep Neural Networks 23 Jun 2023 · 1 repository · arXiv:2306.13474
-
On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes 23 Jun 2023 · 0 repositories · arXiv:2306.13649
-
Incorporating Graph Information in Transformer-based AMR Parsing 23 Jun 2023 · 1 repository · arXiv:2306.13467
-
LLM-Assisted Content Analysis: Using Large Language Models to Support Deductive Coding 23 Jun 2023 · 0 repositories · arXiv:2306.14924
-
Retrieval-Pretrained Transformer: Long-range Language Modeling with Self-retrieval 23 Jun 2023 · 1 repository · arXiv:2306.13421Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
ProRes: Exploring Degradation-aware Visual Prompt for Universal Image Restoration 23 Jun 2023 · 1 repository · arXiv:2306.13653Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Resume Information Extraction via Post-OCR Text Processing 23 Jun 2023 · 0 repositories · arXiv:2306.13775
-
Swin-Free: Achieving Better Cross-Window Attention and Efficiency with Size-varying Window 23 Jun 2023 · 0 repositories · arXiv:2306.13776
-
System-Level Natural Language Feedback 23 Jun 2023 · 1 repository · arXiv:2306.13588Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Double Helix inside the NLP Transformer 23 Jun 2023 · 0 repositories · arXiv:2306.13817
-
Upscaling Global Hourly GPP with Temporal Fusion Transformer (TFT) 23 Jun 2023 · 0 repositories · arXiv:2306.13815
-
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale 23 Jun 2023 · 1 repository · arXiv:2306.15687Syntology 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 3 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
A Comparison of Time-based Models for Multimodal Emotion Recognition 22 Jun 2023 · 0 repositories · arXiv:2306.13076
-
Can a single image processing algorithm work equally well across all phases of DCE-MRI? 22 Jun 2023 · 0 repositories · arXiv:2306.12988
-
Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs 22 Jun 2023 · 1 repository · arXiv:2306.13063Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Cross-lingual Cross-temporal Summarization: Dataset, Models, Evaluation 22 Jun 2023 · 1 repository · arXiv:2306.12916
-
Deep Metric Learning with Soft Orthogonal Proxies 22 Jun 2023 · 0 repositories · arXiv:2306.13055
-
Learning from Visual Observation via Offline Pretrained State-to-Go Transformer 22 Jun 2023 · 0 repositories · arXiv:2306.12860
-
Minimalist and High-Quality Panoramic Imaging with PSF-aware Transformers 22 Jun 2023 · 1 repository · arXiv:2306.12992
-
Named entity recognition in resumes 22 Jun 2023 · 0 repositories · arXiv:2306.13062
-
Prompt to GPT-3: Step-by-Step Thinking Instructions for Humor Generation 22 Jun 2023 · 1 repository · arXiv:2306.13195
-
Quantizable Transformers: Removing Outliers by Helping Attention Heads Do Nothing 22 Jun 2023 · 1 repository · arXiv:2306.12929
-
Rethinking the Backward Propagation for Adversarial Transferability 22 Jun 2023 · 2 repositories · arXiv:2306.12685Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 9 unverified (of 15 harvested samples)
-
Visual Adversarial Examples Jailbreak Aligned Large Language Models 22 Jun 2023 · 1 repository · arXiv:2306.13213
-
ARIES: A Corpus of Scientific Paper Edits Made in Response to Peer Reviews 21 Jun 2023 · 1 repository · arXiv:2306.12587Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Balanced Mixture of SuperNets for Learning the CNN Pooling Architecture 21 Jun 2023 · 1 repository · arXiv:2306.11982
-
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair 21 Jun 2023 · 0 repositories · arXiv:2307.00012
-
Joint Prompt Optimization of Stacked LLMs using Variational Inference 21 Jun 2023 · 1 repository · arXiv:2306.12509Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 16 unverified (of 19 harvested samples)
-
Edge Devices Inference Performance Comparison 21 Jun 2023 · 0 repositories · arXiv:2306.12093
-
Fast Segment Anything 21 Jun 2023 · 1 repository · arXiv:2306.12156
-
GPT-Based Models Meet Simulation: How to Efficiently Use Large-Scale Pre-Trained Language Models Across Simulation Tasks 21 Jun 2023 · 0 repositories · arXiv:2306.13679
-
HSR-Diff:Hyperspectral Image Super-Resolution via Conditional Diffusion Models 21 Jun 2023 · 0 repositories · arXiv:2306.12085
-
Inter-Instance Similarity Modeling for Contrastive Learning 21 Jun 2023 · 1 repository · arXiv:2306.12243
-
Investigating Pre-trained Language Models on Cross-Domain Datasets, a Step Closer to General AI 21 Jun 2023 · 0 repositories · arXiv:2306.12205
-
MSW-Transformer: Multi-Scale Shifted Windows Transformer Networks for 12-Lead ECG Classification 21 Jun 2023 · 0 repositories · arXiv:2306.12098
-
Neural Multigrid Memory For Computational Fluid Dynamics 21 Jun 2023 · 1 repository · arXiv:2306.12545
-
Probing the limit of hydrologic predictability with the Transformer network 21 Jun 2023 · 0 repositories · arXiv:2306.12384
-
Social Media Emotions and IPO Returns 21 Jun 2023 · 0 repositories · arXiv:2306.12602
-
Solving and Generating NPR Sunday Puzzles with Large Language Models 21 Jun 2023 · 1 repository · arXiv:2306.12255
-
StarVQA+: Co-training Space-Time Attention for Video Quality Assessment 21 Jun 2023 · 1 repository · arXiv:2306.12298
-
Towards Accurate Translation via Semantically Appropriate Application of Lexical Constraints 21 Jun 2023 · 1 repository · arXiv:2306.12089
-
What Constitutes Good Contrastive Learning in Time-Series Forecasting? 21 Jun 2023 · 1 repository · arXiv:2306.12086Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Which Spurious Correlations Impact Reasoning in NLI Models? A Visual Interactive Diagnosis through Data-Constrained Counterfactuals 21 Jun 2023 · 0 repositories · arXiv:2306.12146
-
A Novel Counterfactual Data Augmentation Method for Aspect-Based Sentiment Analysis 20 Jun 2023 · 0 repositories · arXiv:2306.11260
-
Comparing Deep Learning Models for the Task of Volatility Prediction Using Multivariate Data 20 Jun 2023 · 0 repositories · arXiv:2306.12446
-
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models 20 Jun 2023 · 0 repositories · arXiv:2306.11698
-
Deep Fusion: Efficient Network Training via Pre-trained Initializations 20 Jun 2023 · 0 repositories · arXiv:2306.11903
-
Democratizing LLMs for Low-Resource Languages by Leveraging their English Dominant Abilities with Linguistically-Diverse Prompts 20 Jun 2023 · 0 repositories · arXiv:2306.11372
-
Event Stream GPT: A Data Pre-processing and Modeling Library for Generative, Pre-trained Transformers over Continuous-time Sequences of Complex Events 20 Jun 2023 · 1 repository · arXiv:2306.11547Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
A GPT-4 Reticular Chemist for Guiding MOF Discovery 20 Jun 2023 · 1 repository · arXiv:2306.14915
-
Harnessing the Power of Adversarial Prompting and Large Language Models for Robust Hypothesis Generation in Astronomy 20 Jun 2023 · 0 repositories · arXiv:2306.11648
-
InRank: Incremental Low-Rank Learning 20 Jun 2023 · 1 repository · arXiv:2306.11250
-
Learning to Generate Better Than Your LLM 20 Jun 2023 · 1 repository · arXiv:2306.11816
-
Multiverse Transformer: 1st Place Solution for Waymo Open Sim Agents Challenge 2023 20 Jun 2023 · 0 repositories · arXiv:2306.11868
-
Surfer: Progressive Reasoning with World Models for Robotic Manipulation 20 Jun 2023 · 0 repositories · arXiv:2306.11335
-
Textbooks Are All You Need 20 Jun 2023 · 0 repositories · arXiv:2306.11644
-
Transforming Graphs for Enhanced Attribute Clustering: An Innovative Graph Transformer-Based Method 20 Jun 2023 · 0 repositories · arXiv:2306.11307
-
Unfolding Framework with Prior of Convolution-Transformer Mixture and Uncertainty Estimation for Video Snapshot Compressive Imaging 20 Jun 2023 · 0 repositories · arXiv:2306.11316
-
A Preliminary Study of ChatGPT on News Recommendation: Personalization, Provider Fairness, Fake News 19 Jun 2023 · 1 repository · arXiv:2306.10702
-
AMRs Assemble! Learning to Ensemble with Autoregressive Models for AMR Parsing 19 Jun 2023 · 1 repository · arXiv:2306.10786
-
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language Models 19 Jun 2023 · 1 repository · arXiv:2306.10968