Methods › General › Output Functions › Softmax › Papers, page 219
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 219 of 375: papers 21,801 to 21,900 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SweCTRL-Mini: a data-transparent Transformer-based large language model for controllable text generation in Swedish 27 Apr 2023 · 1 repository · arXiv:2304.13994
-
TempEE: Temporal-Spatial Parallel Transformer for Radar Echo Extrapolation Beyond Auto-Regression 27 Apr 2023 · 0 repositories · arXiv:2304.14131
-
We're Afraid Language Models Aren't Modeling Ambiguity 27 Apr 2023 · 1 repository · arXiv:2304.14399
-
Prompting GPT-3.5 for Text-to-SQL with De-semanticization and Skeleton Retrieval 26 Apr 2023 · 0 repositories · arXiv:2304.13301
-
Evaluation of GPT-3.5 and GPT-4 for supporting real-world information needs in healthcare delivery 26 Apr 2023 · 0 repositories · arXiv:2304.13714
-
Exploring the Curious Case of Code Prompts 26 Apr 2023 · 1 repository · arXiv:2304.13250Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Extracting Structured Seed-Mediated Gold Nanorod Growth Procedures from Literature with GPT-3 26 Apr 2023 · 0 repositories · arXiv:2304.13846
-
Fine Tuning with Abnormal Examples 26 Apr 2023 · 0 repositories · arXiv:2304.13783
-
HausaNLP at SemEval-2023 Task 12: Leveraging African Low Resource TweetData for Sentiment Analysis 26 Apr 2023 · 1 repository · arXiv:2304.13634
-
Technical Report: Impact of Position Bias on Language Models in Token Classification 26 Apr 2023 · 2 repositories · arXiv:2304.13567
-
The Parrot Dilemma: Human-Labeled vs. LLM-augmented Data in Classification Tasks 26 Apr 2023 · 2 repositories · arXiv:2304.13861Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
ScatterFormer: Locally-Invariant Scattering Transformer for Patient-Independent Multispectral Detection of Epileptiform Discharges 26 Apr 2023 · 1 repository · arXiv:2304.14919Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SIMARA: a database for key-value information extraction from full pages 26 Apr 2023 · 0 repositories · arXiv:2304.13606
-
TextDeformer: Geometry Manipulation using Text Guidance 26 Apr 2023 · 1 repository · arXiv:2304.13348Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
The Closeness of In-Context Learning and Weight Shifting for Softmax Regression 26 Apr 2023 · 0 repositories · arXiv:2304.13276
-
Towards Multi-Modal DBMSs for Seamless Querying of Texts and Tables 26 Apr 2023 · 0 repositories · arXiv:2304.13559
-
AI-assisted coding: Experiments with GPT-4 25 Apr 2023 · 1 repository · arXiv:2304.13187
-
Application of Transformers for Nonlinear Channel Compensation in Optical Systems 25 Apr 2023 · 0 repositories · arXiv:2304.13119
-
CompletionFormer: Depth Completion with Convolutions and Vision Transformers 25 Apr 2023 · 1 repository · arXiv:2304.13030
-
Depth-Relative Self Attention for Monocular Depth Estimation 25 Apr 2023 · 0 repositories · arXiv:2304.12849
-
Detecting Out-of-distribution Data through In-distribution Class Prior 25 Apr 2023 · 1 repository
-
DuETT: Dual Event Time Transformer for Electronic Health Records 25 Apr 2023 · 1 repository · arXiv:2304.13017Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Escaping the sentence-level paradigm in machine translation 25 Apr 2023 · 1 repository · arXiv:2304.12959
-
Introducing MBIB -- the first Media Bias Identification Benchmark Task and Dataset Collection 25 Apr 2023 · 1 repository · arXiv:2304.13148
-
LEMaRT: Label-Efficient Masked Region Transform for Image Harmonization 25 Apr 2023 · 0 repositories · arXiv:2304.13166
-
Measuring Massive Multitask Chinese Understanding 25 Apr 2023 · 2 repositories · arXiv:2304.12986Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
NLP-LTU at SemEval-2023 Task 10: The Impact of Data Augmentation and Semi-Supervised Learning Techniques on Text Classification Performance on an Imbalanced Dataset 25 Apr 2023 · 0 repositories · arXiv:2304.12847
-
SDSC-UNet: Dual Skip Connection ViT-based U-shaped Model for Building Extraction 25 Apr 2023 · 1 repository
-
Semantic Compression With Large Language Models 25 Apr 2023 · 0 repositories · arXiv:2304.12512
-
State Spaces Aren't Enough: Machine Translation Needs Attention 25 Apr 2023 · 0 repositories · arXiv:2304.12776
-
STM-UNet: An Efficient U-shaped Architecture Based on Swin Transformer and Multi-scale MLP for Medical Image Segmentation 25 Apr 2023 · 0 repositories · arXiv:2304.12615
-
SwinFSR: Stereo Image Super-Resolution using SwinIR and Frequency Domain Knowledge 25 Apr 2023 · 0 repositories · arXiv:2304.12556
-
The Potential of Visual ChatGPT For Remote Sensing 25 Apr 2023 · 0 repositories · arXiv:2304.13009
-
Theory of Posterior Concentration for Generalized Bayesian Additive Regression Trees 25 Apr 2023 · 0 repositories · arXiv:2304.12505
-
What does BERT learn about prosody? 25 Apr 2023 · 0 repositories · arXiv:2304.12706
-
AGI: Artificial General Intelligence for Education 24 Apr 2023 · 0 repositories · arXiv:2304.12479
-
Augmentation-based Domain Generalization for Semantic Segmentation 24 Apr 2023 · 0 repositories · arXiv:2304.12122
-
Better Question-Answering Models on a Budget 24 Apr 2023 · 1 repository · arXiv:2304.12370
-
Directed Acyclic Transformer Pre-training for High-quality Non-autoregressive Text Generation 24 Apr 2023 · 1 repository · arXiv:2304.11791Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Explicit Correspondence Matching for Generalizable Neural Radiance Fields 24 Apr 2023 · 1 repository · arXiv:2304.12294Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Extreme Classification for Answer Type Prediction in Question Answering 24 Apr 2023 · 0 repositories · arXiv:2304.12395
-
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM 24 Apr 2023 · 1 repository · arXiv:2304.11872
-
IRNeXt: Rethinking Convolutional Network Design for Image Restoration 24 Apr 2023 · 1 repository
-
Master: Meta Style Transformer for Controllable Zero-Shot and Few-Shot Artistic Style Transfer 24 Apr 2023 · 0 repositories · arXiv:2304.11818
-
MixPro: Data Augmentation with MaskMix and Progressive Attention Labeling for Vision Transformer 24 Apr 2023 · 1 repository · arXiv:2304.12043Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 7 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 8 pointer-only (licence)
-
Multi-cropping Contrastive Learning and Domain Consistency for Unsupervised Image-to-Image Translation 24 Apr 2023 · 0 repositories · arXiv:2304.12235
-
NoiseTrans: Point Cloud Denoising with Transformers 24 Apr 2023 · 0 repositories · arXiv:2304.11812
-
Now You See Me: Robust approach to Partial Occlusions 24 Apr 2023 · 0 repositories · arXiv:2304.11779
-
Once Detected, Never Lost: Surpassing Human Performance in Offline LiDAR based 3D Object Detection 24 Apr 2023 · 2 repositories · arXiv:2304.12315Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
PARAGRAPH2GRAPH: A GNN-based framework for layout paragraph analysis 24 Apr 2023 · 1 repository · arXiv:2304.11810
-
Pre-trained Embeddings for Entity Resolution: An Experimental Analysis [Experiment, Analysis & Benchmark] 24 Apr 2023 · 1 repository · arXiv:2304.12329
-
Rank Flow Embedding for Unsupervised and Semi-Supervised Manifold Learning 24 Apr 2023 · 1 repository · arXiv:2304.12448
-
Vision-based Estimation of Fatigue and Engagement in Cognitive Training Sessions 24 Apr 2023 · 1 repository · arXiv:2304.12470
-
Self-regularised Minimum Latency Training for Streaming Transformer-based Speech Recognition 24 Apr 2023 · 0 repositories · arXiv:2304.11985
-
SocialDial: A Benchmark for Socially-Aware Dialogue Systems 24 Apr 2023 · 1 repository · arXiv:2304.12026
-
Text-to-Audio Generation using Instruction-Tuned LLM and Latent Diffusion Model 24 Apr 2023 · 1 repository · arXiv:2304.13731
-
Transformer-based stereo-aware 3D object detection from binocular images 24 Apr 2023 · 0 repositories · arXiv:2304.11906
-
Universal Domain Adaptation via Compressive Attention Matching 24 Apr 2023 · 0 repositories · arXiv:2304.11862
-
WizardLM: Empowering Large Language Models to Follow Complex Instructions 24 Apr 2023 · 4 repositories · arXiv:2304.12244Syntology official (archive's flag): 1 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Processing Natural Language on Embedded Devices: How Well Do Transformer Models Perform? 23 Apr 2023 · 2 repositories · arXiv:2304.11520
-
Vision Transformer for Efficient Chest X-ray and Gastrointestinal Image Classification 23 Apr 2023 · 0 repositories · arXiv:2304.11529
-
Boosting Theory-of-Mind Performance in Large Language Models via Prompting 22 Apr 2023 · 1 repository · arXiv:2304.11490
-
Dilated-UNet: A Fast and Accurate Medical Image Segmentation Approach using a Dilated Transformer and U-Net Architecture 22 Apr 2023 · 1 repository · arXiv:2304.11450
-
Incomplete Multimodal Learning for Remote Sensing Data Fusion 22 Apr 2023 · 0 repositories · arXiv:2304.11381
-
L3Cube-IndicSBERT: A simple approach for learning cross-lingual sentence representations using multilingual BERT 22 Apr 2023 · 0 repositories · arXiv:2304.11434
-
Vision Transformers, a new approach for high-resolution and large-scale mapping of canopy heights 22 Apr 2023 · 0 repositories · arXiv:2304.11487
-
A Group-Specific Approach to NLP for Hate Speech Detection 21 Apr 2023 · 1 repository · arXiv:2304.11223
-
A Preliminary Study of Deep Learning Sensor Fusion for Pedestrian Detection 21 Apr 2023 · 1 repository
-
BERT Based Clinical Knowledge Extraction for Biomedical Knowledge Graph Construction and Analysis 21 Apr 2023 · 0 repositories · arXiv:2304.10996
-
Building Multimodal AI Chatbots 21 Apr 2023 · 1 repository · arXiv:2305.03512
-
Can GPT-4 Perform Neural Architecture Search? 21 Apr 2023 · 1 repository · arXiv:2304.10970
-
DeformableFormer: Classification of Endoscopic Ultrasound Guided Fine Needle Biopsy in Pancreatic Diseases 21 Apr 2023 · 0 repositories · arXiv:2304.10791
-
Evaluating Transformer Language Models on Arithmetic Operations Using Number Decomposition 21 Apr 2023 · 1 repository · arXiv:2304.10977
-
Exploiting Patch Sizes and Resolutions for Multi-Scale Deep Learning in Mammogram Image Classification 21 Apr 2023 · 0 repositories
-
Inducing anxiety in large language models can induce bias 21 Apr 2023 · 0 repositories · arXiv:2304.11111
-
Multi-Modal Deep Learning for Credit Rating Prediction Using Text and Numerical Data Streams 21 Apr 2023 · 1 repository · arXiv:2304.10740
-
Self-Attention in Colors: Another Take on Encoding Graph Structure in Transformers 21 Apr 2023 · 1 repository · arXiv:2304.10933
-
Text2Time: Transformer-based Article Time Period Prediction 21 Apr 2023 · 0 repositories · arXiv:2304.10859
-
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination 21 Apr 2023 · 0 repositories · arXiv:2304.14347
-
Transformer-based models and hardware acceleration analysis in autonomous driving: A survey 21 Apr 2023 · 0 repositories · arXiv:2304.10891
-
Who's the Best Detective? LLMs vs. MLs in Detecting Incoherent Fourth Grade Math Answers 21 Apr 2023 · 0 repositories · arXiv:2304.11257
-
An Introduction to Transformers 20 Apr 2023 · 0 repositories · arXiv:2304.10557
-
Analyzing FOMC Minutes: Accuracy and Constraints of Language Models 20 Apr 2023 · 0 repositories · arXiv:2304.10164
-
Angle based dynamic learning rate for gradient descent 20 Apr 2023 · 1 repository · arXiv:2304.10457
-
Attention Scheme Inspired Softmax Regression 20 Apr 2023 · 0 repositories · arXiv:2304.10411
-
Contrastive Tuning: A Little Help to Make Masked Autoencoders Forget 20 Apr 2023 · 1 repository · arXiv:2304.10520
-
Domain-specific Continued Pretraining of Language Models for Capturing Long Context in Mental Health 20 Apr 2023 · 0 repositories · arXiv:2304.10447
-
Enhancing object detection robustness: A synthetic and natural perturbation approach 20 Apr 2023 · 0 repositories · arXiv:2304.10622
-
Ensembling Instance and Semantic Segmentation for Panoptic Segmentation 20 Apr 2023 · 0 repositories · arXiv:2304.10326
-
HM-ViT: Hetero-modal Vehicle-to-Vehicle Cooperative perception with vision transformer 20 Apr 2023 · 1 repository · arXiv:2304.10628Syntology 17 ran (of which 7 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 8 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Is Cross-modal Information Retrieval Possible without Training? 20 Apr 2023 · 0 repositories · arXiv:2304.11095
-
Learning CLIP Guided Visual-Text Fusion Transformer for Video-based Pedestrian Attribute Recognition 20 Apr 2023 · 1 repository · arXiv:2304.10091
-
Meta Semantics: Towards better natural language understanding and reasoning 20 Apr 2023 · 0 repositories · arXiv:2304.10663
-
MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models 20 Apr 2023 · 6 repositories · arXiv:2304.10592
-
Movie Box Office Prediction With Self-Supervised and Visually Grounded Pretraining 20 Apr 2023 · 0 repositories · arXiv:2304.10311
-
OptoGPT: A Foundation Model for Inverse Design in Optical Multilayer Thin Film Structures 20 Apr 2023 · 0 repositories · arXiv:2304.10294
-
Performance of ChatGPT on the US Fundamentals of Engineering Exam: Comprehensive Assessment of Proficiency and Potential Implications for Professional Environmental Engineering Practice 20 Apr 2023 · 0 repositories · arXiv:2304.12198
-
Dynamic Graph Representation Learning via Edge Temporal States Modeling and Structure-reinforced Transformer 20 Apr 2023 · 0 repositories · arXiv:2304.10079
-
Safety Assessment of Chinese Large Language Models 20 Apr 2023 · 2 repositories · arXiv:2304.10436
-
SINC: Spatial Composition of 3D Human Motions for Simultaneous Action Generation 20 Apr 2023 · 0 repositories · arXiv:2304.10417